Skip to content

#Deployment

3 today

Sep 17

Thursday

Sep 16

Wednesday
  1. From Megawatts to Tokens: How NVIDIA Maximizes AI Factory Production

    NVIDIA introduced DSX at AI Infra Summit to improve AI-factory token throughput within a fixed power budget. Lambda validated DSX MaxLPS on a five-rack, 19-node HGX B200 cluster, running 19 nodes within the same budget as 16 at full power. Cluster token throughput rose 24%, from around four million to five million tokens per second, with performance per watt up 23%.

Sep 15

Tuesday
  1. Heart of the Matter: How a Major Children’s Hospital Uses Open Source NVIDIA AI for Cardiac Care

    Children's Hospital of Philadelphia (CHOP) uses MONAI, the open-source medical imaging framework co-founded by NVIDIA, to reduce paediatric heart modelling from four hours to seconds, already supporting complex ventricular septal defect surgical planning. CHOP is working with NVIDIA to integrate Newton, an open-source physics engine based on NVIDIA Warp, into SlicerHeart, reducing cardiac-device simulation from up to four hours to near real time.

Sep 12

Saturday

Sep 11

Friday

Sep 10

Thursday
  1. Cloudera and Mistral Partner to Bring Specialized, Sovereign Intelligence to Enterprise Data

    Mistral and Cloudera partnered to integrate Mistral models into Cloudera's hybrid data platform. Enterprises can run inference in private or public clouds, on-premises and fully air-gapped environments with full control. They can train custom models on proprietary data in controlled environments while retaining ownership of data and resulting intelligence. Cloudera runs 30 exabytes of customer-managed data.

Sep 5

Saturday

Sep 3

Thursday

Sep 1

Tuesday

Aug 25

Tuesday
  1. Mistral x HUMAIN

    Mistral and HUMAIN announced a strategic partnership covering infrastructure, advanced models and deployment, initially focusing on cybersecurity and voice, with frontier models strong in Arabic planned. Worth hundreds of millions of euros, it includes exploring HUMAIN data centres and joint market strategies for regulated Saudi industries.

Aug 21

Friday

Aug 18

Tuesday

Aug 14

Friday

Aug 11

Tuesday

Aug 6

Thursday

Jul 30

Thursday

Jul 27

Monday

Jul 16

Thursday

Jul 13

Monday

Jul 10

Friday

Jul 9

Thursday
  1. Your Prompts and Skills need a system of record.

    Mistral Studio now provides central management of Prompts and Skills, with version histories, clear ownership and traceability. Immutable versions, rollback, tags and audit logs make AI behaviour governable and discoverable, while Observability traces production outputs to their versions. Skills can be invoked directly from Studio as MCP servers, ensuring production executes the same governed asset.

Jul 8

Wednesday

Jul 7

Tuesday

Jul 6

Monday

Jun 27

Saturday

Jun 25

Thursday
  1. Optimizing cloud economics with linear elastic caching

    Google Research proposed linear elastic caching, modelling page eviction as a ski-rental problem and using shallow decision trees to predict page TTLs, dynamically resizing caches to minimise total ownership cost. After months in Spanner production, memory fell 15.5%, cache misses rose only 5.5%, TCO dropped around 5% and actual I/O cost impact was just 0.5%. It was also validated on several public cache traces.