NVIDIA AI 工厂如何最大化投资回报:高产出、长寿命、通用性
NVIDIA 提出 AI 工厂投资回报取决于三个因素:每兆瓦产出能力、硬件使用寿命和 token 需求,并以 Vera Rubin NVL72 为例称其每兆瓦吞吐量比 GB300 NVL72 高 30 倍以上,在 DeepSeek V4 Pro 上每百万 token 成本最多低 45 倍。
NVIDIA 提出 AI 工厂投资回报取决于三个因素:每兆瓦产出能力、硬件使用寿命和 token 需求,并以 Vera Rubin NVL72 为例称其每兆瓦吞吐量比 GB300 NVL72 高 30 倍以上,在 DeepSeek V4 Pro 上每百万 token 成本最多低 45 倍。
OpenAI and Synopsys have signed a multi-year strategic partnership to build a specialized AI model for chip design called GPT-Synopsys. Synopsys makes electronic design automation (EDA) tools, the software engineers use to design computer chips and semiconductors. GPT-Synopsys combines OpenAI's AI technology with Synopsys' EDA tools. OpenAI is licensing Synopsys' EDA tools for the project.
Destro emerged from stealth with $8 million in seed funding. Its AI intelligence layer directs both robots and human workers to carry out cross-docking in Yusen Logistics operations. Built on an open-weight vision-language-action model, it uses the Vision operating system to drive robots and the Mothership operating system to coordinate people, goods and vehicles. Its pilot has expanded from 3 robots to 26, with a new 17-robot pilot launched in Southern California.
CoreWeave announced that NVIDIA Vera Rubin NVL72 systems with Spectrum-X 102.4T Ethernet were available on CoreWeave Cloud, with Cognition becoming the first customer to run them in production.
DeepSeek and Huawei have collaborated on programming tools for Ascend chips and released them all as open source. At their core is TileLang, a language that is easier to learn than CUDA. It was originally developed by researchers at Peking University, and DeepSeek has used it for about a year. The companies have also optimised a supernode comprising 128 Ascend 950 chips, with Huawei saying it provided full support. DeepSeek argues that an independent AI chip software ecosystem first needs a general-purpose language that is easy to program and can fully exploit the hardware's performance.
Berlin-based start-up Restate has raised $20 million in a Series A round led by Singular, with participation from Redpoint Ventures and Capital One Ventures. Its durable execution engine helps AI agents' multi-step workflows recover from crashes and network interruptions.
Cerebras Systems CEO and co-founder Andrew Feldman will present “Can AI Keep Scaling?” on the Disrupt Stage at TechCrunch Disrupt 2026.
According to sources, AI inference infrastructure provider Modal Labs is close to completing a $750 million round led by Accel at a $15.75 billion valuation. The size of the round had not previously been reported.
Mistral opened a hub in Munich with a research team focused on Physics AI and industrial AI, and plans to build one gigawatt of European computing capacity by 2030. It will work with BMW on crash simulation and engineering AI, Siemens Energy on industrial AI applications, and the Technical University of Munich (TUM) on automotive aerodynamics digital twins using wind-tunnel facilities.
Google's Project Suncatcher will launch its first test satellite, MVP, on 1 October to validate an orbital AI data-centre concept. Roughly refrigerator-sized, it carries four Google-designed TPU accelerators and about one kilowatt of solar power. Google integrated the chips into a satellite already built by imaging company Planet Labs to accelerate progress; the original plan was to launch two custom satellites in 2027.
NVIDIA launched DSX Ready certification to validate partner products against its DSX AI factory reference design, initially covering battery energy storage systems (BESS) and coolant distribution units (CDU).
NVIDIA explains that deploying physical AI at scale requires safety covering hardware, software, AI behaviour and operating environments, and introduces its full-stack safety system NVIDIA Halos.
NVIDIA held an AI ecosystem event at the Grand Egyptian Museum, with Egyptian Deep Learning Institute participation growing over tenfold in a year. Hassan Allam received a data centre licence from Egypt's National Telecom Regulatory Authority and is developing a new centre with A15 with expected investment of $400 million. NVIDIA says four African AI factories have been announced or launched, with another 656 MW under construction.
NVIDIA highlighted five AI clean-energy companies at New York Climate Week. ThinkLabs AI's NVIDIA CUDA-based digital twins and agents help Southern California Edison cut grid connection application assessments from 30–45 days to two minutes.
Anthropic introduced Projects in Claude Code, allowing one conversation to spawn parallel cloud sessions, pass context between threads and keep running after users leave. Google updated the Antigravity-based harness for Gemini managed agents with Credentials API and Files API, claiming up to 30% lower costs and 22% higher cache-hit rates.
Emerald AI, Google and NVIDIA announced the AI Energy Management Alliance (AEMA) to encourage data centres to dynamically adjust electricity use to grid conditions.
A University of Manchester team used NVIDIA Earth-2's generative Earth-2 CorrDiff model and one year of hourly UK pollution data to train a nationwide air-pollution model at 2–3-square-kilometre resolution, completing training in two days on one eight-GPU Isambard-AI node.
NVIDIA introduced DSX at AI Infra Summit to improve AI-factory token throughput within a fixed power budget. Lambda validated DSX MaxLPS on a five-rack, 19-node HGX B200 cluster, running 19 nodes within the same budget as 16 at full power. Cluster token throughput rose 24%, from around four million to five million tokens per second, with performance per watt up 23%.
NVIDIA announced several advances for Vera Rubin and the DSX platform at AI Infra Summit, focusing on energy-efficiency improvements in token throughput per megawatt.
Mistral and Cloudera partnered to integrate Mistral models into Cloudera's hybrid data platform. Enterprises can run inference in private or public clouds, on-premises and fully air-gapped environments with full control. They can train custom models on proprietary data in controlled environments while retaining ownership of data and resulting intelligence. Cloudera runs 30 exabytes of customer-managed data.
Mistral and HUMAIN announced a strategic partnership covering infrastructure, advanced models and deployment, initially focusing on cybersecurity and voice, with frontier models strong in Arabic planned. Worth hundreds of millions of euros, it includes exploring HUMAIN data centres and joint market strategies for regulated Saudi industries.
With Google's support, UC San Diego plans a data centre using motherboards from 2,000 retired Pixel phones to provide low-cost, low-carbon cloud computing for hundreds of students and faculty. SPEC benchmarks show 25–50 phones roughly equal one modern server. Phones form Kubernetes-managed clusters of 25–50 devices, with launch expected in autumn 2026.
Mistral AI signed a definitive agreement this week to acquire Physics AI pioneer Emmi AI, strengthening its AI transformation services for industrial companies. Founded in Austria, Emmi AI has over 30 researchers and engineers focusing on large engineering models that replace days of computation with real-time simulation and build digital twins. Its co-founders and team will join Mistral's Science and Applied AI teams in May.
Mistral AI announced a multi-year SAP partnership to integrate its models into AI Foundation and jointly develop tailored solutions for complex European industries and the public sector. It will also accelerate vision-language-action model research with Helsing for defence and security, open a German office in the coming months and significantly expand its local team.
Mistral AI, Carbone 4 and France's ADEME completed the first full-lifecycle AI model analysis. By January 2025, Large 2 training and 18 months of use produced 20.4 ktCO₂e, consumed 281,000 cubic metres of water and 660 kg Sb eq of resources.
Mistral AI launched AI for Citizens, helping governments and public institutions strategically apply AI to transform public services, innovate and safeguard competitiveness. It offers open-source models, self-hosting, data sovereignty and tailored R&D, with government and public-sector partners in France, Luxembourg, Singapore, the Netherlands, the UK and Switzerland.