NVIDIA AI 工厂如何最大化投资回报:高产出、长寿命、通用性
NVIDIA 提出 AI 工厂投资回报取决于三个因素:每兆瓦产出能力、硬件使用寿命和 token 需求,并以 Vera Rubin NVL72 为例称其每兆瓦吞吐量比 GB300 NVL72 高 30 倍以上,在 DeepSeek V4 Pro 上每百万 token 成本最多低 45 倍。
NVIDIA 提出 AI 工厂投资回报取决于三个因素:每兆瓦产出能力、硬件使用寿命和 token 需求,并以 Vera Rubin NVL72 为例称其每兆瓦吞吐量比 GB300 NVL72 高 30 倍以上,在 DeepSeek V4 Pro 上每百万 token 成本最多低 45 倍。
NVIDIA's 26th graduate fellowship programme is accepting applications worldwide, focusing on AI, machine learning, autonomous driving, computer graphics, robotics, healthcare and high-performance computing, with awards of up to $60,000 per student.
MSRA – Singapore, Microsoft's first research lab in Southeast Asia, spent its first year focusing on next-generation AI models and agent systems, domain AI, AI-native research practices and talent ecosystems. It works with Singapore's healthcare ecosystem on multimodal medical AI and self-evolving diagnostic agents, and jointly held a logistics and transport AI executive roundtable with EDB in February 2026.
OpenAI apologised for incidents involving Australian government websites and announced stricter safeguards and support measures to help strengthen Australia's cyber defences.
Mistral opened a hub in Munich with a research team focused on Physics AI and industrial AI, and plans to build one gigawatt of European computing capacity by 2030. It will work with BMW on crash simulation and engineering AI, Siemens Energy on industrial AI applications, and the Technical University of Munich (TUM) on automotive aerodynamics digital twins using wind-tunnel facilities.
OpenAI expanded its AI partnership and fellowship programme with the Lenfest Institute, committing $5 million in funding and up to $5 million in software credits and engineering support.
NVIDIA joined a global research consortium including Google DeepMind and EMBL-EBI to release predicted 3D protein complex structures for over 2,800 viruses through the AlphaFold Database, free for any scientist to use.
OpenAI Academy marks its second anniversary, aiming to bring AI skills to more communities.
NVIDIA validation engineer Sakeena Fiza validates new hardware before mass production in its data centre systems lab. One memorable moment was the first successful system-level enumeration of the Rubin GPU. She compares validation to solving crimes, finding and reproducing issues before customers do, across trays, racks, clusters and customers' AI factories.
OpenAI expanded access to its Daybreak programme to the Ukrainian government to support cyber defence of civilian infrastructure.
NVIDIA and partners showcased AI deployment progress in Southeast Asia at Singapore AI Day. Singapore's HTX is researching public-safety AI based on Nemotron 3 Super and Nemotron 3 Nano Omni.
OpenAI and Grab jointly launched regional programme GO Forward with AI to help 30,000 Grab partners in Southeast Asia develop practical AI skills, focusing on real-world application.
oMLX creator and maintainer Jun Kim joined Hugging Face full-time to support MLX. oMLX moves from a side project to a funded long-term effort, remaining Apache 2.0 open-source under Jun's leadership to improve stability and accelerate development.
NVIDIA launched DSX Ready certification to validate partner products against its DSX AI factory reference design, initially covering battery energy storage systems (BESS) and coolant distribution units (CDU).
NVIDIA explains that deploying physical AI at scale requires safety covering hardware, software, AI behaviour and operating environments, and introduces its full-stack safety system NVIDIA Halos.
NVIDIA held an AI ecosystem event at the Grand Egyptian Museum, with Egyptian Deep Learning Institute participation growing over tenfold in a year. Hassan Allam received a data centre licence from Egypt's National Telecom Regulatory Authority and is developing a new centre with A15 with expected investment of $400 million. NVIDIA says four African AI factories have been announced or launched, with another 656 MW under construction.
OpenAI is working with an independent mathematics and AI advisory group to guide review and communication of emerging AI results in the field. Specific members and operational details have not been disclosed.
NVIDIA highlighted five AI clean-energy companies at New York Climate Week. ThinkLabs AI's NVIDIA CUDA-based digital twins and agents help Southern California Edison cut grid connection application assessments from 30–45 days to two minutes.
OpenAI and AARP will provide free hands-on ChatGPT workshops for 1,000 older adults in ten US cities, helping them safely develop practical AI skills.
Emerald AI, Google and NVIDIA announced the AI Energy Management Alliance (AEMA) to encourage data centres to dynamically adjust electricity use to grid conditions.
Mistral announced a partnership with Mozilla. Firefox's AI browsing assistant Smart Window (beta) is now powered by Mistral models, initially for France and North America, with the UK and Germany expected later this year.
A University of Manchester team used NVIDIA Earth-2's generative Earth-2 CorrDiff model and one year of hourly UK pollution data to train a nationwide air-pollution model at 2–3-square-kilometre resolution, completing training in two days on one eight-GPU Isambard-AI node.
Mistral and Cloudera partnered to integrate Mistral models into Cloudera's hybrid data platform. Enterprises can run inference in private or public clouds, on-premises and fully air-gapped environments with full control. They can train custom models on proprietary data in controlled environments while retaining ownership of data and resulting intelligence. Cloudera runs 30 exabytes of customer-managed data.
OpenAI partnered with the US General Services Administration (GSA) to offer eligible federal, state, local and tribal governments $0 licence fees, a 50% usage discount and expanded cyber-defence support.
Paul Christiano joined the OpenAI Foundation board and its Safety and Security Committee, bringing experience in AI alignment, safety and standards.
Mistral announced a €3 billion Series D round at a post-money valuation exceeding €21 billion, led by Samsung Electronics and co-led by Scaleup Europe Fund and PSG Equity. It called this the largest equity financing by a European technology company, funding frontier research, computing, infrastructure and commercial growth. Mistral operates in 20 countries and serves more than 125 global enterprises, including Airbus, ASML and HSBC.
OpenAI expanded journalism support with tools, training and partnerships for students, educators, journalists and news organisations, covering activities from classrooms to newsrooms.
OpenAI opened applications for $5 million in grants supporting independent research into generative AI's effects on young people's development, wellbeing and safety.
OpenAI, AIRPPU and WAN-IFRA launched an AI project to strengthen innovation, resilience and independent journalism in Ukrainian news organisations. The original article did not disclose specific tools, scale or availability details.
Google DeepMind announced the world’s first double-blind evaluation of proprietary frontier AI models, restricting external evaluations to cryptographically isolated environments to prevent models seeing test questions in advance. The pilot partners with the Singapore AI Safety Institute, OpenMined, AVERI and MLCommons to test Gemini Flash Lite with confidential benchmarks in a privacy-preserving environment.
Mistral and HUMAIN announced a strategic partnership covering infrastructure, advanced models and deployment, initially focusing on cybersecurity and voice, with frontier models strong in Arabic planned. Worth hundreds of millions of euros, it includes exploring HUMAIN data centres and joint market strategies for regulated Saudi industries.
Google DeepMind partnered with Fenris Creations to explore AI-driven gameplay prototypes in the EVE Universe, home to EVE Online.
Hugging Face ran the ICML 2026 Open Reproduction Challenge from 15 July to 2 August. Using coding agents such as Claude Code, Codex and Cursor, 1,221 community members reproduced papers and published 6,816 Trackio logs covering 2,226 papers, around a third of the conference total.
At DOE Genesis Mission Summit 2026, Google announced $40 million in AI tokens and cloud credits to support Genesis Mission researchers.
Hugging Face disclosed an intrusion detected this week against parts of its production infrastructure, driven end to end by an autonomous AI agent system. Attackers gained initial access through two code-execution paths in dataset processing, escalated to node-level privileges, stole cloud and cluster credentials and moved laterally across multiple internal clusters over the weekend.
Google DeepMind and A24 announced their first research-focused partnership, with long-term R&D across multiple projects and artists helping shape technologies and workflows. Google also invested in A24. Initial work will bridge frontier technology and next-generation entertainment, with specific goals and technical outputs evolving over time.
Berkeley Artificial Intelligence Research (BAIR) announced its 2026 PhD graduates, whose work covers robotics, LLM reasoning, computer vision, generative models, AI safety and AI for Science. Destinations include OpenAI, Mistral AI, Physical Intelligence and Amazon, faculty positions at UCLA and the University of Chicago, and start-ups.
Google DeepMind announced work with the UK government, Google Cloud, Faculty and planning departments in Barnet, Camden and Dorset on an AI planning-approval prototype, aiming to help officers reduce householder application processing time by 50%.
With Google's support, UC San Diego plans a data centre using motherboards from 2,000 retired Pixel phones to provide low-cost, low-carbon cloud computing for hundreds of students and faculty. SPEC benchmarks show 25–50 phones roughly equal one modern server. Phones form Kubernetes-managed clusters of 25–50 devices, with launch expected in autumn 2026.
Google DeepMind, Schmidt Sciences, Cooperative AI Foundation and ARIA, supported by Google.org, launched up to $10 million in global technical research funding focused on collective behaviour and safety risks in large-scale multi-agent AI systems.