Sep 30, 2026 | Gemini 4 Argon delivers frontier performance in complex workflows across real-world software engineering, enterprise knowledge work like legal and finance, and cybersecurity defense. Today, we’re announcing our new frontier model, Gemini 4 Argon, which is rolling out to a set of trusted cyber defenders through our Fairwind Program.
Microsoft Research released Quine, an AI research system for biological complexity, comprising a biological world model trained on multimodal data spanning sequences, structures, functions, cell states and imaging, and an interactive harness connecting the model, scientific tools, literature and experimental researchers.
Latent Space's AINews rounds up developments from 9/24–9/25, noting positive community feedback since Opus 5.5 launched this week, particularly on generating explainer videos with code.
AMD announced an all-stock acquisition of AI research lab World Labs, co-founded by Fei-Fei Li, for approximately $8.2 billion, with completion expected by year-end. Founded in 2024, World Labs reached a $1 billion valuation within months and launched its first commercial product, Marble, in 2025, generating interactive 3D worlds from prompts.
Google Research published research on long-form video generation, proposing an AI video co-director multi-agent orchestration framework built on Gemini and Veo to plan visual continuity across multi-shot narratives.
Google DeepMind released Gemini 3.8 Live with Live Avatar, adding near-real-time video generation to a native live conversation model to create a dynamic visual avatar that can hear, see and speak. It is available on Gemini Enterprise from today.
Meta positioned Muse as a personal agent at Connect 2026 and announced hardware and feature updates around it. Muse supports voice and live video, runs tasks in the background and will come to all Meta glasses. Each Muse has its own email address, the Mac version supports computer use, and its connector platform has more than 1,500 apps.
Google DeepMind released Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS. The former targets character design and line-by-line performance direction, while the latter targets high-concurrency voiceovers and voice agents.
Xiaomi released the MiMo-V2.6 family, including omnimodal models MiMo-V2.6-Pro and MiMo-V2.6-Flash, plus MiMo-V2.6-Pro-UltraSpeed with up to 20 times faster output.
Google DeepMind and Google Research released WeatherNext 3, calling it the most advanced and accurate global weather model available. It learns directly from real-time geostationary Earth-observation satellite data and generates hourly forecasts, with five-kilometre resolution for key surface variables and roughly five times the overall detail of WeatherNext 2.
Google DeepMind introduced agentic video understanding for Gemini 3.7 Flash, 3.6 Flash and 3.5 Flash-Lite, reducing video analysis token use by up to 88% and costs by up to 66%, while improving accuracy by up to 7%.
Google Research proposed PhotoScan, a deep learning framework that estimates body fat percentage, A/G ratio and V/S ratio from ordinary 2D phone photos to predict insulin resistance.
Google DeepMind released SL2T, a multilingual sign-language-to-text model, bringing the capability to consumer products for the first time. Gboard and Live Transcribe on Pixel 11 support ASL-to-English dictation, with more devices and languages to follow.
Google Research released AMIE (Video), built on Gemini and Project Astra, for real-time video clinical consultations. It can perceive non-verbal cues and guide virtual physical examinations.
NVIDIA released Magpie TTS Multilingual, a 364M-parameter open-weight speech synthesis model supporting English, Spanish, French, German, Italian, Vietnamese, Chinese, Hindi and Japanese, plus newly added Modern Standard Arabic, Korean and Brazilian Portuguese, for 12 languages in total.
Meta released Muse Glimmer, a multimodal model distilled from Muse to 30B parameters under Apache 2.0, targeting local agent use cases such as coding, document analysis and personal assistants.
Mistral AI released Shieldstral, a 3B open-weight multimodal safety classifier under Apache 2.0 that runs on one 16 GB NVIDIA GPU. It frames moderation as policy-adaptive question answering: policies written in natural language at inference time produce calibrated safety scores without retraining, handling text and images uniformly. Official claims say it matches models up to seven times larger on text safety and sets new best results on multimodal moderation benchmarks.
Google DeepMind released Gemini Robotics ER 2 as a high-level brain for robots, supporting video understanding, multi-step task orchestration and multi-robot collaboration, while delegating action execution to lower-level VLA models.
Google DeepMind released Gemini Robotics 2, a next-generation robot intelligence layer achieving whole-body control of a complete humanoid robot for the first time, with fine manipulation using both hands and grippers.