Skip to content

News: selected

Gemini 4 Argon: our next era of frontier intelligence

Sep 30, 2026 | Gemini 4 Argon delivers frontier performance in complex workflows across real-world software engineering, enterprise knowledge work like legal and finance, and cybersecurity defense. Today, we’re announcing our new frontier model, Gemini 4 Argon, which is rolling out to a set of trusted cyber defenders through our Fairwind Program.

Google DeepMind

Build a multi-agent music production pipeline on Amazon Bedrock AgentCore Runtime Instances

The official AWS blog demonstrates deploying a three-agent music production pipeline on Amazon Bedrock AgentCore Runtime Instances. The composition agent uses Claude Sonnet 4.6 to generate a music brief and runs ACE-Step on the instance's NVIDIA L4 to render audio. The delivery agent reads .wav files from the shared volume for measurement and DSP processing, while the compliance agent independently remeasures them and checks harmonic similarity against a music catalogue.

AWS Machine Learning Blog

  1. From Training to Production, NVIDIA and CoreWeave Close the Loop on Agentic AI

    CoreWeave announced that NVIDIA Vera Rubin NVL72 systems with Spectrum-X 102.4T Ethernet were available on CoreWeave Cloud, with Cognition becoming the first customer to run them in production.

    NVIDIA BlogIndustry
  2. Introducing SynthID Bio

    Google DeepMind has released SynthID Bio, bringing watermarking technology to synthetic biology. It embeds imperceptible signatures in biological code so that watermarks can be verified in both digital models and the physical proteins synthesised from them, while preserving their biological function.

    Google DeepMindModels
  3. Bring near-Astra intelligence to everyday work with GPT-6.1 Sol on Amazon Bedrock

    GPT-6.1 Sol is now generally available on Amazon Bedrock, offering stronger reasoning for agentic coding, computer use and everyday professional workloads. OpenAI says it matches GPT-6 Astra on DeepSWE v1.1 at roughly one-fifth the cost per task, and exceeds GPT-6 Sol’s best result by 6.4 percentage points.

    AWS Machine Learning BlogProducts

DevDay 2026 Recap

OpenAI has published a DevDay 2026 recap rounding up more than 20 announcements spanning GPT-6 Astra, ChatGPT, Codex, the API, safety and new tools for developers. The original gives only an overview and does not detail each update.

OpenAI News

  1. NVIDIA Kumo Tabular Sets a New Accuracy-Efficiency Frontier for Tabular Prediction

    NVIDIA released Kumo Tabular, an open-source tabular foundation model that predicts labels for new rows in a single forward pass given labelled rows, without training, tuning or feature engineering. It supports classification and regression, offers three sizes from 28M to 215M, and was pretrained solely on artificially generated tables. It uses the commercially usable OpenMDW-1.1 licence and ranks first on TabArena, BeyondArena, TALENT and ScoringBench.

    Hugging Face BlogModels
  2. Introducing Quine: An AI research system designed for the complexity of biology

    Microsoft Research released Quine, an AI research system for biological complexity, comprising a biological world model trained on multimodal data spanning sequences, structures, functions, cell states and imaging, and an interactive harness connecting the model, scientific tools, literature and experimental researchers.

    Microsoft ResearchModels
  3. Introducing dots

    OpenAI has released dots, a proactive assistant that keeps work moving on complex projects and everyday tasks. The company says dots lets users stay in control as work progresses.

    OpenAI NewsProducts
  4. Grok 4.7 is now available on Amazon Bedrock

    xAI’s Grok 4.7 is available on Amazon Bedrock with a 500K-token context window, four configurable reasoning levels—low, medium, high and xhigh—and support for the Responses, Chat Completions and Converse APIs.

    AWS Machine Learning BlogModels
  5. Introducing Claude Sonnet 5.5 on AWS

    AWS announced Claude Sonnet 5.5 availability on Amazon Bedrock and Claude Platform on AWS, positioning it as a more efficient Sonnet for coding and knowledge work, with lower cost per task and faster speeds for most tasks.

    AWS Machine Learning BlogModels

Holo4: powering generalist computer-use agents

H released the Holo4 agent model family, comprising 27B dense and 35B-A3B MoE versions. Both are available through H Models API, with weights open-sourced on Hugging Face in BF16, FP8, NVFP4 and 4-bit GGUF formats.

Hugging Face Blog

Introducing Gemini 3.8 Live with Live Avatar

Google DeepMind released Gemini 3.8 Live with Live Avatar, adding near-real-time video generation to a native live conversation model to create a dynamic visual avatar that can hear, see and speak. It is available on Gemini Enterprise from today.

Google DeepMind

  1. Automating coherent long-form video generation

    Google Research published research on long-form video generation, proposing an AI video co-director multi-agent orchestration framework built on Gemini and Veo to plan visual continuity across multi-shot narratives.

    Google ResearchPapers
  2. AI-powered fuzzing with the GitHub Security Lab Taskflow Agent

    GitHub Security Lab released Fuzzing Taskflow, an autonomous fuzzing pipeline for C/C++ projects. Given a GitHub repository, it identifies entry points, analyses build systems, writes harnesses, runs AFL++, reads coverage reports and improves harnesses, then classifies every crash and generates vulnerability reports.

    GitHub Blog · AI & MLProducts
  1. Offloaded inference for real-world physical AI robotics

    Microsoft Research systematically measured mobile manipulation robot workloads and found that offloading physical AI inference from onboard GPUs to edge or cloud GPUs can improve task success, support larger models and extend battery life.

    Microsoft ResearchPapers
  2. Advancing Private AI Compute with secure, server-side memory

    Google DeepMind announced an update to Private AI Compute that brings persistent, cross-device AI memory to the cloud while maintaining device-level privacy standards. Data is sealed in encrypted storage, with decryption keys retained only on user devices. When a model needs access, an end-to-end encrypted channel connects to a cloud secure enclave, where data is temporarily decrypted in isolated memory, then re-encrypted immediately after new context is saved.

    Google DeepMindIndustry

Introducing GPT-6 Sol and Luna

OpenAI released GPT-6 Sol and Luna, bringing frontier intelligence to everyday work with different balances between capabilities and cost.

OpenAI News

  1. Gemini 3.8 text-to-speech says hello

    Google DeepMind released Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS. The former targets character design and line-by-line performance direction, while the latter targets high-concurrency voiceovers and voice agents.

    Google DeepMindModels