Skip to content

All AI news

4 today

Oct 1

ThursdayToday4 items
  1. Gemini 4 Argon: our next era of frontier intelligence

    Sep 30, 2026 | Gemini 4 Argon delivers frontier performance in complex workflows across real-world software engineering, enterprise knowledge work like legal and finance, and cybersecurity defense. Today, we’re announcing our new frontier model, Gemini 4 Argon, which is rolling out to a set of trusted cyber defenders through our Fairwind Program.

Sep 30

Wednesday
  1. Build a multi-agent music production pipeline on Amazon Bedrock AgentCore Runtime Instances

    The official AWS blog demonstrates deploying a three-agent music production pipeline on Amazon Bedrock AgentCore Runtime Instances. The composition agent uses Claude Sonnet 4.6 to generate a music brief and runs ACE-Step on the instance's NVIDIA L4 to render audio. The delivery agent reads .wav files from the shared volume for measurement and DSP processing, while the compliance agent independently remeasures them and checks harmonic similarity against a music catalogue.

  2. How Diffusion Controller unifies and simplifies AI image generation

    Google Research proposed Diffusion Controller, reframing diffusion-model denoising as a continuous control problem. A lightweight steering-damper network dynamically adjusts generation trajectories while the base model remains frozen. Evaluated on Stable Diffusion v1.4 using HPS-v2, its fully unlocked version achieved a 90% win rate against the baseline, and it supports customised control of closed models without access to internal weights.

  3. Prompt engineering fundamentals for Amazon Quick

    Prompt engineering determines the quality of Amazon Quick's AI responses to natural language requests. The first instalment of an official two-part series covers principles shared across components and reusable frameworks. It introduces CRISPE, covering context and constraints, roles and responsibilities, intent and inputs, steps and scope, and emphasises specificity, business context and examples over abstract descriptions.

Sep 29

Tuesday
  1. NVIDIA Kumo Tabular Sets a New Accuracy-Efficiency Frontier for Tabular Prediction

    NVIDIA released Kumo Tabular, an open-source tabular foundation model that predicts labels for new rows in a single forward pass given labelled rows, without training, tuning or feature engineering. It supports classification and regression, offers three sizes from 28M to 215M, and was pretrained solely on artificially generated tables. It uses the commercially usable OpenMDW-1.1 licence and ranks first on TabArena, BeyondArena, TALENT and ScoringBench.

  2. Getting the Source Right, Not Just the Fact: Source-Aware Verification for MCP Agents

    Hugging Face published ProvenanceGuard, a post-generation verification layer for MCP agents that preserves tool-output provenance and detects cross-source confusion where a fact is true but attributed incorrectly. Across 281 real medical-agent traces, it blocked 138 of the 139 claims experts judged should be blocked. Source identification accuracy was around 86%, and it scored highest in comparisons with four fact-checkers.

  3. Introducing GPT-6.1 Sol

    OpenAI has released GPT-6.1 Sol, which it says offers intelligence close to Astra and is aimed at coding, computer use and professional work. The model's API input and output token prices are one fifth of Astra's standard price.

  4. DevDay 2026 Recap

    OpenAI has published a DevDay 2026 recap rounding up more than 20 announcements spanning GPT-6 Astra, ChatGPT, Codex, the API, safety and new tools for developers. The original gives only an overview and does not detail each update.

  5. One year in: How Microsoft Research Asia – Singapore is advancing research, partnership and talent for real-world impact

    MSRA – Singapore, Microsoft's first research lab in Southeast Asia, spent its first year focusing on next-generation AI models and agent systems, domain AI, AI-native research practices and talent ecosystems. It works with Singapore's healthcare ecosystem on multimodal medical AI and self-evolving diagnostic agents, and jointly held a logistics and transport AI executive roundtable with EDB in February 2026.

  6. Generate images and video with vLLM-Omni on SageMaker AI – Part 2

    AWS published a tutorial deploying two endpoints from the same vLLM-Omni DLC image on SageMaker AI: a real-time endpoint running FLUX.2-klein-4B for images and an asynchronous endpoint running Wan2.1-VACE-1.3B for image-conditioned video. A text prompt first generates a PNG, which is sent with a motion prompt through Amazon S3 to the video endpoint; the resulting MP4 is retrieved from S3. The example includes a command-line workflow and an optional Streamlit interface.

Sep 28

Monday
  1. Hallo, Deutschland!

    Mistral opened a hub in Munich with a research team focused on Physics AI and industrial AI, and plans to build one gigawatt of European computing capacity by 2030. It will work with BMW on crash simulation and engineering AI, Siemens Energy on industrial AI applications, and the Technical University of Munich (TUM) on automotive aerodynamics digital twins using wind-tunnel facilities.

  2. Are you a Codex Original?

    OpenAI launched the next phase of Codex Originals, seeking real stories and projects from developers, researchers and creators using Codex. Interested participants can submit their experiences and project descriptions through the form below.

Sep 26

Saturday