Introducing the Agents API
OpenAI released the Agents API, a managed service powered by the Codex harness for building and launching cloud agents. It supports orchestration, long-running sessions and tool calls.
OpenAI released the Agents API, a managed service powered by the Codex harness for building and launching cloud agents. It supports orchestration, long-running sessions and tool calls.
An MIT researcher used GPT-5.6 Sol with Codex to autonomously run quantum computing experiments, analyse results and calibrate qubits.
OpenAI released ChatGPT Images 2.5, turning ideas, sketches and reference photos into more personalised, polished images. The company says the results better reflect users’ ideas.
1Password engineers use Codex to rapidly build production-ready features and internal tools while maintaining strict security policies, improving engineering efficiency by 21%.
GitHub released a research preview of Project HydraFusion, using runtime multi-model orchestration to deliver frontier-level coding through Copilot. Users can enable it via /experimental in GitHub Copilot CLI, paying each model's standard rates for tokens actually consumed.
Copilot app runs multiple agent sessions simultaneously in separate Git worktrees, preserving context without interference. The sessions view shows titles and progress, allowing switching and resuming without restating tasks. The article demonstrates funded sort development, accessibility review and testing in parallel in tailspin-toys.
Hugging Face released funes, an open-source persistent memory layer for Claude Code, Codex, pi, Hermes and other coding agents. It indexes existing local sessions and performs embedding and reranking locally by default.
Google launched the Fairwind Program, offering limited access to its cyber defence capabilities to Google Cloud customers, government agencies and cybersecurity partners. The initial offering includes Gemini 3.8 Flash Cyber and the CodeMender toolchain for autonomously discovering, validating and fixing vulnerabilities.
Google DeepMind introduced agentic video understanding for Gemini 3.7 Flash, 3.6 Flash and 3.5 Flash-Lite, reducing video analysis token use by up to 88% and costs by up to 66%, while improving accuracy by up to 7%.
Hugging Face’s WebAI team released @huggingface/kernels, a lightweight library for loading and running optimised WebGPU kernels from the Hugging Face Hub. An initial collection of 207 kernels is also available as separate repositories under Apache-2.0.
Hugging Face introduced gr.Workflow in Gradio to describe multi-step AI pipelines as graphs of typed nodes. Gradio provides a draggable canvas where each node can run and every intermediate result is visible.
Mistral released Agentic Search, a retrieval layer that lets models find, examine and verify information in multi-step retrieval loops. It is provided through Mistral Search Toolkit and built into Libraries in Studio and Vibe.
Hugging Face's OlmoEarth Studio now computes and exports embeddings. Users choose an area, period, encoder variant and resolution through UI or API and receive Cloud-Optimized GeoTIFF (COG) results.
Mistral announced three AI sovereignty initiatives. Mistral Regional Endpoints are generally available, letting customers choose inference in Europe or the US. Mistral Priority Tier is in public preview, offering committed service levels, custom rate limits and uptime SLAs for critical workloads.
Baseten became a new Inference Provider on Hugging Face Hub, initially offering conversation and text generation with open-weight models including Kimi K3, DeepSeek V4 Flash and GLM-5.2.
Microsoft Research released the open-source Orchard framework, centred on the Kubernetes environment service Orchard Env. It reuses environments, data pipelines and evaluation workflows across tasks and supports training agents directly within real deployment frameworks such as Codex, OpenClaw and ZeroClaw.
Hugging Face introduced Nunchaku Lite into Diffusers, allowing Nunchaku quantised checkpoints to load directly with from_pretrained(), without custom pipelines or local CUDA compilation.
Hugging Face released Grabette, an open-source system recording manipulation demonstrations with a handheld gripper and two cameras, producing robot-ready datasets without robots or teleoperation equipment. The handheld hardware costs about €490 in materials, with the accompanying motorised Gripette gripper around €120. Hardware CAD, Raspberry Pi collection software and browser-based processing are all open-source.
Google DeepMind launched the ATL Saathi pilot in India, a Gemini-powered web app providing Atal Tinkering Labs educators with a 24/7 lesson preparation and training assistant, initially covering 100 schools.
Mistral Studio now provides central management of Prompts and Skills, with version histories, clear ownership and traceability. Immutable versions, rollback, tags and audit logs make AI behaviour governable and discoverable, while Observability traces production outputs to their versions. Skills can be invoked directly from Studio as MCP servers, ensuring production executes the same governed asset.
Hugging Face announced that vLLM’s transformers modelling backend now matches or exceeds the throughput of vLLM’s handwritten native implementations across several LLM architectures. It uses torch.fx for static graph analysis and ast to rewrite source code, dynamically applying inference-related layer fusion at runtime to match custom-code performance.
Hugging Face and Amazon SageMaker AI introduced deep-link integration. Clicking Customize on SageMaker AI or Deploy on SageMaker AI on supported model pages opens SageMaker Studio with the model preloaded and environment configured automatically.
At Build 2026, Microsoft announced Foundry Managed Compute and a Hugging Face model collection, with weekly updates to open-weight models and one-click deployment to Foundry's managed GPU platform.
Hugging Face released LeRobot v0.6.0, introducing world model policies VLA-JEPA, FastWAM and LingBot-VA, new VLAs GR00T N1.7, MolmoAct2, EO-1, EVO1 and Multitask DiT, and a unified reward model API with Robometer and TOPReward for the robotics learning loop.
Hugging Face and SkyPilot released an integration allowing a Hugging Face Bucket or any model, dataset or Space repository to be mounted into SkyPilot tasks with an hf:// URL and an existing HF_TOKEN, running across more than 20 clouds, Kubernetes, Slurm and local environments.
Hugging Face introduced a new kernel repository type for Kernels, along with trusted publishers and code signing. By default, only trusted publishers' kernels load; other sources require explicit trust_remote_code=True.
Hugging Face and Cerebras jointly demonstrated a real-time speech-to-speech pipeline, accelerating Gemma 4 31B inference with Cerebras and combining Nvidia Parakeet speech recognition with Alibaba Qwen3TTS synthesis.
Google Research expanded building-level roof reflectivity (albedo) data from 14 cities in 2024 to over 50 cities in nine countries, accessible through the new Heat Resilience Earth Engine App.
Every Eval Ever (EEE) and Hugging Face Community Evals are now interoperable, allowing cross-publication of evaluations with links to complete records. EEE stores around 229,000 results across more than 22,000 models and 2,200 benchmarks from 31 reporting formats.
Google DeepMind integrated computer use as a built-in tool in Gemini 3.5 Flash. Previously, it was available only as the standalone Gemini 2.5 computer use model.
Mistral AI added new Connectors capabilities. Admin controls for workspace- or organisation-level access and individual tool toggles are generally available, as are API keys scoped to connectors.
Google launched a public preview of Agentic RAG-powered cross-corpus retrieval on Gemini Enterprise Agent Platform. Roles including Orchestrator, Planner, Query Rewriter, Search Fanout and Sufficient Context Agent collaborate on multi-source, multi-hop queries.
Google Research open-sourced its hydrological modelling framework on GitHub under Apache 2.0, enabling national weather and hydrology agencies to integrate AI flood forecasting. The Python package uses PyTorch and an LSTM architecture, can train or fine-tune on Caravan data, and includes interactive tutorial notebooks and videos.
Google Research unveiled the experimental Gemini for Science toolkit at I/O 2026, including Computational Discovery based on ERA and AlphaEvolve, Hypothesis Generation based on Co-Scientist and Literature Insights based on NotebookLM.
Mistral released an AI stack for industrial engineering at AI Now Summit 2026, partnering with Airbus, BMW and ASML to optimise design, simulation and production while retaining control over proprietary data and IP.
Mistral upgraded Le Chat to Vibe, a unified AI agent covering work and coding, retaining all existing conversations, settings and subscription plans.
Mistral AI released the public preview of Search Toolkit, a composable framework for production search pipelines in AI applications. It unifies ingestion, retrieval and evaluation through shared interfaces, is open-source and deploys in cloud, local or edge environments.
After incorporating Emmi AI, Mistral launched physical AI capabilities for AI-native industrial engineering with partners including ASML, Airbus, Safran and Siemens Energy. The model predicts physical fields directly from geometry and boundary conditions in seconds through one forward pass on a single GPU, versus hours to weeks per design variant in traditional CFD/FEM. It accelerates design iteration while retaining traditional solvers for validation and edge cases.
Mistral released Connectors in Studio. Built-in connectors and custom MCP are available through API/SDK for all model and agent calls, currently in public preview.
Google DeepMind's Co-Scientist accelerates cellular ageing reversal research, scanning tens of thousands of papers to propose over 20 testable new genetic factors. Several were validated in labs as driving cells towards younger states and improving overall function. It also cuts screening data analysis from six months to days.