Cognition helps Devin test its own work with GPT‑6 Astra
Cognition uses GPT-6 Astra to improve Devin's ability to test software and demonstrate that it works, aiming to reduce engineers' code review burden and speed up delivery.
Cognition uses GPT-6 Astra to improve Devin's ability to test software and demonstrate that it works, aiming to reduce engineers' code review burden and speed up delivery.
César de la Fuente's lab uses Codex and ChatGPT to search the genomes of living and extinct organisms for candidate antimicrobial molecules to combat drug-resistant infections.
OpenAI announced Data agent in ChatGPT Work is available to everyone, connecting enterprise data, discovering insights and building interactive dashboards with AI through natural language.
OpenAI launched ChatGPT for financial services, with built-in financial data and access to GPT-6 Astra for research, modelling and creating client-ready materials.
OpenAI released the Agents API, a managed service powered by the Codex harness for building and launching cloud agents. It supports orchestration, long-running sessions and tool calls.
An MIT researcher used GPT-5.6 Sol with Codex to autonomously run quantum computing experiments, analyse results and calibrate qubits.
Google DeepMind launched AlphaGenome Atlas, a platform containing predicted effects for 9 billion single-nucleotide variants across the human genome. At 1 PB, it is over 30 times the size of the AlphaFold Database.
OpenAI released ChatGPT Images 2.5, turning ideas, sketches and reference photos into more personalised, polished images. The company says the results better reflect users’ ideas.
1Password engineers use Codex to rapidly build production-ready features and internal tools while maintaining strict security policies, improving engineering efficiency by 21%.
GitHub released a research preview of Project HydraFusion, using runtime multi-model orchestration to deliver frontier-level coding through Copilot. Users can enable it via /experimental in GitHub Copilot CLI, paying each model's standard rates for tokens actually consumed.
Copilot app runs multiple agent sessions simultaneously in separate Git worktrees, preserving context without interference. The sessions view shows titles and progress, allowing switching and resuming without restating tasks. The article demonstrates funded sort development, accessibility review and testing in parallel in tailspin-toys.
Hugging Face released funes, an open-source persistent memory layer for Claude Code, Codex, pi, Hermes and other coding agents. It indexes existing local sessions and performs embedding and reranking locally by default.
Google launched the Fairwind Program, offering limited access to its cyber defence capabilities to Google Cloud customers, government agencies and cybersecurity partners. The initial offering includes Gemini 3.8 Flash Cyber and the CodeMender toolchain for autonomously discovering, validating and fixing vulnerabilities.
Google DeepMind introduced agentic video understanding for Gemini 3.7 Flash, 3.6 Flash and 3.5 Flash-Lite, reducing video analysis token use by up to 88% and costs by up to 66%, while improving accuracy by up to 7%.
Hugging Face’s WebAI team released @huggingface/kernels, a lightweight library for loading and running optimised WebGPU kernels from the Hugging Face Hub. An initial collection of 207 kernels is also available as separate repositories under Apache-2.0.
Voice Arena and Hugging Face added Monsoon en-IN and Monsoon hi-IN evaluation datasets to Open ASR Leaderboard, making Hindi its first Indian language.
Hugging Face introduced gr.Workflow in Gradio to describe multi-step AI pipelines as graphs of typed nodes. Gradio provides a draggable canvas where each node can run and every intermediate result is visible.
Liquid AI released DSpark draft model checkpoints for LFM2.5-1.2B-Instruct, LFM2.5-2.6B and LFM2.5-8B-A1B. Speculative decoding accelerates decoding without changing output quality, increasing throughput by up to 3.18 times on GPUs and 2.87 times on devices.
Mistral released Agentic Search, a retrieval layer that lets models find, examine and verify information in multi-step retrieval loops. It is provided through Mistral Search Toolkit and built into Libraries in Studio and Vibe.
Sentence Transformers v6.0 adds a fourth model type, MultiVectorEncoder, for ColBERT-style late-interaction retrieval. It directly loads PyLate, Stanford-NLP ColBERT checkpoints and visual document retrieval models from colpali-engine.
Hugging Face's OlmoEarth Studio now computes and exports embeddings. Users choose an area, period, encoder variant and resolution through UI or API and receive Cloud-Optimized GeoTIFF (COG) results.
Mistral announced three AI sovereignty initiatives. Mistral Regional Endpoints are generally available, letting customers choose inference in Europe or the US. Mistral Priority Tier is in public preview, offering committed service levels, custom rate limits and uptime SLAs for critical workloads.
NVIDIA released Magpie TTS Multilingual, a 364M-parameter open-weight speech synthesis model supporting English, Spanish, French, German, Italian, Vietnamese, Chinese, Hindi and Japanese, plus newly added Modern Standard Arabic, Korean and Brazilian Portuguese, for 12 languages in total.
Baseten became a new Inference Provider on Hugging Face Hub, initially offering conversation and text generation with open-weight models including Kimi K3, DeepSeek V4 Flash and GLM-5.2.
Microsoft Research released the open-source Orchard framework, centred on the Kubernetes environment service Orchard Env. It reuses environments, data pipelines and evaluation workflows across tasks and supports training agents directly within real deployment frameworks such as Codex, OpenClaw and ZeroClaw.
Hugging Face introduced Nunchaku Lite into Diffusers, allowing Nunchaku quantised checkpoints to load directly with from_pretrained(), without custom pipelines or local CUDA compilation.
Hugging Face released Grabette, an open-source system recording manipulation demonstrations with a handheld gripper and two cameras, producing robot-ready datasets without robots or teleoperation equipment. The handheld hardware costs about €490 in materials, with the accompanying motorised Gripette gripper around €120. Hardware CAD, Raspberry Pi collection software and browser-based processing are all open-source.
Google DeepMind launched the ATL Saathi pilot in India, a Gemini-powered web app providing Atal Tinkering Labs educators with a 24/7 lesson preparation and training assistant, initially covering 100 schools.
Mistral Studio now provides central management of Prompts and Skills, with version histories, clear ownership and traceability. Immutable versions, rollback, tags and audit logs make AI behaviour governable and discoverable, while Observability traces production outputs to their versions. Skills can be invoked directly from Studio as MCP servers, ensuring production executes the same governed asset.
Hugging Face announced that vLLM’s transformers modelling backend now matches or exceeds the throughput of vLLM’s handwritten native implementations across several LLM architectures. It uses torch.fx for static graph analysis and ast to rewrite source code, dynamically applying inference-related layer fusion at runtime to match custom-code performance.
Hugging Face and Amazon SageMaker AI introduced deep-link integration. Clicking Customize on SageMaker AI or Deploy on SageMaker AI on supported model pages opens SageMaker Studio with the model preloaded and environment configured automatically.
At Build 2026, Microsoft announced Foundry Managed Compute and a Hugging Face model collection, with weekly updates to open-weight models and one-click deployment to Foundry's managed GPU platform.
Hugging Face released LeRobot v0.6.0, introducing world model policies VLA-JEPA, FastWAM and LingBot-VA, new VLAs GR00T N1.7, MolmoAct2, EO-1, EVO1 and Multitask DiT, and a unified reward model API with Robometer and TOPReward for the robotics learning loop.
Hugging Face and SkyPilot released an integration allowing a Hugging Face Bucket or any model, dataset or Space repository to be mounted into SkyPilot tasks with an hf:// URL and an existing HF_TOKEN, running across more than 20 clouds, Kubernetes, Slurm and local environments.
Hugging Face introduced a new kernel repository type for Kernels, along with trusted publishers and code signing. By default, only trusted publishers' kernels load; other sources require explicit trust_remote_code=True.
Hugging Face and Cerebras jointly demonstrated a real-time speech-to-speech pipeline, accelerating Gemma 4 31B inference with Cerebras and combining Nvidia Parakeet speech recognition with Alibaba Qwen3TTS synthesis.
Google Research expanded building-level roof reflectivity (albedo) data from 14 cities in 2024 to over 50 cities in nine countries, accessible through the new Heat Resilience Earth Engine App.
Every Eval Ever (EEE) and Hugging Face Community Evals are now interoperable, allowing cross-publication of evaluations with links to complete records. EEE stores around 229,000 results across more than 22,000 models and 2,200 benchmarks from 31 reporting formats.
Google DeepMind integrated computer use as a built-in tool in Gemini 3.5 Flash. Previously, it was available only as the standalone Gemini 2.5 computer use model.
Mistral AI added new Connectors capabilities. Admin controls for workspace- or organisation-level access and individual tool toggles are generally available, as are API keys scoped to connectors.