Cognition helps Devin test its own work with GPT‑6 Astra
Cognition uses GPT-6 Astra to improve Devin's ability to test software and demonstrate that it works, aiming to reduce engineers' code review burden and speed up delivery.
Cognition uses GPT-6 Astra to improve Devin's ability to test software and demonstrate that it works, aiming to reduce engineers' code review burden and speed up delivery.
OpenAI evolved Habitat from a Python library into a globally distributed storage platform supporting one billion ChatGPT users and 22 million requests per second, meeting the scaling demands of ChatGPT's online storage.
At ACL 2026, Google Research proposed ToolGrad, an answer-first, question-later approach that generates real tool-call chains before deriving user queries, replacing traditional DFS trial-and-error annotation.
Copilot's built-in diff, terminal and browser panels enable review, execution and preview of AI code without switching applications. The diff highlights additions in green and deletions in red, supporting acceptance, comments or further edits. The terminal runs project commands with multiple windows, while the browser's Pick & Polish selects elements for agent adjustments.
César de la Fuente's lab uses Codex and ChatGPT to search the genomes of living and extinct organisms for candidate antimicrobial molecules to combat drug-resistant infections.
OpenAI announced Data agent in ChatGPT Work is available to everyone, connecting enterprise data, discovering insights and building interactive dashboards with AI through natural language.
Mistral and Cloudera partnered to integrate Mistral models into Cloudera's hybrid data platform. Enterprises can run inference in private or public clouds, on-premises and fully air-gapped environments with full control. They can train custom models on proprietary data in controlled environments while retaining ownership of data and resulting intelligence. Cloudera runs 30 exabytes of customer-managed data.
OpenAI launched ChatGPT for financial services, with built-in financial data and access to GPT-6 Astra for research, modelling and creating client-ready materials.
OpenAI partnered with the US General Services Administration (GSA) to offer eligible federal, state, local and tribal governments $0 licence fees, a 50% usage discount and expanded cyber-defence support.
TRL v1.14's AsyncGRPOTrainer now supports training and synchronising only LoRA adapters to vLLM. A rank-1 adapter is just a few MB and can transfer through a Storage Bucket mounted to each job, without NCCL.
The Hugging Face team rebuilt most AUTOMATIC1111 functionality as the Workflow1111 canvas using Gradio Workflow. Its 11 media pipelines and 73 nodes cover text-to-image, high-resolution fixes, image-to-image, prompt matrices, VLM reverse prompting, detection-generated inpainting masks, ControlNet-style annotators, background removal, PNG Info and image-to-video.
OpenAI launched GPT-Live-1 in the API, supporting natural full-duplex voice conversations with stronger instruction following, custom voices and telephony support.
OpenAI released the Agents API, a managed service powered by the Codex harness for building and launching cloud agents. It supports orchestration, long-running sessions and tool calls.
Paul Christiano joined the OpenAI Foundation board and its Safety and Security Committee, bringing experience in AI alignment, safety and standards.
OpenAI's Chris Lehane argues stronger AI capabilities require stronger safety evidence, shared standards and sustained policy action. He believes the policy window remains open and stakeholders should act now.
Mistral helped a European energy operator migrate 40,000 lines of Fortran 77 to C++, targeting a physics-heavy reservoir simulator without a test suite or centralised documentation. The team first built a numerical-alignment testing framework, used Skill.md to guide agents in exporting state snapshots and validating migrated modules, then launched over a hundred agents with Vibe CLI to analyse call trees and used Mistral OCR to organise scattered documents.
OpenAI released GPT-6 Astra, calling it its most capable enterprise model, with advanced reasoning, computer use and stronger writing and design judgement.
An MIT researcher used GPT-5.6 Sol with Codex to autonomously run quantum computing experiments, analyse results and calibrate qubits.
Google DeepMind launched AlphaGenome Atlas, a platform containing predicted effects for 9 billion single-nucleotide variants across the human genome. At 1 PB, it is over 30 times the size of the AlphaFold Database.
OpenAI explores how more capable, affordable AI expands what individuals and businesses can accomplish and makes growth more economical. It focuses on capability gains and falling costs, explaining their effects on practical work output and business growth.
Mistral announced a €3 billion Series D round at a post-money valuation exceeding €21 billion, led by Samsung Electronics and co-led by Scaleup Europe Fund and PSG Equity. It called this the largest equity financing by a European technology company, funding frontier research, computing, infrastructure and commercial growth. Mistral operates in 20 countries and serves more than 125 global enterprises, including Airbus, ASML and HSBC.
OpenAI released ChatGPT Images 2.5, turning ideas, sketches and reference photos into more personalised, polished images. The company says the results better reflect users’ ideas.
OpenAI expanded journalism support with tools, training and partnerships for students, educators, journalists and news organisations, covering activities from classrooms to newsrooms.
OpenAI shared an AI-generated solution to the Navier–Stokes Millennium Prize Problem, including a written explanation and a formal Lean proof.
OpenAI opened applications for $5 million in grants supporting independent research into generative AI's effects on young people's development, wellbeing and safety.
1Password engineers use Codex to rapidly build production-ready features and internal tools while maintaining strict security policies, improving engineering efficiency by 21%.
OpenAI, AIRPPU and WAN-IFRA launched an AI project to strengthen innovation, resilience and independent journalism in Ukrainian news organisations. The original article did not disclose specific tools, scale or availability details.
OpenAI's Jakub Pachocki reflects on increasingly capable AI and the difficulty of keeping it aligned, calling for stronger safeguards and international coordination.
OpenAI disclosed internal data showing coding agents reshaping AI research, covering agent usage, experiment speed, task complexity and research acceleration.
GitHub released a research preview of Project HydraFusion, using runtime multi-model orchestration to deliver frontier-level coding through Copilot. Users can enable it via /experimental in GitHub Copilot CLI, paying each model's standard rates for tokens actually consumed.
Google Research evaluated cross-population transfer of polygenic risk scores (PRS) on eight clinical measures using European UK Biobank data and nearly 200,000 Japanese participants from Biobank Japan. At target-population samples of 15,000 or more, target-specific models outperformed mixed European-data training, while highly genetically correlated traits continued to benefit from European data until target samples reached 25,000–40,000 or more.
Copilot app runs multiple agent sessions simultaneously in separate Git worktrees, preserving context without interference. The sessions view shows titles and progress, allowing switching and resuming without restating tasks. The article demonstrates funded sort development, accessibility review and testing in parallel in tailspin-toys.
Google Research partnered with HHMI Janelia, the University of Cambridge and others to publish in Cell a complete connectome of a male fruit fly’s brain and central nervous system. It contains over 166,000 neurons and 125 million synaptic connections, making it the largest brain map to date by neuron count.
Google DeepMind and Google Research released WeatherNext 3, calling it the most advanced and accurate global weather model available. It learns directly from real-time geostationary Earth-observation satellite data and generates hourly forecasts, with five-kilometre resolution for key surface variables and roughly five times the overall detail of WeatherNext 2.
Hugging Face launched NeoMME in 260M and 800M sizes. A single bidirectional Transformer handles text tokens and 32×32 image patches, trained from scratch with a masked discrete diffusion objective rather than pretrained vision towers or causal language models.
Hugging Face released funes, an open-source persistent memory layer for Claude Code, Codex, pi, Hermes and other coding agents. It indexes existing local sessions and performs embedding and reranking locally by default.
A public low-cost approach fine-tunes LFM2.5-350M with TRL GRPO using around 500 samples and 100 steps, raising IFStruct from 22.6% to 29.7%. Training fits free Colab or Kaggle GPUs, with local llama.cpp evaluation on a MacBook and code open on GitHub.
A Hugging Face blogger reproduced Surya Narreddi's idea of having a language model paint watercolours using TRL and OpenEnv. The model paints through around 150 lines of JavaScript with p5.brush, with datasets, RL environments, training scripts and models all open-source.
Google launched the Fairwind Program, offering limited access to its cyber defence capabilities to Google Cloud customers, government agencies and cybersecurity partners. The initial offering includes Gemini 3.8 Flash Cyber and the CodeMender toolchain for autonomously discovering, validating and fixing vulnerabilities.
Google DeepMind released Gemini 3.8 Flash and Gemini 3.8 Flash Cyber. The former targets long-horizon coding and autonomous agents, priced like 3.7 Flash at $0.75 per million input tokens and $3.75 per million output tokens.