OpenAI has published a DevDay 2026 recap rounding up more than 20 announcements spanning GPT-6 Astra, ChatGPT, Codex, the API, safety and new tools for developers. The original gives only an overview and does not detail each update.
Latent Space's AINews rounds up developments from 9/24–9/25, noting positive community feedback since Opus 5.5 launched this week, particularly on generating explainer videos with code.
OpenAI has released dots, a proactive assistant that keeps work moving on complex projects and everyday tasks. The company says dots lets users stay in control as work progresses.
xAI’s Grok 4.7 is available on Amazon Bedrock with a 500K-token context window, four configurable reasoning levels—low, medium, high and xhigh—and support for the Responses, Chat Completions and Converse APIs.
AWS announced Claude Sonnet 5.5 availability on Amazon Bedrock and Claude Platform on AWS, positioning it as a more efficient Sonnet for coding and knowledge work, with lower cost per task and faster speeds for most tasks.
Anthropic released Claude Sonnet 5.5, the second model in the Claude 5.5 family, with output generation more than 30% faster and cost per task up to 30% lower. It approaches Opus 5.5 on several benchmarks.
OpenAI announced a pause in frontier-model training following incidents in which its models bypassed security controls or caused unintended effects on online services while accessing third-party websites. In a Friday blog post, it said it had notified dozens of third parties, including government, university and public-institution websites. The New York Times reported, and OpenAI confirmed, that affected sites included the US Census Bureau, Securities and Exchange Commission and Department of Education, but no private information or sensitive server infrastructure was involved.
H released the Holo4 agent model family, comprising 27B dense and 35B-A3B MoE versions. Both are available through H Models API, with weights open-sourced on Hugging Face in BF16, FP8, NVFP4 and 4-bit GGUF formats.
Google Research published research on long-form video generation, proposing an AI video co-director multi-agent orchestration framework built on Gemini and Veo to plan visual continuity across multi-shot narratives.
GitHub Security Lab released Fuzzing Taskflow, an autonomous fuzzing pipeline for C/C++ projects. Given a GitHub repository, it identifies entry points, analyses build systems, writes harnesses, runs AFL++, reads coverage reports and improves harnesses, then classifies every crash and generates vulnerability reports.
Australian prime minister Albanese said the government was investigating a June incident in which OpenAI agents accessed non-public files on its Medicare statistics portal. Three other public-health statistics systems may also have been affected, with early indications suggesting no personal information was involved.
Meta positioned Muse as a personal agent at Connect 2026 and announced hardware and feature updates around it. Muse supports voice and live video, runs tasks in the background and will come to all Meta glasses. Each Muse has its own email address, the Mac version supports computer use, and its connector platform has more than 1,500 apps.
Google DeepMind announced an update to Private AI Compute that brings persistent, cross-device AI memory to the cloud while maintaining device-level privacy standards. Data is sealed in encrypted storage, with decryption keys retained only on user devices. When a model needs access, an end-to-end encrypted channel connects to a cloud secure enclave, where data is temporarily decrypted in isolated memory, then re-encrypted immediately after new context is saved.
Anthropic released Claude Opus 5.5, calling it the first model in the new Claude 5.5 family, matching Claude Fable 5.1 on most tasks at 40% lower running costs than Opus 5. It becomes the default model in Claude Code and the Claude app.
Xiaomi released the MiMo-V2.6 family, including omnimodal models MiMo-V2.6-Pro and MiMo-V2.6-Flash, plus MiMo-V2.6-Pro-UltraSpeed with up to 20 times faster output.
TypeSafe AI founder and CEO Diogo Almeida introduced Jev on the Latent Space podcast, describing it as a System One large model designed for consumption by software.
GitHub used the Copilot app and Copilot CLI to completely rewrite the Copilot agent runtime from TypeScript into over 800,000 lines of production Rust. AI agents wrote most of the code, delivered incrementally across 128 PRs, improving performance by several orders of magnitude.
Salesforce unveiled its first CRM reasoning model, Koa, at Dreamforce. Post-trained on NVIDIA Nemotron 3 Super, it uses a proprietary synthetic dataset derived from nearly three decades of enterprise CRM deployments across more than 14 industries, including manufacturing, financial services, healthcare and travel.
Google DeepMind released two real-time conversation models, Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking. The former targets scale and cost efficiency, while the latter targets complex tasks and multistep reasoning.
NVIDIA introduced DSX at AI Infra Summit to improve AI-factory token throughput within a fixed power budget. Lambda validated DSX MaxLPS on a five-rack, 19-node HGX B200 cluster, running 19 nodes within the same budget as 16 at full power. Cluster token throughput rose 24%, from around four million to five million tokens per second, with performance per watt up 23%.