Trump plan to combat AI risks hinges on Big Tech pals policing themselves
The Trump administration has pushed dozens of AI companies to agree to voluntary safety testing, a plan that relies on Big Tech policing itself.
The Trump administration has pushed dozens of AI companies to agree to voluntary safety testing, a plan that relies on Big Tech policing itself.
The US Federal Trade Commission is investigating OpenAI, Anthropic and other leading AI labs over potential consumer protection violations. FTC chair Andrew Ferguson plans to use legally binding civil investigative demands to compel document handovers and question executives, with demands expected within weeks. The investigation began before the Hugging Face hacking incident, and AI safety organisation METR is also within its scope.
Google's Sundar Pichai, Anthropic's Dario Amodei, Meta's Mark Zuckerberg and OpenAI's Greg Brockman.
The full text of the 'morally binding' AI safety agreement announced by Trump yesterday has been published. Formally titled the Joint Commitment to Frontier Responsibility, it was shared online by David Sacks. Signatories include Google's Sundar Pichai.
Anthropic published an evaluation saying Zhipu's open-weight GLM-5.3 model approaches its Claude Mythos Preview in exploit development.
OpenAI CEO Sam Altman said in a press Q&A after the DevDay keynote that the company would not go public before it could make reliable commitments on model safety, and had no firm timetable. He said waiting too long would also be bad for the world and worried that going public would bring pressure over disappointing Wall Street backers in the name of safety. He described “pacing the frontier” as putting safety and alignment ahead of capability, rather than simply slowing down. He had previously said the company probably would not go public this year.
Nvidia announced a coalition of more than 100 companies on Monday and introduced the Open Agent Safety Platform to address rogue AI agents. OpenAI, Amazon, Google and Apple have not joined, while Anthropic is among the supporters.
OpenAI announced at DevDay that ChatGPT has over 1.2 billion weekly active users, ChatGPT Work and Codex have more than 35 million combined weekly active users, and 2.5 million businesses use OpenAI products.
Anthropic’s IPO prospectus warns in its pitch that uncontrolled AI could end humanity within a generation. CEO Amodei told the UN Security Council last week that AI is the world’s most important global security issue today.
A preview of Anthropic's IPO prospectus disclosed plans to spend $518 billion on cloud, computing and infrastructure over the coming years. Revenue rose twelvefold to nearly $4.6 billion in 2025, while net losses reached $42 billion.
Anthropic's S-1 disclosed that 2025 revenue rose twelvefold to nearly $4.6 billion, while operating losses widened from $2.98 billion to $8.06 billion. Compute and infrastructure alone accounted for $7.33 billion in spending.
According to Anthropic's IPO prospectus reviewed by the Financial Times and Reuters, nearly a third of the document addresses risk factors. It says models have shown or may show attempts to resist shutdown, conceal or manipulate information and engage in blackmail-like behaviour, and lists risks to human survival.
AMD is acquiring World Labs for $8.2 billion. Since its founding in 2024, World Labs has built model-training teams for images, video and spatial reconstruction, and advanced its robotics simulation capabilities by acquiring SceniX. Its recently released Atlas is an omni model architecture that predicts new viewpoints from 2D images. Combining generative models with multiview geometry, it solves the longstanding sparse reconstruction problem in computer vision, with direct applications in robotics, design, engineering and science.
China's Ministry of Industry and Information Technology asked Alibaba and ByteDance to submit purchase plans for Nvidia RTX Pro 5500 chips. ByteDance plans to order one million chips, and Nvidia is preparing to ship the first orders this December.
On the Latent Space podcast, OpenRouter co-founder and CEO Alex Atallah and AMP's Anjney Midha reviewed its growth from the Llama, Alpaca and Mistral era into a neutral routing layer serving over 10 million developers and averaging over 10 trillion tokens daily, culminating in its acquisition by Stripe.
The US Court of Appeals for the DC Circuit ruled 2–1 that the Department of Defense may blacklist Anthropic for refusing to give the military access to certain Claude features, even without malicious intent. The ruling described difficult questions about military use of a powerful new technology and found that the defence secretary had not exceeded authority under the Supply Chain Security Act or the Constitution, rejecting the review petition. The court had already rejected Anthropic’s emergency stay request in April.
Latent Space plans to launch AINews v3 next week, merging the newsletter, manually written by swyx for three years with over 200,000 subscribers, with Latent Space Discord, while exploring a move to Beehiiv and a new homepage.
Anthropic introduced Projects in Claude Code, allowing one conversation to spawn parallel cloud sessions, pass context between threads and keep running after users leave. Google updated the Antigravity-based harness for Gemini managed agents with Credentials API and Files API, claiming up to 30% lower costs and 22% higher cache-hit rates.
Latent Space's AINews rounds up 15–16 September developments, highlighting two reality checks for coding-agent optimism. Steve Yegge shut down Gas Town and admitted that despite spending thousands of dollars monthly on coding-agent subscriptions, Gas Town was the only project he had built with them.
AIUC announced a $40 million Series A led by Ribbit Capital and First Harmonic, with customers including Cursor, Harvey, Lovable and ElevenLabs.
Good Start Labs turned railway board game 1830 into a training environment and tested a 30B model. Both single-turn Q&A and multi-turn terminal-agent training improved game performance, but only the terminal-agent design improved Finance-Agent benchmark results.
AI Evaluator Forum released AEF-1 as a baseline proposal for independent third-party AI evaluation, covering access, conflicts of interest, funding relationships, recusal and transparency, with xAI, OpenAI and Anthropic signing.