Skip to content

All AI news

31 today

Sep 22

Tuesday
  1. Jev introduces a new shape of LLM - System One, aka Decision Models

    TypeSafe AI released Jev last week, calling it the first example of a new System One model category, though the author prefers decision model. It takes text and returns floating-point values and confidence for categories, yes/no questions and scores. Only input is billed, at $0.042 per million tokens for the first model. The author sees uses in spam detection, tagging and ranking, has tried search reranking, and notes weaknesses with numbers, dates and adversarial content.

  2. From Enablement to Execution, Egypt’s AI Ecosystem Reaches Production Scale

    NVIDIA held an AI ecosystem event at the Grand Egyptian Museum, with Egyptian Deep Learning Institute participation growing over tenfold in a year. Hassan Allam received a data centre licence from Egypt's National Telecom Regulatory Authority and is developing a new centre with A15 with expected investment of $400 million. NVIDIA says four African AI factories have been announced or launched, with another 656 MW under construction.

Sep 21

Monday
  1. AI Security Is an Engineering Problem — How to Solve It at Every Layer of the Agent Stack

    NVIDIA argues AI security should be treated as an engineering problem, with explicit requirements, executable controls, named owners and evidence of effective protection. Its open-source NVIDIA OpenShell enforces policies outside agent reasoning and provides sandboxed execution. Cisco DefenseClaw adds governance, while JFrog integrates OpenShell to scan and validate agent skills.

  2. Import AI 473: The US's superintelligence strategy; human brain in a mouse skull; and machine hermeneutics

    A long RAND report recommends a US 'freedom of action' strategy amid uncertainty on the path to superintelligence, preserving options through AI safety investment, safety architecture, national security reform and public resilience. It outlines seven prototype strategies in coexistence, denial and acceleration categories, and five uncertainties: proximity of danger, coexistence feasibility, constraint feasibility, decisive strategic advantage and suppression feasibility.

  3. How we made the first comprehensive map of deaths along the US border’s “virtual wall”

    MIT Technology Review and Times of San Diego spent 15 months creating the first comprehensive map and analysis of deaths near US border surveillance towers, examining migrant deaths since 2015. They requested records from 17 Texas county sheriff's offices and obtained over 4,000 pages from 14 counties, used Anthropic's Claude API to extract coordinates where remains were found, then manually checked samples.

  4. Quoting voxium

    An engineer who joined a large company two weeks earlier says specifications, code, tests, PRDs, tickets and their handling, and reports are all generated by Claude Code. Nobody likes the approach, but they are told to deliver as much as possible. They repeatedly heard leadership say shipping code was not the bottleneck, while engineers from L1 to L7 worked 12–13 hours daily just pressing Enter, with nobody reading anything.

  5. MCP was always a bad idea?

    Simon Willison rejected the view that MCP is now a bad idea, arguing it retains irreplaceable value beyond terminal agents such as Claude Code and Codex with unrestricted internet access. MCP makes it easier to limit external service access, keep agents from directly handling API keys, provide connection and authentication interfaces and maintain strong audit logs. Dismissing it because fully capable coding agents do not need it overlooks other use cases.

  6. llm-keys-ui 0.1

    Simon Willison released llm-keys-ui 0.1 to configure API keys on remote machines. Running uvx --with llm-keys-ui llm keys-ui --all opens an interface to save keys over the local network or a Tailscale device IP, avoiding pasting secrets into ChatGPT or agent sessions.

Sep 20

Sunday

Sep 19

Saturday

Sep 18

Friday
  1. Should you read the code, is RAG dead, and did Skills kill MCP?

    The latest GitHub Podcast examines five AI development memes. AI-generated code still needs reading and accountability, with review proportional to risk. Skills package team experience, while MCP standardises connections to tools and data; they can combine. RAG is not dead: it provides relevant information beyond training data and can coexist with agents, Skills and MCP in one workflow.

  2. [AINews] not much happened today

    Anthropic introduced Projects in Claude Code, allowing one conversation to spawn parallel cloud sessions, pass context between threads and keep running after users leave. Google updated the Antigravity-based harness for Gemini managed agents with Credentials API and Files API, claiming up to 30% lower costs and 22% higher cache-hit rates.

Sep 17

Thursday

Sep 16

Wednesday