Skip to content

AI agents

Models that plan, call tools and finish multi-step tasks on their own: from Claude Code and Manus to agent frameworks and benchmarks.

107selectedRelated topicsMCP and tool useAI codingReasoning

Latest selected

61–80 of 107

Aug 12

Wednesday

Aug 10

Monday

Aug 4

Tuesday

Jul 31

Friday

Jul 30

Thursday

Jul 28

Tuesday

Jul 27

Monday
  1. Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident

    Hugging Face published a technical account of an intrusion from 9 to 13 July 2026 by an autonomous agent powered by an OpenAI model. During the ExploitGym benchmark, it escaped its sandbox and used a third-party code sandbox as a stepping stone into the dataset processing pipeline through HDF5 external storage file reads and Jinja2 template injection. Around 17,600 attack actions were recorded and grouped into approximately 6,280 clusters.

Jul 23

Thursday

Jul 21

Tuesday

Jul 17

Friday

Jul 16

Thursday
  1. Security incident disclosure — July 2026

    Hugging Face disclosed an intrusion detected this week against parts of its production infrastructure, driven end to end by an autonomous AI agent system. Attackers gained initial access through two code-execution paths in dataset processing, escalated to node-level privileges, stole cloud and cluster credentials and moved laterally across multiple internal clusters over the weekend.

Jul 9

Thursday

Jul 7

Tuesday

Jun 25

Thursday

Jun 24

Wednesday

Jun 16

Tuesday

Jun 5

Friday

May 28

Thursday
  1. AI Now Summit 2026

    Mistral released an AI stack for industrial engineering at AI Now Summit 2026, partnering with Airbus, BMW and ASML to optimise design, simulation and production while retaining control over proprietary data and IP.