Skip to content

AI agents

Models that plan, call tools and finish multi-step tasks on their own: from Claude Code and Manus to agent frameworks and benchmarks.

107selectedRelated topicsMCP and tool useAI codingReasoning

Latest selected

81–100 of 107

May 28

Thursday

May 22

Friday

May 20

Wednesday

May 18

Monday

May 17

Sunday

May 16

Saturday

May 12

Tuesday

May 6

Wednesday

Apr 30

Thursday
  1. Enabling a new model for healthcare with AI co-clinician

    Google DeepMind announced AI co-clinician research to explore AI agents assisting patient care under clinical supervision. In blinded assessments of 98 real primary-care queries, 97 responses had no critical errors, and doctors preferred them to existing evidence-synthesis tools. Across 140 consultation skills, AI matched or exceeded primary-care doctors on 68, but expert doctors were better overall at recognising red flags and guiding key physical examinations.

Apr 27

Monday

Apr 22

Wednesday

Mar 29

Sunday
  1. Reimagining the mouse pointer for the AI era

    Google DeepMind unveiled an experimental Gemini-powered AI pointer that understands not only what it points at but what it means to the user. The team proposed four interaction principles: avoiding workflow interruption, capturing nearby visual and semantic context, supporting natural shorthand such as “this” and “that”, and turning pixels into actionable entities such as places, dates and objects.

Mar 25

Wednesday

Mar 18

Wednesday
  1. Introducing Forge

    Mistral AI launched Forge, a system for enterprises to build frontier-class AI models on proprietary knowledge, supporting pre-training, post-training and reinforcement learning. It handles dense and MoE architectures and, when needed, multimodal inputs, with training and governance on companies' own infrastructure.

Mar 11

Wednesday

Jan 28

Wednesday

Dec 9

Tuesday

Oct 24

Friday