Skip to content

#MCP and tool use

2 today

Oct 1

ThursdayToday2 items

Sep 30

Wednesday
  1. Build a multi-agent music production pipeline on Amazon Bedrock AgentCore Runtime Instances

    The official AWS blog demonstrates deploying a three-agent music production pipeline on Amazon Bedrock AgentCore Runtime Instances. The composition agent uses Claude Sonnet 4.6 to generate a music brief and runs ACE-Step on the instance's NVIDIA L4 to render audio. The delivery agent reads .wav files from the shared volume for measurement and DSP processing, while the compliance agent independently remeasures them and checks harmonic similarity against a music catalogue.

  2. AI-powered app maker Wabi pivots to a messaging experience

    AI start-up Wabi, which previously let users build apps with prompts, has pivoted to messaging with Wabi 2.0, positioned as a personal agent that does things for users and instantly builds the interfaces they need. Users can generate apps such as calorie trackers and weightlifting logs within conversations. Access is currently offered only through invitation codes distributed on X.

  3. Prompt engineering fundamentals for Amazon Quick

    Prompt engineering determines the quality of Amazon Quick's AI responses to natural language requests. The first instalment of an official two-part series covers principles shared across components and reusable frameworks. It introduces CRISPE, covering context and constraints, roles and responsibilities, intent and inputs, steps and scope, and emphasises specificity, business context and examples over abstract descriptions.

Sep 29

Tuesday
  1. Getting the Source Right, Not Just the Fact: Source-Aware Verification for MCP Agents

    Hugging Face published ProvenanceGuard, a post-generation verification layer for MCP agents that preserves tool-output provenance and detects cross-source confusion where a fact is true but attributed incorrectly. Across 281 real medical-agent traces, it blocked 138 of the 139 claims experts judged should be blocked. Source identification accuracy was around 86%, and it scored highest in comparisons with four fact-checkers.

  2. Manus 2.0 lets users edit videos, host multiplayer games, and run agents remotely from their phone

    Manus released version 2.0, expanding its AI agents into a platform with video editing, multiplayer game hosting, and personal agents that run through a phone number and can be controlled remotely from a phone. In the test configuration, the new Cascade agent framework used 23.2% fewer tokens than the previous system and reduced operating costs by 32%.

Sep 28

Monday
  1. Bluesky reply bot checker

    Simon Willison used Opus 5.5 to build a tool analysing any Bluesky account for automated reply-bot signals. These include replying within seconds, never posting original content or image links and replying only to high-follower accounts, and using question marks in replies. Bluesky's free API makes such investigations more feasible than on Twitter.

Sep 26

Saturday

Sep 25

Friday

Sep 24

Thursday

Sep 22

Tuesday

Sep 21

Monday
  1. MCP was always a bad idea?

    Simon Willison rejected the view that MCP is now a bad idea, arguing it retains irreplaceable value beyond terminal agents such as Claude Code and Codex with unrestricted internet access. MCP makes it easier to limit external service access, keep agents from directly handling API keys, provide connection and authentication interfaces and maintain strong audit logs. Dismissing it because fully capable coding agents do not need it overlooks other use cases.

  2. llm-keys-ui 0.1

    Simon Willison released llm-keys-ui 0.1 to configure API keys on remote machines. Running uvx --with llm-keys-ui llm keys-ui --all opens an interface to save keys over the local network or a Tailscale device IP, avoiding pasting secrets into ChatGPT or agent sessions.

Sep 19

Saturday

Sep 18

Friday
  1. Should you read the code, is RAG dead, and did Skills kill MCP?

    The latest GitHub Podcast examines five AI development memes. AI-generated code still needs reading and accountability, with review proportional to risk. Skills package team experience, while MCP standardises connections to tools and data; they can combine. RAG is not dead: it provides relevant information beyond training data and can coexist with agents, Skills and MCP in one workflow.

Sep 11

Friday

Sep 10

Thursday
  1. Rebuilding AUTOMATIC1111 with Gradio Workflow

    The Hugging Face team rebuilt most AUTOMATIC1111 functionality as the Workflow1111 canvas using Gradio Workflow. Its 11 media pipelines and 73 nodes cover text-to-image, high-resolution fixes, image-to-image, prompt matrices, VLM reverse prompting, detection-generated inpainting masks, ControlNet-style annotators, background removal, PNG Info and image-to-video.

Sep 3

Thursday

Aug 25

Tuesday

Aug 20

Thursday

Jul 9

Thursday
  1. Your Prompts and Skills need a system of record.

    Mistral Studio now provides central management of Prompts and Skills, with version histories, clear ownership and traceability. Immutable versions, rollback, tags and audit logs make AI behaviour governable and discoverable, while Observability traces production outputs to their versions. Skills can be invoked directly from Studio as MCP servers, ensuring production executes the same governed asset.

Jun 25

Thursday

Jun 24

Wednesday

May 28

Thursday

May 22

Friday

Apr 27

Monday