3 articles
Demystifying evals for AI agents

Anthropic shares principles for building effective AI agents, focusing on context engineering — the practice of providing agents with the right information, tools, and instructions to produce reliable results. Covers prompt design, tool definitions, and managing long-running agent sessions.
Anthropic describes the engineering behind their multi-agent Research feature, where a planning agent decomposes complex queries and spawns parallel search agents. The post covers architectural principles, prompting strategies, and evaluation methods for reliable multi-agent systems.


Demystifying evals for AI agents
Anthropic shares principles for building effective AI agents, focusing on context engineering — the practice of providing agents with the right information, tools, and instructions to produce reliable results. Covers prompt design, tool definitions, and managing long-running agent sessions.

Anthropic describes the engineering behind their multi-agent Research feature, where a planning agent decomposes complex queries and spawns parallel search agents. The post covers architectural principles, prompting strategies, and evaluation methods for reliable multi-agent systems.