Sub-Agents in Claude Code 2.1: When and How to Spawn Them
A practical engineering deep dive into Claude Code 2.1 sub-agents, covering architecture, tradeoffs, and what production teams need to know about agent orchestration.
Agentic AI, LLM engineering, and the models behind modern automation — multi-agent systems, LLM evaluation and comparisons, RAG, fine-tuning, AI infrastructure, security, and production AI engineering.
From the blog
A practical engineering deep dive into Claude Code 2.1 sub-agents, covering architecture, tradeoffs, and what production teams need to know about agent orchestration.
Infrastructure-level look at MCP 1.0 spec, including Model Context Protocol, deployment topology, region availability, and cost considerations.
A practical engineering deep dive into Claude Opus 4.7 tool use, covering architecture, tradeoffs, and what production teams need to know about function calling.
A practical engineering deep dive into Constitutional AI 3.0 production, covering architecture, tradeoffs, and what production teams need to know about AI guardrails.
Google's A2A protocol lets agents discover and call other agents across vendor boundaries. The spec, the agent-card flow, and what it competes with in the agent stack.
Bedrock AgentCore adds memory primitives, tool gateways, and runtime services to the Bedrock agent stack. The architecture explained for real production builders.
Three reranker families dominate 2026 with different pricing and accuracy profiles. Quality on real RAG datasets, latency, and price-per-million ranked pairs measured.
Stuffing the whole memory buffer into the prompt is a token-bill grenade waiting to explode. The memory-aware prompt patterns that cut cost by 4x in production.
What does 10,000 minutes of voice AI per month really cost on Vapi vs CallSphere? Here is the side-by-side, line by line.