Skills in Claude Code 2.1: Loadable Capability Packs Explained
A practical engineering deep dive into Claude Code 2.1 skills, covering architecture, tradeoffs, and what production teams need to know about developer tooling.
Agentic AI, LLM engineering, and the models behind modern automation — multi-agent systems, LLM evaluation and comparisons, RAG, fine-tuning, AI infrastructure, security, and production AI engineering.
From the blog
A practical engineering deep dive into Claude Code 2.1 skills, covering architecture, tradeoffs, and what production teams need to know about developer tooling.
Infrastructure-level look at MCP signed registry, including supply chain security AI, deployment topology, region availability, and cost considerations.
Azure AI Foundry's connected agents and shared memory make multi-agent orchestration on Azure finally workable. The patterns and gotchas to know before going live.
The CrewAI tools marketplace crossed 500 tools in April 2026. The most-installed tools and the ones quietly powering production crews across industries today.
Embedding model picks in 2026 with real numbers. text-embedding-3, Voyage, Cohere v4, and BGE-M3 across English, multilingual, and code retrieval workloads measured.
A semantic cache layer in front of memory recall cuts tail latency and saves embedding API spend significantly. Where it works, where it lies, and how to debug it.
How Singapore-based companies are using ChatGPT Operator 2.0 for cross-border APAC workflows — pricing, latency, and regulatory considerations in 2026.
Trace every turn, every tool call, every LLM round-trip with OpenTelemetry shipped to Langfuse. Find latency outliers, debug hallucinations, and watch p95 stay under 800ms.
Cloudflare Calls is anycast WebRTC across 330+ cities. Fastly is leaner and faster on the compute edge. Here is how each fits a 2026 global voice-AI deployment.