Agentic RAG: Adaptive Retrieval Patterns That Scale
How agentic RAG beats naive pipelines: agent-controlled retrieval, query routing, verify-then-retrieve loops, and guardrails that prevent infinite loops.
Lab // Technical notes
Technical deep dives into AI agent engineering — architecture patterns, protocol internals, and the implementation details behind production-grade systems.
How agentic RAG beats naive pipelines: agent-controlled retrieval, query routing, verify-then-retrieve loops, and guardrails that prevent infinite loops.
Tool poisoning hides malicious instructions in MCP tool descriptions models trust but users never see. A working exploit, the rug pull, and defenses in code.
The 2026-07-28 MCP RC drops the initialize handshake and Mcp-Session-Id, moves client info to _meta, and adds routing headers. What to refactor.
Code mode lets an agent orchestrate many tools in a sandbox and return only results to the model. ~80% lower inference cost and how it works in production.
Claude Fable 5 ships effort and task_budget — two API primitives that govern cost, latency, and reliability in agent loops. A production wiring guide.
MCP's 2026-07-28 release candidate locked on May 21. Tasks graduated to a first-class extension. Here is what every server author has 10 weeks to refactor.
Microsoft's CVE-2026-26030 and CVE-2026-25592 turn one injected prompt into calc.exe. The tool registry is now the attack surface — here's the chain.
A2A v1.0 ships in April 2026 under the AAIF. Agent Cards, a five-state task lifecycle, and the orchestrator code that ties them together.
PASTE-style speculative tool execution cuts agent task time 48.5%. A working ~300 LOC harness over the Anthropic SDK with rollback and benchmarks.
CaMeL splits the LLM in two and wraps every value in a capability. 77% of AgentDojo tasks complete with provable prompt-injection defense.