Vigyata.AI
Is this your channel?

Anthropic Just Killed Tool Calling

57.8K views· 1,621 likes· 13:40· Feb 18, 2026

🛍️ Products Mentioned (16)

Anthropic's latest Sonnet 4.6 release quietly introduced programmatic tool calling; a feature that lets AI agents write code instead of JSON to invoke tools, slashing token usage by up to 98% while improving accuracy. I break down why this "code mode" approach outperforms traditional tool calling, how companies like Cloudflare and Anthropic are already implementing it, and why this could become the new industry standard just like MCP did. If you're building AI agents, this might be the most important under-the-radar upgrade of the year. https://platform.claude.com/docs/en/agents-and-tools/tool-use/programmatic-tool-calling https://www.anthropic.com/engineering/effective-context-engineering-for-ai-agents https://claude.com/blog/improved-web-search-with-dynamic-filtering https://www.anthropic.com/news/model-context-protocol https://platform.claude.com/docs/en/agents-and-tools/agent-skills/overview https://claude.com/blog/equipping-agents-for-the-real-world-with-agent-skills https://www.anthropic.com/engineering/advanced-tool-use My Dictation App: www.whryte.com Website: https://engineerprompt.ai/ RAG Beyond Basics Course: https://prompt-s-site.thinkific.com/courses/rag Signup for Newsletter, localgpt: https://tally.so/r/3y9bb0 Let's Connect: 🦾 Discord: https://discord.com/invite/t4eYQRUcXB ☕ Buy me a Coffee: https://ko-fi.com/promptengineering |🔴 Patreon: https://www.patreon.com/PromptEngineering 💼Consulting: https://calendly.com/engineerprompt/consulting-call 📧 Business Contact: engineerprompt@gmail.com Become Member: http://tinyurl.com/y5h28s6h 💻 Pre-configured localGPT VM: https://bit.ly/localGPT (use Code: PromptEngineering for 50% off). Signup for Newsletter, localgpt: https://tally.so/r/3y9bb0

About This Video

With Sonnet 4.6, Anthropic quietly shipped something that I think is way more important than most people realize: programmatic tool calling. The core idea is simple—stop forcing agents to emit verbose JSON tool calls, and instead let them do what they’re actually trained to do: write code. In a sandbox, the agent can orchestrate tool invocations, control ordering, and keep all the intermediate tool I/O out of the context window. You only feed the final result back, which means drastically fewer tokens and less context pollution. In this video I break down why this matters in the bigger “context engineering” problem—especially now that MCP-style setups load tons of tool definitions and tool outputs into the prompt. I walk through the timeline of “code mode” adoption (Cloudflare, then Anthropic) and why I expect this to become a standard the same way MCP and agent skills did. Finally, I cover Anthropic’s improved web search + dynamic filtering, which uses programmatic tool calling to post-process search results before they hit the context window. On benchmarks like BrowserComp and DeepSearch QA, they report meaningful accuracy gains with fewer input tokens—though I also explain why token cost doesn’t always go down if the model writes a lot of filtering code (Opus is the obvious example).

Frequently Asked Questions

🎬 More from Prompt Engineering