4 ms·
If context is the bottleneck, MCP is dead. MCP can use 10k tokens. Everything good happens in the first 100k tokens. It's more context efficient to code a cu
by heyrhett 1y ago
If context is the bottleneck, MCP is dead.
MCP can use 10k tokens. Everything good happens in the first 100k tokens.
It's more context efficient to code a custom binary and prompt the LLM how to use the binary when needed.