InsForge has updated its MCPMark benchmark results, now using Anthropic's Claude Sonnet 4.6. The benchmarks, which compare InsForge's MCP layer against Supabase's MCP layer on 21 real-world database tasks, show InsForge maintaining its lead in accuracy and token efficiency. With Claude Sonnet 4.6, InsForge achieved 28% higher Pass⁴ accuracy and used 2.4 times fewer tokens than Supabase MCP, widening the efficiency gap from previous tests. AI
IMPACT Highlights the growing importance of structured backend context for LLM agents, as more capable models amplify the cost of not providing it.
RANK_REASON Updated benchmark results comparing two MCP layers using a new model version. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →