Cognition SWE-2 reportedly scores 50 percent on Frontier Code at 70 percent lower cost. What is verified, how coding benchmarks mislead, and how to evaluate it.
Grok 4.7 from xAI explained: what is disclosed on architecture, the CursorBench 46 percent result, API pricing near $2 per million tokens, and how it compares.