
News12 days ago
Copilot HydraFusion Routes Your Code Through Multiple AI Models and Two of Three Benchmarks Got Worse
GitHub's new Copilot feature picks between AI models per task and claims a 67% cost reduction. We read the benchmark tables: it beat Opus 5 on TerminalBench, lost 1.5 points on DeepSWE, and the internal benchmark was a wash.
Sep 12, 20265 min