Google Gemini 2.5 Pro Leads in Coding and Reasoning Benchmarks
- Gemini 2.5 Pro tops the WebDev Arena leaderboard, surpassing Claude in coding tasks.
- The AI model supports a 1 million token context window, expandable to 2 million.
- Achieved highest scores on reasoning benchmarks, including MENSA IQ tests and Humanity’s Last Exam.
- Scored 86.7% on AIME math test and 84.0% on GPQA science assessment.
- Handles up to 30,000 lines of code with advanced multimodal abilities.
Google’s Gemini 2.5 Pro has emerged as a leader in coding and reasoning capabilities, outperforming competitors like Claude with its extensive token support and high benchmark scores.
Source (2.6)https://decrypt.co/318416/googles-gemini-2-5-pro-tops-coding-charts-mensa-tests-ai-iq-battle?rand=52368