Scenario guide
Best AI models for Research / Analyst
A research assistant that reads long PDFs and answers nuanced questions. Weighted toward reasoning, knowledge, long-context, and a saturation-resistant frontier capability score so the ranking stays meaningful as MMLU-style evals saturate.
Rankings use the same scenario weights and cost blending as the interactive leaderboard on AI Model Analyzer. Data is min-max normalised per benchmark; missing scores are skipped without penalty.
- 1Gemini 3 ProGoogleScore 90.9Q 96.3In $1.25/M
- 2GPT-5.5OpenAIScore 87.9Q 93.4In $1.50/M
- 3Gemini 3 FlashGoogleScore 79.8Q 81.2In $0.30/M
- 4Gemini 2.5 ProGoogleScore 78.5Q 82.5In $1.25/M
- 5GPT-5.4OpenAIScore 75.2Q 79.2In $1.50/M
- 6DeepSeek V3DeepSeekScore 73.6Q 73.6In $0.27/M
- 7Gemini 2.0 FlashGoogleScore 72.3Q 70.2In $0.10/M
- 8Claude Opus 4.7AnthropicScore 72.2Q 80.2In $15.00/M
- 9Qwen3 235BAlibaba (Qwen)Score 69.4Q 68.3In $0.20/M
- 10Claude Opus 4.6AnthropicScore 69.0Q 76.6In $15.00/M
- 11Grok 3xAIScore 68.4Q 72.8In $3.00/M
- 12o3-miniOpenAIScore 68.2Q 70.5In $1.10/M
- 13o1OpenAIScore 67.1Q 74.5In $15.00/M
- 14DeepSeek R1DeepSeekScore 67.1Q 67.9In $0.55/M
- 15o3OpenAIScore 66.2Q 72.6In $10.00/M