BrowseComp

    🏆 Leaderboard

    As of August 20, 2026, Kimi K3 is #1 for BrowseComp at 91.2%. Ranked by the BrowseComp score 59 models in this index have a published BrowseComp score. Methodology: BrowseComp (OpenAI) (https://openai.com/index/browsecomp/). BrowseComp leaderboard with live API prices. BrowseComp — evaluates web browsing and information retrieval across complex multi-step research tasks. Official methodology: BrowseComp (OpenAI) (https://openai.com/index/browsecomp/).

    Updated August 20, 2026282 models33 providers
    Kimi K3OSS
    Moonshot AI · Open Source
    91.2%$18.00
    Claude Opus 5
    Anthropic · Proprietary
    90.8%$30.00
    GPT-5.6 Sol
    OpenAI · Proprietary
    90.4%$35.00
    GPT-5.5 Pro
    OpenAI · Proprietary
    90.1%$540.00
    Claude Mythos 5
    Anthropic · Proprietary
    88%$60.00
    GPT-5.6 Terra
    OpenAI · Proprietary
    87.5%$14.00
    Claude Mythos Preview
    Anthropic · Proprietary
    86.9%$60.00
    Kimi K2.6OSS
    Moonshot AI · Open Source
    86.3%$4.93
    Gemini 3.1 Pro
    Google · Proprietary
    85.9%$17.50
    Seed 2.1 Turbo
    ByteDance · Proprietary
    84.9%$3.00
    Claude Sonnet 5
    Anthropic · Proprietary
    84.7%$12.00
    GPT-5.5
    OpenAI · Proprietary
    84.4%$35.00
    Claude Opus 4.8
    Anthropic · Proprietary
    84.3%$30.00
    Hy3OSS
    Tencent · Open Source
    84.2%$0.66
    Claude Opus 4.6
    Anthropic · Proprietary
    84%$30.00
    MiniMax M3OSS
    MiniMax · Open Source
    83.5%$1.50
    DeepSeek-V4-Pro-MaxOSS
    DeepSeek · Open Source
    83.4%$5.22
    GPT-5.6 Luna
    OpenAI · Proprietary
    83.3%$1.40
    GPT-5.4
    OpenAI · Proprietary
    82.7%$17.50
    GLM-5.1OSS
    Z AI · Open Source
    79.3%$5.80
    Claude Opus 4.7
    Anthropic · Proprietary
    79.3%$30.00
    GPT-5.2 Pro
    OpenAI · Proprietary
    77.9%$189.00
    Inkling-SmallOSS
    Thinking Machines · Open Source · via Thinking Machines Lab
    77.4%$1.50
    Seed 2.0 Pro
    ByteDance · Proprietary
    77.3%$3.50
    MiniMax M2.5OSS
    MiniMax · Open Source
    76.3%$1.50
    GLM-5OSS
    Z AI · Open Source
    75.9%$4.20
    Kimi K2.5OSS
    Moonshot AI · Open Source
    74.9%$3.68
    Claude Sonnet 4.6
    Anthropic · Proprietary
    74.7%$18.00
    DeepSeek-V4-Flash-MaxOSS
    DeepSeek · Open Source
    73.2%$0.42
    Step-3.5-FlashOSS
    StepFun · Open Source
    69%$0.50
    Qwen3.5-397B-A17BOSS
    Qwen · Open Source
    69%$4.20
    GPT-5.2
    OpenAI · Proprietary
    65.8%$15.75
    Qwen3.5-122B-A10BOSS
    Qwen · Open Source
    63.8%$3.60
    MiniMax M2.1OSS
    MiniMax · Open Source
    62%$1.50
    Qwen3.5-35B-A3BOSS
    Qwen · Open Source
    61%$2.25
    Qwen3.5-27BOSS
    Qwen · Open Source
    61%$2.70
    Kimi K2-Thinking-0905OSS
    Moonshot AI · Open Source
    60.2%$2.47
    MiMo-V2-FlashOSS
    Xiaomi · Open Source
    58.3%$0.40
    LongCat-Flash-Thinking-2601OSS
    Meituan · Open Source
    56.6%$1.50
    GPT-5
    OpenAI · Proprietary
    54.9%$11.25
    DeepSeek-V4-Flash-0423OSS
    DeepSeek · Open Source
    53.5%$0.30
    GLM-4.7OSS
    Z AI · Open Source
    52%$2.80
    o4-mini
    OpenAI · Proprietary
    51.5%$5.50
    DeepSeek-V3.2OSS
    DeepSeek · Open Source · via OpenRouter
    51.4%$0.57
    o3
    OpenAI · Proprietary
    49.7%$10.00
    Solar Pro 4
    Upstage · Proprietary
    49.2%$1.50
    Mistral Medium 3.5OSS
    Mistral · Open Source · via Mistral AI
    48.6%$9.00
    GLM-4.6OSS
    Z AI · Open Source
    45.1%$2.80
    Grok 4 Fast
    xAI · Proprietary
    44.9%$0.70
    Nemotron 3 Ultra (550B A55B)OSS
    NVIDIA · Open Source · via OpenRouter
    44.4%$2.70
    Showing 150 of 282 models

    Next step

    You found the model. Now ship the product.

    Auth, billing, and the API layer are already decided. Fork 8 finished AI apps, or have us build the first version with you.

    The index

    All Large Language Models

    282 models across 33 providers. Search or jump to a lab — every model page stays linked here.

    Baidu

    1 models

    Inception

    1 models

    inclusionAI

    1 models

    LG AI Research

    1 models

    Nous Research

    1 models

    Sakana AI

    1 models

    StepFun

    1 models

    Tencent

    1 models

    Thinking Machines

    1 models

    Unisound

    1 models

    Upstage

    1 models

    Which model leads BrowseComp right now?

    As of August 20, 2026, Kimi K3 by Moonshot AI is #1 for BrowseComp at 91.2%. Ranked by the BrowseComp score This board also tracks BrowseComp. Next on the same board: Claude Opus 5 and GPT-5.6 Sol. This browsecomp leaderboard ranks models by BrowseComp. Scores come from public evals. Prices are the live API rates in the table above.

    Sources: BrowseComp (OpenAI) (https://openai.com/index/browsecomp/); OpenAI API pricing (https://developers.openai.com/api/docs/pricing); Anthropic Claude API pricing (https://platform.claude.com/docs/en/about-claude/pricing); Gemini API pricing (https://ai.google.dev/gemini-api/docs/pricing)

    Top 8 for browsecomp. Ranked by the BrowseComp score Input and output are dollars per million tokens.
    RankModelBrowseCompInput /MOutput /M
    1Kimi K391.2%$3.00$15.00
    2Claude Opus 590.8%$5.00$25.00
    3GPT-5.6 Sol90.4%$5.00$30.00
    4GPT-5.5 Pro90.1%$60.00$480.00
    5Claude Mythos 588%$10.00$50.00
    6GPT-5.6 Terra87.5%$2.00$12.00
    7Claude Mythos Preview86.9%$10.00$50.00
    8Kimi K2.686.3%$0.96$3.97

    BrowseComp FAQ

    Who ranks #1 on the BrowseComp leaderboard?

    As of August 20, 2026, Kimi K3 by Moonshot AI ranks #1 on BrowseComp at 91.2%. API pricing is $3.00/M input and $15.00/M output.

    What are the top models on BrowseComp?

    The current BrowseComp ranking as of August 20, 2026 is 1. Kimi K3 at 91.2%; 2. Claude Opus 5 at 90.8%; 3. GPT-5.6 Sol at 90.4%.

    Which browsecomp model is the cheapest?

    Nemotron 3.5 Lightning (30B A3B) is the cheapest scored model on this browsecomp leaderboard at $0.05/M input and $0.20/M output ($0.25 blended). Kimi K3 still leads BrowseComp at 91.2%.

    Should I always pick the #1 BrowseComp model?

    Not automatically. Kimi K3 leads BrowseComp, but a cheaper scored model can be the better production choice if the quality gap is small. Use the table to weigh BrowseComp against input/output price, context window, and related evals.

    How often is the BrowseComp leaderboard updated?

    Scores and API prices on this page are refreshed from published evals and provider rates. The snapshot is labeled August 20, 2026. Treat it as a current index, not a one-off blog post.

    What is BrowseComp?

    BrowseComp — evaluates web browsing and information retrieval across complex multi-step research tasks. This page ranks models that have published a BrowseComp score, with live API token prices on the same row. Official methodology: BrowseComp (OpenAI) (https://openai.com/index/browsecomp/).

    Where is the BrowseComp leaderboard?

    This page is the BrowseComp leaderboard. Models are sorted by BrowseComp, with input and output token prices on the same row so you can weigh score against cost. Official boards often omit price; that comparison is the point of this index.

    How is this BrowseComp ranking different from the official board?

    The official BrowseComp (OpenAI) page owns the methodology. This page keeps the published BrowseComp score next to live API $/M so you can pick a production SKU, not only a trophy number. Source: https://openai.com/index/browsecomp/