BrowseComp

    ๐Ÿ† Leaderboard

    As of August 20, 2026, Kimi K3 is #1 for BrowseComp at 91.2%. Ranked by the BrowseComp score BrowseComp leaderboard for web-research agents, plus API pricing. Rank LLMs on multi-step browsing and retrieval.

    Updated August 20, 2026282 models33 providers
    Kimi K3OSS
    Moonshot AI ยท Open Source
    91.2%42.7%84.2%$18.00
    Claude Opus 5
    Anthropic ยท Proprietary
    90.8%56.7%85.8%$30.00
    GPT-5.6 Sol
    OpenAI ยท Proprietary
    90.4%71.6%โ€”$35.00
    GPT-5.5 Pro
    OpenAI ยท Proprietary
    90.1%โ€”โ€”$540.00
    Claude Mythos 5
    Anthropic ยท Proprietary
    88%โ€”โ€”$60.00
    GPT-5.6 Terra
    OpenAI ยท Proprietary
    87.5%43.1%โ€”$14.00
    Claude Mythos Preview
    Anthropic ยท Proprietary
    86.9%โ€”โ€”$60.00
    Kimi K2.6OSS
    Moonshot AI ยท Open Source
    86.3%38.7%โ€”$4.93
    Gemini 3.1 Pro
    Google ยท Proprietary
    85.9%77.3%69.2%$17.50
    Seed 2.1 Turbo
    ByteDance ยท Proprietary
    84.9%โ€”80.3%$3.00
    Claude Sonnet 5
    Anthropic ยท Proprietary
    84.7%25%โ€”$12.00
    GPT-5.5
    OpenAI ยท Proprietary
    84.4%โ€”75.3%$35.00
    Claude Opus 4.8
    Anthropic ยท Proprietary
    84.3%39.5%82.2%$30.00
    Hy3OSS
    Tencent ยท Open Source
    84.2%โ€”79.1%$0.66
    Claude Opus 4.6
    Anthropic ยท Proprietary
    84%41.0%62.7%$30.00
    MiniMax M3OSS
    MiniMax ยท Open Source
    83.5%โ€”74.2%$1.50
    DeepSeek-V4-Pro-MaxOSS
    DeepSeek ยท Open Source
    83.4%57.9%73.6%$5.22
    GPT-5.6 Luna
    OpenAI ยท Proprietary
    83.3%41.7%โ€”$1.40
    GPT-5.4
    OpenAI ยท Proprietary
    82.7%44.8%67.2%$17.50
    GLM-5.1OSS
    Z AI ยท Open Source
    79.3%37.3%71.8%$5.80
    Claude Opus 4.7
    Anthropic ยท Proprietary
    79.3%50.6%77.3%$30.00
    GPT-5.2 Pro
    OpenAI ยท Proprietary
    77.9%โ€”77.8%$189.00
    Inkling-SmallOSS
    Thinking Machines ยท Open Source ยท via Thinking Machines Lab
    77.4%19.5%79.6%$1.50
    Seed 2.0 Pro
    ByteDance ยท Proprietary
    77.3%โ€”โ€”$3.50
    MiniMax M2.5OSS
    MiniMax ยท Open Source
    76.3%โ€”โ€”$1.50
    GLM-5OSS
    Z AI ยท Open Source
    75.9%โ€”67.8%$4.20
    Kimi K2.5OSS
    Moonshot AI ยท Open Source
    74.9%โ€”โ€”$3.68
    Claude Sonnet 4.6
    Anthropic ยท Proprietary
    74.7%29%61.3%$18.00
    DeepSeek-V4-Flash-MaxOSS
    DeepSeek ยท Open Source
    73.2%34.1%69%$0.42
    Step-3.5-FlashOSS
    StepFun ยท Open Source
    69%โ€”โ€”$0.50
    Qwen3.5-397B-A17BOSS
    Qwen ยท Open Source
    69%โ€”โ€”$4.20
    GPT-5.2
    OpenAI ยท Proprietary
    65.8%35.4%60.6%$15.75
    Qwen3.5-122B-A10BOSS
    Qwen ยท Open Source
    63.8%โ€”โ€”$3.60
    MiniMax M2.1OSS
    MiniMax ยท Open Source
    62%โ€”โ€”$1.50
    Qwen3.5-35B-A3BOSS
    Qwen ยท Open Source
    61%โ€”โ€”$2.25
    Qwen3.5-27BOSS
    Qwen ยท Open Source
    61%โ€”โ€”$2.70
    Kimi K2-Thinking-0905OSS
    Moonshot AI ยท Open Source
    60.2%โ€”โ€”$2.47
    MiMo-V2-FlashOSS
    Xiaomi ยท Open Source
    58.3%โ€”โ€”$0.40
    LongCat-Flash-Thinking-2601OSS
    Meituan ยท Open Source
    56.6%โ€”โ€”$1.50
    GPT-5
    OpenAI ยท Proprietary
    54.9%50.6%โ€”$11.25
    DeepSeek-V4-Flash-0423OSS
    DeepSeek ยท Open Source
    53.5%28.9%67.4%$0.30
    GLM-4.7OSS
    Z AI ยท Open Source
    52%31.5%โ€”$2.80
    o4-mini
    OpenAI ยท Proprietary
    51.5%21.9%โ€”$5.50
    DeepSeek-V3.2OSS
    DeepSeek ยท Open Source ยท via OpenRouter
    51.4%โ€”โ€”$0.57
    o3
    OpenAI ยท Proprietary
    49.7%53%โ€”$10.00
    Solar Pro 4
    Upstage ยท Proprietary
    49.2%โ€”โ€”$1.50
    Mistral Medium 3.5OSS
    Mistral ยท Open Source ยท via Mistral AI
    48.6%โ€”โ€”$9.00
    GLM-4.6OSS
    Z AI ยท Open Source
    45.1%โ€”โ€”$2.80
    Grok 4 Fast
    xAI ยท Proprietary
    44.9%95%โ€”$0.70
    Nemotron 3 Ultra (550B A55B)OSS
    NVIDIA ยท Open Source ยท via OpenRouter
    44.4%โ€”โ€”$2.70
    Showing 1โ€“50 of 282 models

    Next step

    You found the model. Now ship the product.

    Auth, billing, and the API layer are already decided. Fork 8 finished AI apps, or have us build the first version with you.

    The index

    All Large Language Models

    282 models across 33 providers. Search or jump to a lab โ€” every model page stays linked here.

    Baidu

    1 models

    Inception

    1 models

    inclusionAI

    1 models

    LG AI Research

    1 models

    Nous Research

    1 models

    Sakana AI

    1 models

    StepFun

    1 models

    Tencent

    1 models

    Thinking Machines

    1 models

    Unisound

    1 models

    Upstage

    1 models

    Which model leads BrowseComp right now?

    As of August 20, 2026, Kimi K3 by Moonshot AI is #1 for BrowseComp at 91.2%. Ranked by the BrowseComp score This board also tracks BrowseComp, SimpleQA, MCP Atlas. Next on the same board: Claude Opus 5 and GPT-5.6 Sol. Related leaders: DeepSeek-V3.2-Exp on SimpleQA at 97.1%; Muse Spark 1.1 on MCP Atlas at 88.1%. This browsecomp leaderboard ranks models by BrowseComp. Scores come from public evals. Prices are the live API rates in the table above.

    Sources: BrowseComp (OpenAI) (https://openai.com/index/browsecomp/); OpenAI API pricing (https://developers.openai.com/api/docs/pricing); Anthropic Claude API pricing (https://platform.claude.com/docs/en/about-claude/pricing); Gemini API pricing (https://ai.google.dev/gemini-api/docs/pricing)

    Top 8 for browsecomp. Ranked by the BrowseComp score Input and output are dollars per million tokens.
    RankModelBrowseCompInput /MOutput /M
    1Kimi K391.2%$3.00$15.00
    2Claude Opus 590.8%$5.00$25.00
    3GPT-5.6 Sol90.4%$5.00$30.00
    4GPT-5.5 Pro90.1%$60.00$480.00
    5Claude Mythos 588%$10.00$50.00
    6GPT-5.6 Terra87.5%$2.00$12.00
    7Claude Mythos Preview86.9%$10.00$50.00
    8Kimi K2.686.3%$0.96$3.97

    BrowseComp FAQ

    Who ranks #1 on the BrowseComp leaderboard?

    As of August 20, 2026, Kimi K3 by Moonshot AI ranks #1 on BrowseComp at 91.2%. API pricing is $3.00/M input and $15.00/M output.

    What are the top models on BrowseComp?

    The current BrowseComp ranking as of August 20, 2026 is 1. Kimi K3 at 91.2%; 2. Claude Opus 5 at 90.8%; 3. GPT-5.6 Sol at 90.4%.

    Which browsecomp model is the cheapest?

    Nemotron 3.5 Lightning (30B A3B) is the cheapest scored model on this browsecomp leaderboard at $0.05/M input and $0.20/M output ($0.25 blended). Kimi K3 still leads BrowseComp at 91.2%.

    Should I always pick the #1 BrowseComp model?

    Not automatically. Kimi K3 leads BrowseComp, but a cheaper scored model can be the better production choice if the quality gap is small. Use the table to weigh BrowseComp against input/output price, context window, and related evals.

    How often is the BrowseComp leaderboard updated?

    Scores and API prices on this page are refreshed from published evals and provider rates. The snapshot is labeled August 20, 2026. Treat it as a current index, not a one-off blog post.

    What is BrowseComp?

    BrowseComp tests whether a model can search the web, follow sources, and answer hard research questions that need more than one lookup.

    Is BrowseComp a search-engine ranking?

    No. It ranks the model that is doing the browsing, not Google or Bing. Use it when you are buying an agent that researches on the user's behalf.

    Why is SimpleQA on this page?

    SimpleQA checks short-form factuality. A model that browses well should also keep hallucination rates down on easier questions.