BrowseComp

    ๐Ÿ† Leaderboard

    As of August 31, 2026, Kimi K3 is #1 for BrowseComp at 91.2%. Ranked by the BrowseComp score 62 models in this index have a published BrowseComp score. Methodology: BrowseComp (OpenAI) (https://openai.com/index/browsecomp/). BrowseComp leaderboard with live API prices. BrowseComp โ€” evaluates web browsing and information retrieval across complex multi-step research tasks. Official methodology: BrowseComp (OpenAI) (https://openai.com/index/browsecomp/).

    Updated August 31, 2026345 models36 providers

    BrowseComp vs price

    61 models. Left is cheaper. Up is a higher score. The line is the best score you can buy at each price โ€” a dot under it is a worse deal than something on the line.

    Best score at each price#1 on this board
    Kimi K3OSS
    Moonshot AI ยท Open Source
    91.2%$18.00
    Claude Opus 5
    Anthropic ยท Proprietary
    90.8%$30.00
    GPT-5.6 Sol
    OpenAI ยท Proprietary
    90.4%$35.00
    GPT-5.5 Pro
    OpenAI ยท Proprietary
    90.1%$540.00
    Claude Mythos 5
    Anthropic ยท Proprietary
    88%$60.00
    GPT-5.6 Terra
    OpenAI ยท Proprietary
    87.5%$14.00
    Claude Mythos Preview
    Anthropic ยท Proprietary
    86.9%$60.00
    Kimi K2.6OSS
    Moonshot AI ยท Open Source
    86.3%$4.93
    Seed 2.1 Pro
    ByteDance ยท Proprietary
    86.2%$5.34
    Gemini 3.1 Pro
    Google ยท Proprietary
    85.9%$17.50
    Seed 2.1 Turbo
    ByteDance ยท Proprietary
    84.9%$3.00
    Claude Sonnet 5
    Anthropic ยท Proprietary
    84.7%$12.00
    GPT-5.5
    OpenAI ยท Proprietary
    84.4%$35.00
    Claude Opus 4.8
    Anthropic ยท Proprietary
    84.3%$30.00
    Hy3OSS
    Tencent ยท Open Source
    84.2%$0.66
    Claude Opus 4.6
    Anthropic ยท Proprietary
    84%$30.00
    MiniMax M3OSS
    MiniMax ยท Open Source
    83.5%$1.50
    DeepSeek-V4-Pro-MaxOSS
    DeepSeek ยท Open Source
    83.4%$5.22
    GPT-5.6 Luna
    OpenAI ยท Proprietary
    83.3%$1.40
    GPT-5.4
    OpenAI ยท Proprietary
    82.7%$17.50
    GLM-5.1OSS
    Z AI ยท Open Source
    79.3%$5.80
    Claude Opus 4.7
    Anthropic ยท Proprietary
    79.3%$30.00
    GPT-5.2 Pro
    OpenAI ยท Proprietary
    77.9%$189.00
    Inkling-SmallOSS
    Thinking Machines ยท Open Source ยท via Thinking Machines Lab
    77.4%$1.50
    Seed 2.0 Pro
    ByteDance ยท Proprietary
    77.3%$3.50
    MiniMax M2.5OSS
    MiniMax ยท Open Source
    76.3%$1.50
    GLM-5OSS
    Z AI ยท Open Source
    75.9%$4.20
    Kimi K2.5OSS
    Moonshot AI ยท Open Source
    74.9%$3.68
    Claude Sonnet 4.6
    Anthropic ยท Proprietary
    74.7%$18.00
    DeepSeek-V4-Flash-MaxOSS
    DeepSeek ยท Open Source
    73.2%$0.42
    Step-3.5-FlashOSS
    StepFun ยท Open Source
    69%$0.50
    Qwen3.5-397B-A17BOSS
    Qwen ยท Open Source
    69%$4.20
    GPT-5.2
    OpenAI ยท Proprietary
    65.8%$15.75
    Qwen3.5-122B-A10BOSS
    Qwen ยท Open Source
    63.8%$3.60
    MiniMax M2.1OSS
    MiniMax ยท Open Source
    62%$1.50
    Qwen3.5-35B-A3BOSS
    Qwen ยท Open Source
    61%$2.25
    Qwen3.5-27BOSS
    Qwen ยท Open Source
    61%$2.70
    Kimi K2-Thinking-0905OSS
    Moonshot AI ยท Open Source
    60.2%$2.47
    MiMo-V2-FlashOSS
    Xiaomi ยท Open Source
    58.3%$0.40
    LongCat-Flash-Thinking-2601OSS
    Meituan ยท Open Source
    56.6%$1.50
    GPT-5
    OpenAI ยท Proprietary
    54.9%$11.25
    DeepSeek-V4-Flash-0423OSS
    DeepSeek ยท Open Source
    53.5%$0.30
    GLM-4.7OSS
    Z AI ยท Open Source
    52%$2.80
    o4-mini
    OpenAI ยท Proprietary
    51.5%$5.50
    DeepSeek-V3.2OSS
    DeepSeek ยท Open Source ยท via OpenRouter
    51.4%$0.57
    o3
    OpenAI ยท Proprietary
    49.7%$10.00
    Sarvam-105BOSS
    Sarvam AI ยท Open Source
    49.5%$1.60
    Solar Pro 4
    Upstage ยท Proprietary
    49.2%$1.50
    Mistral Medium 3.5OSS
    Mistral ยท Open Source ยท via Mistral AI
    48.6%$9.00
    GLM-4.6OSS
    Z AI ยท Open Source
    45.1%$2.80
    Showing 1โ€“50 of 345 models

    Next step

    You found the model. Now ship the product.

    Auth, billing, and the API layer are already decided. Fork 8 finished AI apps, or have us build the first version with you.

    The index

    All Large Language Models

    345 models across 36 providers. Search or jump to a lab โ€” every model page stays linked here.

    Baidu

    2 models

    Inception

    1 models

    inclusionAI

    1 models

    LG AI Research

    1 models

    Liquid AI

    2 models

    Nous Research

    1 models

    OpenBMB

    1 models

    Sakana AI

    1 models

    Sarvam AI

    2 models

    Tencent

    1 models

    Thinking Machines

    1 models

    Unisound

    1 models

    Upstage

    1 models

    Which model leads BrowseComp right now?

    As of August 31, 2026, Kimi K3 by Moonshot AI is #1 for BrowseComp at 91.2%. Ranked by the BrowseComp score This board also tracks BrowseComp. Next on the same board: Claude Opus 5 and GPT-5.6 Sol. This browsecomp leaderboard ranks models by BrowseComp. Scores come from public evals. Prices are the live API rates in the table above.

    Sources: BrowseComp (OpenAI) (https://openai.com/index/browsecomp/); OpenAI API pricing (https://developers.openai.com/api/docs/pricing); Anthropic Claude API pricing (https://platform.claude.com/docs/en/about-claude/pricing); Gemini API pricing (https://ai.google.dev/gemini-api/docs/pricing)

    Top 8 for browsecomp. Ranked by the BrowseComp score Input and output are dollars per million tokens.
    RankModelBrowseCompInput /MOutput /M
    1Kimi K391.2%$3.00$15.00
    2Claude Opus 590.8%$5.00$25.00
    3GPT-5.6 Sol90.4%$5.00$30.00
    4GPT-5.5 Pro90.1%$60.00$480.00
    5Claude Mythos 588%$10.00$50.00
    6GPT-5.6 Terra87.5%$2.00$12.00
    7Claude Mythos Preview86.9%$10.00$50.00
    8Kimi K2.686.3%$0.96$3.97

    BrowseComp FAQ

    Who ranks #1 on the BrowseComp leaderboard?

    As of August 31, 2026, Kimi K3 by Moonshot AI ranks #1 on BrowseComp at 91.2%. API pricing is $3.00/M input and $15.00/M output.

    What are the top models on BrowseComp?

    The current BrowseComp ranking as of August 31, 2026 is 1. Kimi K3 at 91.2%; 2. Claude Opus 5 at 90.8%; 3. GPT-5.6 Sol at 90.4%.

    Which browsecomp model is the cheapest?

    Nemotron 3.5 Lightning (30B A3B) is the cheapest scored model on this browsecomp leaderboard at $0.05/M input and $0.20/M output ($0.25 blended). Kimi K3 still leads BrowseComp at 91.2%.

    Should I always pick the #1 BrowseComp model?

    Not automatically. Kimi K3 leads BrowseComp, but a cheaper scored model can be the better production choice if the quality gap is small. Use the table to weigh BrowseComp against input/output price, context window, and related evals.

    How often is the BrowseComp leaderboard updated?

    Scores and API prices on this page are refreshed from published evals and provider rates. The snapshot is labeled August 31, 2026. Treat it as a current index, not a one-off blog post.

    What is BrowseComp?

    BrowseComp โ€” evaluates web browsing and information retrieval across complex multi-step research tasks. This page ranks models that have published a BrowseComp score, with live API token prices on the same row. Official methodology: BrowseComp (OpenAI) (https://openai.com/index/browsecomp/).

    Where is the BrowseComp leaderboard?

    This page is the BrowseComp leaderboard. Models are sorted by BrowseComp, with input and output token prices on the same row so you can weigh score against cost. Official boards often omit price; that comparison is the point of this index.

    How is this BrowseComp ranking different from the official board?

    The official BrowseComp (OpenAI) page owns the methodology. This page keeps the published BrowseComp score next to live API $/M so you can pick a production SKU, not only a trophy number. Source: https://openai.com/index/browsecomp/