Video-MME

    ๐Ÿ† Leaderboard

    As of August 20, 2026, Gemini 1.5 Pro is #1 for Video-MME at 75%. Ranked by the Video-MME score 7 models in this index have a published Video-MME score. Video-MME leaderboard: rank models by Video-MME next to live API token prices.

    Updated August 20, 2026282 models33 providers
    Gemini 1.5 Pro
    Google ยท Proprietary
    75%$12.50
    Qwen2.5 VL 72B InstructOSS
    Qwen ยท Open Source ยท via OpenRouter
    73.5%$1.80
    GPT-4o
    OpenAI ยท Proprietary
    71.9%$12.50
    Gemini 1.5 Flash
    Google ยท Proprietary
    70.3%$0.75
    GPT-4o mini
    OpenAI ยท Proprietary
    64.8%$0.75
    Claude 3.5 Sonnet
    Anthropic ยท Proprietary
    60%$18.00
    Claude 3.5 Sonnet
    Anthropic ยท Proprietary
    60%$18.00
    ChatGPT-4o Latest
    OpenAI ยท Proprietary
    โ€”$12.50
    Claude 3 Haiku
    Anthropic ยท Proprietary
    โ€”$1.50
    Claude 3 Opus
    Anthropic ยท Proprietary
    โ€”$90.00
    Claude 3 Sonnet
    Anthropic ยท Proprietary
    โ€”$18.00
    Claude 3.5 Haiku
    Anthropic ยท Proprietary
    โ€”$4.80
    Claude 3.7 Sonnet
    Anthropic ยท Proprietary
    โ€”$18.00
    Claude Fable 5
    Anthropic ยท Proprietary
    โ€”$60.00
    Claude Haiku 4.5
    Anthropic ยท Proprietary
    โ€”$6.00
    Claude Mythos 5
    Anthropic ยท Proprietary
    โ€”$60.00
    Claude Mythos Preview
    Anthropic ยท Proprietary
    โ€”$60.00
    Claude Opus 4
    Anthropic ยท Proprietary
    โ€”$90.00
    Claude Opus 4.1
    Anthropic ยท Proprietary
    โ€”$90.00
    Claude Opus 4.5
    Anthropic ยท Proprietary
    โ€”$30.00
    Claude Opus 4.6
    Anthropic ยท Proprietary
    โ€”$30.00
    Claude Opus 4.7
    Anthropic ยท Proprietary
    โ€”$30.00
    Claude Opus 4.8
    Anthropic ยท Proprietary
    โ€”$30.00
    Claude Opus 5
    Anthropic ยท Proprietary
    โ€”$30.00
    Claude Sonnet 4
    Anthropic ยท Proprietary
    โ€”$18.00
    Claude Sonnet 4.5
    Anthropic ยท Proprietary
    โ€”$18.00
    Claude Sonnet 4.6
    Anthropic ยท Proprietary
    โ€”$18.00
    Claude Sonnet 5
    Anthropic ยท Proprietary
    โ€”$12.00
    Command A+OSS
    Cohere ยท Open Source
    โ€”$12.50
    Command R+OSS
    Cohere ยท Open Source
    โ€”$1.25
    Composer 2
    Cursor ยท Proprietary
    โ€”$3.00
    Composer 2 Fast
    Cursor ยท Proprietary
    โ€”$9.00
    DeepSeek R1 Distill Llama 70BOSS
    DeepSeek ยท Open Source
    โ€”$0.50
    DeepSeek R1 Distill Qwen 32BOSS
    DeepSeek ยท Open Source
    โ€”$0.30
    DeepSeek-R1OSS
    DeepSeek ยท Open Source
    โ€”$2.74
    DeepSeek-R1-0528OSS
    DeepSeek ยท Open Source
    โ€”$2.74
    DeepSeek-V2.5OSS
    DeepSeek ยท Open Source
    โ€”$0.42
    DeepSeek-V3OSS
    DeepSeek ยท Open Source
    โ€”$1.37
    DeepSeek-V3 0324OSS
    DeepSeek ยท Open Source
    โ€”$1.42
    DeepSeek-V3.1OSS
    DeepSeek ยท Open Source
    โ€”$1.27
    DeepSeek-V3.2OSS
    DeepSeek ยท Open Source ยท via OpenRouter
    โ€”$0.57
    DeepSeek-V3.2 (Non-thinking)OSS
    DeepSeek ยท Open Source
    โ€”$0.70
    DeepSeek-V3.2-ExpOSS
    DeepSeek ยท Open Source
    โ€”$0.68
    DeepSeek-V4-Flash-0423OSS
    DeepSeek ยท Open Source
    โ€”$0.30
    DeepSeek-V4-Flash-0731OSS
    DeepSeek ยท Open Source
    โ€”$1.76
    DeepSeek-V4-Flash-MaxOSS
    DeepSeek ยท Open Source
    โ€”$0.42
    DeepSeek-V4-Pro-0813OSS
    DeepSeek ยท Open Source
    โ€”$5.28
    DeepSeek-V4-Pro-MaxOSS
    DeepSeek ยท Open Source
    โ€”$5.22
    Devstral Medium
    Mistral ยท Proprietary ยท via Mistral AI
    โ€”$2.40
    Devstral Small 1.1OSS
    Mistral ยท Open Source ยท via Mistral AI
    โ€”$0.40
    Showing 1โ€“50 of 282 models

    Next step

    You found the model. Now ship the product.

    Auth, billing, and the API layer are already decided. Fork 8 finished AI apps, or have us build the first version with you.

    The index

    All Large Language Models

    282 models across 33 providers. Search or jump to a lab โ€” every model page stays linked here.

    Baidu

    1 models

    Inception

    1 models

    inclusionAI

    1 models

    LG AI Research

    1 models

    Nous Research

    1 models

    Sakana AI

    1 models

    StepFun

    1 models

    Tencent

    1 models

    Thinking Machines

    1 models

    Unisound

    1 models

    Upstage

    1 models

    Which model leads Video-MME right now?

    As of August 20, 2026, Gemini 1.5 Pro by Google is #1 for Video-MME at 75%. Ranked by the Video-MME score This board also tracks Video-MME. Next on the same board: Qwen2.5 VL 72B Instruct and GPT-4o. This video-mme leaderboard ranks models by Video-MME. Scores come from public evals. Prices are the live API rates in the table above.

    Sources: OpenAI API pricing (https://developers.openai.com/api/docs/pricing); Anthropic Claude API pricing (https://platform.claude.com/docs/en/about-claude/pricing); Gemini API pricing (https://ai.google.dev/gemini-api/docs/pricing)

    Top 7 for video-mme. Ranked by the Video-MME score Input and output are dollars per million tokens.
    RankModelVideo-MMEInput /MOutput /M
    1Gemini 1.5 Pro75%$2.50$10.00
    2Qwen2.5 VL 72B Instruct73.5%$0.80$1.00
    3GPT-4o71.9%$2.50$10.00
    4Gemini 1.5 Flash70.3%$0.15$0.60
    5GPT-4o mini64.8%$0.15$0.60
    6Claude 3.5 Sonnet60%$3.00$15.00
    7Claude 3.5 Sonnet60%$3.00$15.00

    Video-MME FAQ

    Who ranks #1 on the Video-MME leaderboard?

    As of August 20, 2026, Gemini 1.5 Pro by Google ranks #1 on Video-MME at 75%. API pricing is $2.50/M input and $10.00/M output.

    What are the top models on Video-MME?

    The current Video-MME ranking as of August 20, 2026 is 1. Gemini 1.5 Pro at 75%; 2. Qwen2.5 VL 72B Instruct at 73.5%; 3. GPT-4o at 71.9%.

    Which video-mme model is the cheapest?

    Gemini 1.5 Flash is the cheapest scored model on this video-mme leaderboard at $0.15/M input and $0.60/M output ($0.75 blended). Gemini 1.5 Pro still leads Video-MME at 75%.

    Should I always pick the #1 Video-MME model?

    Not automatically. Gemini 1.5 Pro leads Video-MME, but a cheaper scored model can be the better production choice if the quality gap is small. Use the table to weigh Video-MME against input/output price, context window, and related evals.

    How often is the Video-MME leaderboard updated?

    Scores and API prices on this page are refreshed from published evals and provider rates. The snapshot is labeled August 20, 2026. Treat it as a current index, not a one-off blog post.

    What is Video-MME?

    Video-MME is a public LLM eval (the Video-MME score). This page ranks models that have published a score, next to live API prices.

    Where is the Video-MME leaderboard?

    This page is the Video-MME leaderboard. Models are sorted by Video-MME, with input and output token prices on the same row so you can weigh score against cost. Official boards often omit price; that comparison is the point of this index.

    How is this Video-MME ranking different from the official board?

    Official eval pages own the methodology. This page keeps the published Video-MME score next to live API $/M so you can pick a production SKU, not only a trophy number.