FinanceAgent v1.1

    🏆 Leaderboard

    As of August 20, 2026, GPT-5.4 is #1 for FinanceAgent v1.1 at 79.6%. Ranked by FinanceAgent: financial analysis tasks, not a generic chat vibe. 39 models in this index have a published FinanceAgent v1.1 score. FinanceAgent v1.1 leaderboard with live API prices. FinanceAgent v1.1 — evaluates financial analysis, modeling, and data interpretation.

    Updated August 20, 2026282 models33 providers
    GPT-5.4
    OpenAI · Proprietary
    79.6%$17.50
    GPT-5.3 Codex
    OpenAI · Proprietary
    79.2%$15.75
    Claude Opus 4.7
    Anthropic · Proprietary
    64.4%$30.00
    Claude Sonnet 4.6
    Anthropic · Proprietary
    63.3%$18.00
    Claude Opus 4.6
    Anthropic · Proprietary
    60.7%$30.00
    DeepSeek-V4-Pro-0813OSS
    DeepSeek · Open Source
    60.4%$5.28
    GPT-5.5
    OpenAI · Proprietary
    60%$35.00
    Gemini 3.1 Pro
    Google · Proprietary
    59.7%$17.50
    GPT-5.2
    OpenAI · Proprietary
    58.5%$15.75
    Gemini 3.5 Flash
    Google · Proprietary
    57.9%$10.50
    GLM-5.1OSS
    Z AI · Open Source
    57.7%$5.80
    Kimi K2.6OSS
    Moonshot AI · Open Source
    57.1%$4.93
    Seed 2.1 Turbo
    ByteDance · Proprietary
    56%$3.00
    GPT-5.1
    OpenAI · Proprietary
    55.3%$11.25
    Gemini 3 Pro
    Google · Proprietary
    55.1%$14.00
    Qwen3.6 Plus
    Qwen · Proprietary
    54.6%$3.50
    Claude Opus 4.8
    Anthropic · Proprietary
    53.9%$30.00
    Grok 4.3
    xAI · Proprietary
    53.8%$3.75
    Nemotron 3 Ultra (550B A55B)OSS
    NVIDIA · Open Source · via OpenRouter
    53.7%$2.70
    GPT-5.4 mini
    OpenAI · Proprietary
    53.4%$5.25
    Grok-4.1 Fast Reasoning
    xAI · Proprietary
    52.5%$0.70
    GPT-5
    OpenAI · Proprietary
    52.1%$11.25
    GPT-5 mini
    OpenAI · Proprietary
    51.9%$2.25
    Gemma 4 31BOSS
    Google · Open Source
    50.8%$0.54
    MiniMax M2.7OSS
    MiniMax · Open Source
    48.4%$1.50
    GPT-5.4 nano
    OpenAI · Proprietary
    47.8%$1.45
    Gemini 3 Flash
    Google · Proprietary
    47.6%$3.50
    Gemini 3.1 Flash-Lite
    Google · Proprietary
    46.1%$1.75
    Mistral Medium 3.5OSS
    Mistral · Open Source · via Mistral AI
    46.1%$9.00
    Grok-4 Fast Reasoning
    xAI · Proprietary
    46.1%$0.70
    GLM-4.7OSS
    Z AI · Open Source
    46.0%$2.80
    Grok-4.1 Fast Non-Reasoning
    xAI · Proprietary
    44.4%$0.70
    Qwen3 Max
    Qwen · Proprietary
    44.3%$5.50
    Gemini 2.5 Pro
    Google · Proprietary
    41.6%$11.25
    MiniMax M2.5OSS
    MiniMax · Open Source
    38.6%$1.50
    GLM-4.6OSS
    Z AI · Open Source
    36.5%$2.80
    MiniMax M2.1OSS
    MiniMax · Open Source
    33.4%$1.50
    GPT OSS 120BOSS
    OpenAI · Open Source
    21.5%$0.54
    GPT-4o
    OpenAI · Proprietary
    8.1%$12.50
    ChatGPT-4o Latest
    OpenAI · Proprietary
    $12.50
    Claude 3 Haiku
    Anthropic · Proprietary
    $1.50
    Claude 3 Opus
    Anthropic · Proprietary
    $90.00
    Claude 3 Sonnet
    Anthropic · Proprietary
    $18.00
    Claude 3.5 Haiku
    Anthropic · Proprietary
    $4.80
    Claude 3.5 Sonnet
    Anthropic · Proprietary
    $18.00
    Claude 3.5 Sonnet
    Anthropic · Proprietary
    $18.00
    Claude 3.7 Sonnet
    Anthropic · Proprietary
    $18.00
    Claude Fable 5
    Anthropic · Proprietary
    $60.00
    Claude Haiku 4.5
    Anthropic · Proprietary
    $6.00
    Claude Mythos 5
    Anthropic · Proprietary
    $60.00
    Showing 150 of 282 models

    Next step

    You found the model. Now ship the product.

    Auth, billing, and the API layer are already decided. Fork 8 finished AI apps, or have us build the first version with you.

    The index

    All Large Language Models

    282 models across 33 providers. Search or jump to a lab — every model page stays linked here.

    Baidu

    1 models

    Inception

    1 models

    inclusionAI

    1 models

    LG AI Research

    1 models

    Nous Research

    1 models

    Sakana AI

    1 models

    StepFun

    1 models

    Tencent

    1 models

    Thinking Machines

    1 models

    Unisound

    1 models

    Upstage

    1 models

    Which model leads FinanceAgent v1.1 right now?

    As of August 20, 2026, GPT-5.4 by OpenAI is #1 for FinanceAgent v1.1 at 79.6%. Ranked by FinanceAgent: financial analysis tasks, not a generic chat vibe. This board also tracks FinanceAgent v1.1. Next on the same board: GPT-5.3 Codex and Claude Opus 4.7. This financeagent v1.1 leaderboard ranks models by FinanceAgent v1.1. Scores come from public evals. Prices are the live API rates in the table above.

    Sources: OpenAI API pricing (https://developers.openai.com/api/docs/pricing); Anthropic Claude API pricing (https://platform.claude.com/docs/en/about-claude/pricing); Gemini API pricing (https://ai.google.dev/gemini-api/docs/pricing)

    Top 8 for financeagent v1.1. Ranked by FinanceAgent: financial analysis tasks, not a generic chat vibe. Input and output are dollars per million tokens.
    RankModelFinanceAgent v1.1Input /MOutput /M
    1GPT-5.479.6%$2.50$15.00
    2GPT-5.3 Codex79.2%$1.75$14.00
    3Claude Opus 4.764.4%$5.00$25.00
    4Claude Sonnet 4.663.3%$3.00$15.00
    5Claude Opus 4.660.7%$5.00$25.00
    6DeepSeek-V4-Pro-081360.4%$1.32$3.96
    7GPT-5.560%$5.00$30.00
    8Gemini 3.1 Pro59.7%$2.50$15.00

    FinanceAgent v1.1 FAQ

    Who ranks #1 on the FinanceAgent v1.1 leaderboard?

    As of August 20, 2026, GPT-5.4 by OpenAI ranks #1 on FinanceAgent v1.1 at 79.6%. API pricing is $2.50/M input and $15.00/M output.

    What are the top models on FinanceAgent v1.1?

    The current FinanceAgent v1.1 ranking as of August 20, 2026 is 1. GPT-5.4 at 79.6%; 2. GPT-5.3 Codex at 79.2%; 3. Claude Opus 4.7 at 64.4%.

    Which financeagent v1.1 model is the cheapest?

    Gemma 4 31B is the cheapest scored model on this financeagent v1.1 leaderboard at $0.14/M input and $0.40/M output ($0.54 blended). GPT-5.4 still leads FinanceAgent v1.1 at 79.6%.

    Should I always pick the #1 FinanceAgent v1.1 model?

    Not automatically. GPT-5.4 leads FinanceAgent v1.1, but a cheaper scored model can be the better production choice if the quality gap is small. Use the table to weigh FinanceAgent v1.1 against input/output price, context window, and related evals.

    How often is the FinanceAgent v1.1 leaderboard updated?

    Scores and API prices on this page are refreshed from published evals and provider rates. The snapshot is labeled August 20, 2026. Treat it as a current index, not a one-off blog post.

    What is FinanceAgent v1.1?

    FinanceAgent v1.1 — evaluates financial analysis, modeling, and data interpretation. This page ranks models that have published a FinanceAgent v1.1 score, with live API token prices on the same row.

    Where is the FinanceAgent v1.1 leaderboard?

    This page is the FinanceAgent v1.1 leaderboard. Models are sorted by FinanceAgent v1.1, with input and output token prices on the same row so you can weigh score against cost. Official boards often omit price; that comparison is the point of this index.

    How is this FinanceAgent v1.1 ranking different from the official board?

    Official eval pages own the methodology. This page keeps the published FinanceAgent v1.1 score next to live API $/M so you can pick a production SKU, not only a trophy number.