K-EXAONE-236B-A23B vs Llama-3.3 Nemotron Super 49B v1

    Comparison

    As of August 30, 2026, Llama-3.3 Nemotron Super 49B v1 is 69% cheaper per million tokens than K-EXAONE-236B-A23B. K-EXAONE-236B-A23B is $0.60 / $1.00 per million input/output tokens versus Llama-3.3 Nemotron Super 49B v1 at $0.10 / $0.40. Sources: https://anotherwrapper.com/tools/llm-pricing and https://build.nvidia.com/nvidia/llama-3_3-nemotron-super-49b-v1

    Sources: AnotherWrapper LLM pricing (https://anotherwrapper.com/tools/llm-pricing); NVIDIA pricing (https://build.nvidia.com/nvidia/llama-3_3-nemotron-super-49b-v1)

    K-EXAONE-236B-A23B

    LG AI Research

    $1.60blended / 1M

    Input

    $0.60

    Output

    $1.00

    32.8K ctx|Proprietary
    69%
    cheaper
    Better value

    Llama-3.3 Nemotron Super 49B v1

    NVIDIA

    $0.50blended / 1M

    Input

    $0.10

    Output

    $0.40

    131.1K ctx|Open Source
    Save $1.10 per million tokens by choosing Llama-3.3 Nemotron Super 49B v1 over K-EXAONE-236B-A23B.

    K-EXAONE-236B-A23B vs Llama-3.3 Nemotron Super 49B v1 comparison

    Token pricing, specs, and benchmarks side by side

    Shared benchmarks

    Each bar is share of the current leader on that eval. Click a row to open the board.

    K-EXAONE-236B-A23BLlama-3.3 Nemotron Super 49B v1

    Score vs price

    AIME 2025. Left is cheaper. Up is a higher score. The line is the best score you can buy at each price.

    Best score at each priceK-EXAONE-236B-A23BLlama-3.3 Nemotron Super 49B v1
    Metric
    K-EXAONE-236B-A23B
    Llama-3.3 Nemotron Super 49B v1
    Provider
    Provider
    LG AI Research
    NVIDIA
    License
    Proprietary
    Open Source
    Release Date
    2025-12-31
    2025-03-18
    Pricing (per 1M tokens)
    Input Price+500%
    $0.60
    $0.10
    Output Price+150%
    $1.00
    $0.40
    Blended (1M + 1M)
    $1.60
    $0.50
    Model Details
    Context Window
    32.8K
    131.1K
    Max Output Tokens
    32.8K
    N/A
    Knowledge Cutoff
    2025-10-01
    2023-12-31
    Throughput
    50 tok/s
    N/A
    Benchmarks
    66.7%
    67.3%
    85.7%
    92.8%
    58.4%
    96.6%
    88.3%
    73.7%
    91.3%
    83.8%
    91.7%
    73.2%
    Benchmark Wins
    K-EXAONE-236B-A23B 1|0 Llama-3.3 Nemotron Super 49B v1

    Verdict

    K-EXAONE-236B-A23B vs Llama-3.3 Nemotron Super 49B v1: the bottom line

    Llama-3.3 Nemotron Super 49B v1 offers significantly lower pricing, while K-EXAONE-236B-A23B leads on benchmark performance. Your choice depends on whether cost efficiency or raw capability matters more for your use case.

    K-EXAONE-236B-A23B vs Llama-3.3 Nemotron Super 49B v1 FAQ

    Which is cheaper, K-EXAONE-236B-A23B or Llama-3.3 Nemotron Super 49B v1?

    As of August 30, 2026, Llama-3.3 Nemotron Super 49B v1 is 69% cheaper on blended LLM API pricing. Llama-3.3 Nemotron Super 49B v1 costs $0.10 per million input tokens and $0.40 per million output tokens ($0.50 blended 1M-in + 1M-out). K-EXAONE-236B-A23B costs $0.60 / $1.00 per million tokens ($1.60 blended). Sources: https://anotherwrapper.com/tools/llm-pricing and https://build.nvidia.com/nvidia/llama-3_3-nemotron-super-49b-v1. The cheapest LLM for your app still depends on how many output tokens you generate.

    How much does K-EXAONE-236B-A23B vs Llama-3.3 Nemotron Super 49B v1 cost per million tokens?

    K-EXAONE-236B-A23B LG AI Research API pricing is $0.60 input and $1.00 output per million tokens via LG AI Research. Llama-3.3 Nemotron Super 49B v1 NVIDIA API pricing is $0.10 input and $0.40 output per million tokens via NVIDIA. Use those four numbers, not ChatGPT Plus or Claude Pro subscription prices, when you are comparing APIs.

    Which is better for coding, K-EXAONE-236B-A23B or Llama-3.3 Nemotron Super 49B v1?

    Neither K-EXAONE-236B-A23B nor Llama-3.3 Nemotron Super 49B v1 has a published SWE-bench or LiveCodeBench score on this page yet. Use the comparison table for the evals that do exist.

    What is the context window for K-EXAONE-236B-A23B vs Llama-3.3 Nemotron Super 49B v1?

    K-EXAONE-236B-A23B supports a 33K token context window, while Llama-3.3 Nemotron Super 49B v1 supports 131K tokens. Llama-3.3 Nemotron Super 49B v1 offers a larger context window. Filling a larger window bills more input tokens, so the cheaper-per-million model can still cost more on long documents.

    Which model performs better on benchmarks, K-EXAONE-236B-A23B or Llama-3.3 Nemotron Super 49B v1?

    K-EXAONE-236B-A23B leads on 1 of 1 shared benchmarks versus Llama-3.3 Nemotron Super 49B v1's 0 wins. Check the comparison table for GPQA Diamond, SWE-bench, MMLU, HLE, and the other evals we track. Token price and benchmark score together are the usual LLM comparison, not either number alone.

    Is K-EXAONE-236B-A23B or Llama-3.3 Nemotron Super 49B v1 better for production use?

    Both K-EXAONE-236B-A23B and Llama-3.3 Nemotron Super 49B v1 are production API models. For cost-sensitive production traffic, Llama-3.3 Nemotron Super 49B v1 has better economics on the blended 1M-in + 1M-out scale. For maximum capability, weight the benchmark table for your domain. Many production stacks route cheap models for drafts and a frontier model for the hard turn, which is usually cheaper than sending everything to the expensive API.

    Can I switch between K-EXAONE-236B-A23B and Llama-3.3 Nemotron Super 49B v1 in my app?

    Yes. Call K-EXAONE-236B-A23B through LG AI Research and Llama-3.3 Nemotron Super 49B v1 through NVIDIA with separate API keys, or through a gateway that already wraps both. Keep prompts in tokens, not characters, when you estimate the invoice. AnotherWrapper templates can swap providers without rewriting auth and billing.

    How accurate is this K-EXAONE-236B-A23B vs Llama-3.3 Nemotron Super 49B v1 pricing data?

    List rates are the latest published LG AI Research API pricing and NVIDIA API pricing, quoted per million tokens. Official rate sheets: https://anotherwrapper.com/tools/llm-pricing and https://build.nvidia.com/nvidia/llama-3_3-nemotron-super-49b-v1. Prompt caching, batch APIs, and committed-use discounts can change the invoice. Benchmark scores are from official publications and independent evals, updated as new numbers land.

    What is the output token limit for K-EXAONE-236B-A23B and Llama-3.3 Nemotron Super 49B v1?

    K-EXAONE-236B-A23B supports up to 33K output tokens per request, while Llama-3.3 Nemotron Super 49B v1 supports up to an unspecified number of output tokens. Output tokens are usually the expensive half of API pricing, so a higher max output cap is a capability, not a discount.

    Which model has better throughput, K-EXAONE-236B-A23B or Llama-3.3 Nemotron Super 49B v1?

    K-EXAONE-236B-A23B has a measured throughput of approximately 50 tokens/second. Throughput data for Llama-3.3 Nemotron Super 49B v1 is not yet available. Throughput (tokens per second) and time-to-first-token are why two models with the same per-million-token price can still feel different in chat and agents.

    How do I switch the models in this K-EXAONE-236B-A23B vs Llama-3.3 Nemotron Super 49B v1 comparison?

    Use the model pickers at the top of this LLM comparison. Search any model in the index and the URL updates. Open K-EXAONE-236B-A23B or Llama-3.3 Nemotron Super 49B v1 from the sidebar for the single-model API pricing page, or go back to the LLM pricing table to sort by cheapest blended cost.

    K-EXAONE-236B-A23B vs Llama-3.3 Nemotron Super 49B v1 pricing

    This comparison covers API pricing, context window, throughput, and benchmark scores. K-EXAONE-236B-A23B costs $0.60/M input and $1.00/M output. Llama-3.3 Nemotron Super 49B v1 costs $0.10/M input and $0.40/M output.

    Rank both on the LLM leaderboard, SWE-bench, and Humanity's Last Exam.

    The index

    All Large Language Models

    343 models across 36 providers. Search or jump to a lab — every model page stays linked here.

    Baidu

    2 models

    Inception

    1 models

    inclusionAI

    1 models

    LG AI Research

    0 models

    Liquid AI

    2 models

    Nous Research

    1 models

    OpenBMB

    1 models

    Sakana AI

    1 models

    Sarvam AI

    2 models

    Tencent

    1 models

    Thinking Machines

    1 models

    Unisound

    1 models

    Upstage

    1 models

    Next step

    You found the model. Now ship the product.

    Auth, billing, and the API layer are already decided. Fork 8 finished AI apps, or have us build the first version with you.