FindCostShortlistChangesModels

Model API prices & token cost calculator

Find a model

→ 1043 quotes match

Real cost

Compare below

Estimate, not a bill: token and request charges only, in the quote’s own currency. Excludes tools, storage, search, taxes and platform fees. Different models tokenize the same task differently.

Shortlist 0/3

Add quotes from the results with “+ Shortlist”, or let us pick three.

Today

1043 matching · per 1M tokens

Paired input/output rates per provider endpoint for tracked models. Purchased through OpenRouter; endpoints need review after 24 hours.

  • 1.Mistral: Mistral Nemo

    DekaLLM via OpenRouter · dekallm/fp8

    fp8
    Input / 1M
    $0.018
    Output / 1M
    $0.03
    Cache read / 1M
    —
    Context
    131K
    Your cost / mo
    $0.24
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by DekaLLM through OpenRouter, not a direct-provider contract. Published rates for route dekallm/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    dekallm/fp8 · fp8

    Declared parameters: frequency_penalty, logit_bias, logprobs, max_tokens, presence_penalty, response_format, seed, stop, structured_outputs, temperature, top_logprobs, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: mistralai/mistral-nemo
    Protocol: openai
    Provider routing: DekaLLM
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "mistralai/mistral-nemo",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "DekaLLM"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "mistralai/mistral-nemo". Route requests to provider "DekaLLM" by setting the `provider.order` field to ["DekaLLM"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.018 / 1M input tokens, $0.03 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints.
    Price facts
    Input: $0.018 / 1M tokens
    Output: $0.03 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:28.816Z
    Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints
    Price Since Endpoints on OpenRouter
  • 2.Mistral: Mistral Nemo

    DeepInfra via OpenRouter · deepinfra/fp8

    fp8
    Input / 1M
    $0.019
    Output / 1M
    $0.03
    Cache read / 1M
    —
    Context
    131K
    Your cost / mo
    $0.25
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    deepinfra/fp8 · fp8

    Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: mistralai/mistral-nemo
    Protocol: openai
    Provider routing: DeepInfra
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "mistralai/mistral-nemo",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "DeepInfra"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "mistralai/mistral-nemo". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.019 / 1M input tokens, $0.03 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints.
    Price facts
    Input: $0.019 / 1M tokens
    Output: $0.03 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:28.816Z
    Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints
    Price Since Endpoints on OpenRouter
  • 3.Meta: Llama 3.1 8B Instruct

    DeepInfra via OpenRouter · deepinfra/fp8

    fp8
    Input / 1M
    $0.02
    Output / 1M
    $0.04
    Cache read / 1M
    —
    Context
    131K
    Your cost / mo
    $0.28
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    deepinfra/fp8 · fp8

    Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, response_format, seed, stop, temperature, top_k, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: meta-llama/llama-3.1-8b-instruct
    Protocol: openai
    Provider routing: DeepInfra
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "meta-llama/llama-3.1-8b-instruct",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "DeepInfra"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "meta-llama/llama-3.1-8b-instruct". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.02 / 1M input tokens, $0.04 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints.
    Price facts
    Input: $0.02 / 1M tokens
    Output: $0.04 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:26.014Z
    Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints
    Price Since Endpoints on OpenRouter
  • 4.Meta: Llama 3.1 8B Instruct

    Novita via OpenRouter · novita/fp8

    fp8
    Input / 1M
    $0.02
    Output / 1M
    $0.05
    Cache read / 1M
    —
    Context
    16K
    Your cost / mo
    $0.30
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by Novita through OpenRouter, not a direct-provider contract. Published rates for route novita/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    novita/fp8 · fp8

    Declared parameters: frequency_penalty, logprobs, max_tokens, presence_penalty, repetition_penalty, response_format, seed, stop, temperature, top_k, top_logprobs, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: meta-llama/llama-3.1-8b-instruct
    Protocol: openai
    Provider routing: Novita
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "meta-llama/llama-3.1-8b-instruct",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "Novita"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "meta-llama/llama-3.1-8b-instruct". Route requests to provider "Novita" by setting the `provider.order` field to ["Novita"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.02 / 1M input tokens, $0.05 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints.
    Price facts
    Input: $0.02 / 1M tokens
    Output: $0.05 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:26.014Z
    Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints
    Price Since Endpoints on OpenRouter
  • 5.Mistral: Mistral Nemo

    Parasail via OpenRouter · parasail/fp8

    fp8
    Input / 1M
    $0.03
    Output / 1M
    $0.03
    Cache read / 1M
    —
    Context
    131K
    Your cost / mo
    $0.36
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by Parasail through OpenRouter, not a direct-provider contract. Published rates for route parasail/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    parasail/fp8 · fp8

    Declared parameters: frequency_penalty, logit_bias, logprobs, max_tokens, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_logprobs, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: mistralai/mistral-nemo
    Protocol: openai
    Provider routing: Parasail
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "mistralai/mistral-nemo",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "Parasail"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "mistralai/mistral-nemo". Route requests to provider "Parasail" by setting the `provider.order` field to ["Parasail"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.03 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints.
    Price facts
    Input: $0.03 / 1M tokens
    Output: $0.03 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:28.816Z
    Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints
    Price Since Endpoints on OpenRouter
  • 6.OpenAI: gpt-oss-20b

    Darkbloom via OpenRouter · darkbloom/fp8

    fp8
    Input / 1M
    $0.018
    Output / 1M
    $0.09
    Cache read / 1M
    $0.009
    Context
    131K
    Your cost / mo
    $0.36
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by Darkbloom through OpenRouter, not a direct-provider contract. Published rates for route darkbloom/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    $0.009
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    darkbloom/fp8 · fp8

    Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: openai/gpt-oss-20b
    Protocol: openai
    Provider routing: Darkbloom
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "openai/gpt-oss-20b",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "Darkbloom"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "Darkbloom" by setting the `provider.order` field to ["Darkbloom"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.018 / 1M input tokens, $0.09 / 1M output tokens, $0.009 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints.
    Price facts
    Input: $0.018 / 1M tokens
    Output: $0.09 / 1M tokens
    Cache read: $0.009 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:16.420Z
    Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints
    Price Since Endpoints on OpenRouter
  • 7.IBM: Granite 4.0 Micro

    Cloudflare via OpenRouter · cloudflare

    Input / 1M
    $0.017
    Output / 1M
    $0.112
    Cache read / 1M
    —
    Context
    131K
    Your cost / mo
    $0.394
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by Cloudflare through OpenRouter, not a direct-provider contract. Published rates for route cloudflare; unknown quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    cloudflare · unknown

    Declared parameters: frequency_penalty, logit_bias, logprobs, max_tokens, min_p, presence_penalty, repetition_penalty, response_format, seed, stop, temperature, top_k, top_logprobs, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: ibm-granite/granite-4.0-h-micro
    Protocol: openai
    Provider routing: Cloudflare
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "ibm-granite/granite-4.0-h-micro",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "Cloudflare"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "ibm-granite/granite-4.0-h-micro". Route requests to provider "Cloudflare" by setting the `provider.order` field to ["Cloudflare"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.017 / 1M input tokens, $0.112 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/ibm-granite/granite-4.0-h-micro/endpoints.
    Price facts
    Input: $0.017 / 1M tokens
    Output: $0.112 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:24.978Z
    Source: https://openrouter.ai/api/v1/models/ibm-granite/granite-4.0-h-micro/endpoints
    Price Since Endpoints on OpenRouter
  • 8.OpenAI: gpt-oss-20b

    AkashML via OpenRouter · akashml/fp4

    fp4
    Input / 1M
    $0.02
    Output / 1M
    $0.10
    Cache read / 1M
    —
    Context
    131K
    Your cost / mo
    $0.40
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by AkashML through OpenRouter, not a direct-provider contract. Published rates for route akashml/fp4; fp4 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    akashml/fp4 · fp4

    Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: openai/gpt-oss-20b
    Protocol: openai
    Provider routing: AkashML
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "openai/gpt-oss-20b",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "AkashML"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "AkashML" by setting the `provider.order` field to ["AkashML"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.02 / 1M input tokens, $0.10 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints.
    Price facts
    Input: $0.02 / 1M tokens
    Output: $0.10 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:16.420Z
    Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints
    Price Since Endpoints on OpenRouter
  • 9.Nex AGI: Nex-N2.5-Mini

    Nex AGI via OpenRouter · nex-agi/bf16

    bf16
    Input / 1M
    $0.025
    Output / 1M
    $0.10
    Cache read / 1M
    $0.0025
    Context
    262K
    Your cost / mo
    $0.45
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by Nex AGI through OpenRouter, not a direct-provider contract. Published rates for route nex-agi/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    $0.0025
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    nex-agi/bf16 · bf16

    Declared parameters: include_reasoning, logprobs, max_tokens, reasoning, reasoning_effort, structured_outputs, temperature, top_k, top_logprobs, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: nex-agi/nex-n2.5-mini
    Protocol: openai
    Provider routing: Nex AGI
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "nex-agi/nex-n2.5-mini",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "Nex AGI"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "nex-agi/nex-n2.5-mini". Route requests to provider "Nex AGI" by setting the `provider.order` field to ["Nex AGI"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.025 / 1M input tokens, $0.10 / 1M output tokens, $0.0025 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/nex-agi/nex-n2.5-mini/endpoints.
    Price facts
    Input: $0.025 / 1M tokens
    Output: $0.10 / 1M tokens
    Cache read: $0.0025 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:30.211Z
    Source: https://openrouter.ai/api/v1/models/nex-agi/nex-n2.5-mini/endpoints
    Price Since Endpoints on OpenRouter
  • 10.Sao10K: Llama 3 8B Lunaris

    DeepInfra via OpenRouter · deepinfra/turbo

    fp8
    Input / 1M
    $0.04
    Output / 1M
    $0.05
    Cache read / 1M
    —
    Context
    8K
    Your cost / mo
    $0.50
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/turbo; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    deepinfra/turbo · fp8

    Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, seed, stop, temperature, top_k, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: sao10k/l3-lunaris-8b
    Protocol: openai
    Provider routing: DeepInfra
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "sao10k/l3-lunaris-8b",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "DeepInfra"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "sao10k/l3-lunaris-8b". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.04 / 1M input tokens, $0.05 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints.
    Price facts
    Input: $0.04 / 1M tokens
    Output: $0.05 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:44.773Z
    Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints
    Price Since Endpoints on OpenRouter
  • 11.Sao10K: Llama 3 8B Lunaris

    Parasail via OpenRouter · parasail/bf16

    bf16
    Input / 1M
    $0.04
    Output / 1M
    $0.05
    Cache read / 1M
    —
    Context
    8K
    Your cost / mo
    $0.50
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by Parasail through OpenRouter, not a direct-provider contract. Published rates for route parasail/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    parasail/bf16 · bf16

    Declared parameters: frequency_penalty, logit_bias, logprobs, max_tokens, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_logprobs, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: sao10k/l3-lunaris-8b
    Protocol: openai
    Provider routing: Parasail
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "sao10k/l3-lunaris-8b",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "Parasail"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "sao10k/l3-lunaris-8b". Route requests to provider "Parasail" by setting the `provider.order` field to ["Parasail"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.04 / 1M input tokens, $0.05 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints.
    Price facts
    Input: $0.04 / 1M tokens
    Output: $0.05 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:44.773Z
    Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints
    Price Since Endpoints on OpenRouter
  • 12.OpenAI: gpt-oss-20b

    CoreWeave via OpenRouter · coreweave/fp4

    fp4
    Input / 1M
    $0.03
    Output / 1M
    $0.13
    Cache read / 1M
    $0.03
    Context
    131K
    Your cost / mo
    $0.56
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by CoreWeave through OpenRouter, not a direct-provider contract. Published rates for route coreweave/fp4; fp4 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    $0.03
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    coreweave/fp4 · fp4

    Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: openai/gpt-oss-20b
    Protocol: openai
    Provider routing: CoreWeave
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "openai/gpt-oss-20b",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "CoreWeave"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "CoreWeave" by setting the `provider.order` field to ["CoreWeave"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.13 / 1M output tokens, $0.03 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints.
    Price facts
    Input: $0.03 / 1M tokens
    Output: $0.13 / 1M tokens
    Cache read: $0.03 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:16.420Z
    Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints
    Price Since Endpoints on OpenRouter
  • 13.Qwen: Qwen3.7 Flash

    Alibaba via OpenRouter · alibaba

    Input / 1M
    $0.03
    Output / 1M
    $0.13
    Cache read / 1M
    $0.006
    Context
    1M
    Your cost / mo
    $0.56
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by Alibaba through OpenRouter, not a direct-provider contract. Published rates for route alibaba; unknown quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    $0.006
    Cache write / 1M
    $0.038
    Per request
    Not published
    Route
    alibaba · unknown

    Declared parameters: include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, response_format, seed, temperature, tool_choice, tools, top_logprobs, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: qwen/qwen3.7-flash
    Protocol: openai
    Provider routing: Alibaba
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "qwen/qwen3.7-flash",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "Alibaba"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "qwen/qwen3.7-flash". Route requests to provider "Alibaba" by setting the `provider.order` field to ["Alibaba"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.13 / 1M output tokens, $0.006 / 1M cache read tokens, $0.038 / 1M cache write tokens. Source: https://openrouter.ai/api/v1/models/qwen/qwen3.7-flash/endpoints.
    Price facts
    Input: $0.03 / 1M tokens
    Output: $0.13 / 1M tokens
    Cache read: $0.006 / 1M tokens
    Cache write: $0.038 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:43.500Z
    Source: https://openrouter.ai/api/v1/models/qwen/qwen3.7-flash/endpoints
    Price Since Endpoints on OpenRouter
  • 14.OpenAI: gpt-oss-20b

    DekaLLM via OpenRouter · dekallm/bf16

    bf16
    Input / 1M
    $0.029
    Output / 1M
    $0.14
    Cache read / 1M
    —
    Context
    131K
    Your cost / mo
    $0.57
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by DekaLLM through OpenRouter, not a direct-provider contract. Published rates for route dekallm/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    dekallm/bf16 · bf16

    Declared parameters: frequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_logprobs, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: openai/gpt-oss-20b
    Protocol: openai
    Provider routing: DekaLLM
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "openai/gpt-oss-20b",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "DekaLLM"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "DekaLLM" by setting the `provider.order` field to ["DekaLLM"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.029 / 1M input tokens, $0.14 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints.
    Price facts
    Input: $0.029 / 1M tokens
    Output: $0.14 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:16.420Z
    Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints
    Price Since Endpoints on OpenRouter
  • 15.OpenAI: gpt-oss-20b

    DeepInfra via OpenRouter · deepinfra/bf16

    bf16
    Input / 1M
    $0.03
    Output / 1M
    $0.14
    Cache read / 1M
    —
    Context
    131K
    Your cost / mo
    $0.58
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    deepinfra/bf16 · bf16

    Declared parameters: frequency_penalty, include_reasoning, logit_bias, max_tokens, min_p, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: openai/gpt-oss-20b
    Protocol: openai
    Provider routing: DeepInfra
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "openai/gpt-oss-20b",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "DeepInfra"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.14 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints.
    Price facts
    Input: $0.03 / 1M tokens
    Output: $0.14 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:16.420Z
    Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints
    Price Since Endpoints on OpenRouter
  • 16.Inference.net: Schematron V2 Turbo

    InferenceNet via OpenRouter · inference-net

    Input / 1M
    $0.03
    Output / 1M
    $0.15
    Cache read / 1M
    $0.03
    Context
    128K
    Your cost / mo
    $0.60
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by InferenceNet through OpenRouter, not a direct-provider contract. Published rates for route inference-net; unknown quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    $0.03
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    inference-net · unknown

    Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: inference-net/schematron-v2-turbo
    Protocol: openai
    Provider routing: InferenceNet
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "inference-net/schematron-v2-turbo",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "InferenceNet"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "inference-net/schematron-v2-turbo". Route requests to provider "InferenceNet" by setting the `provider.order` field to ["InferenceNet"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.15 / 1M output tokens, $0.03 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/inference-net/schematron-v2-turbo/endpoints.
    Price facts
    Input: $0.03 / 1M tokens
    Output: $0.15 / 1M tokens
    Cache read: $0.03 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:25.607Z
    Source: https://openrouter.ai/api/v1/models/inference-net/schematron-v2-turbo/endpoints
    Price Since Endpoints on OpenRouter
  • 17.OpenAI: gpt-oss-20b

    Parasail via OpenRouter · parasail/fp4

    fp4
    Input / 1M
    $0.03
    Output / 1M
    $0.15
    Cache read / 1M
    $0.02
    Context
    131K
    Your cost / mo
    $0.60
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by Parasail through OpenRouter, not a direct-provider contract. Published rates for route parasail/fp4; fp4 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    $0.02
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    parasail/fp4 · fp4

    Declared parameters: frequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: openai/gpt-oss-20b
    Protocol: openai
    Provider routing: Parasail
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "openai/gpt-oss-20b",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "Parasail"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "Parasail" by setting the `provider.order` field to ["Parasail"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.15 / 1M output tokens, $0.02 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints.
    Price facts
    Input: $0.03 / 1M tokens
    Output: $0.15 / 1M tokens
    Cache read: $0.02 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:16.420Z
    Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints
    Price Since Endpoints on OpenRouter
  • 18.Sao10K: Llama 3 8B Lunaris

    Novita via OpenRouter · novita/bf16

    bf16
    Input / 1M
    $0.05
    Output / 1M
    $0.05
    Cache read / 1M
    —
    Context
    8K
    Your cost / mo
    $0.60
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by Novita through OpenRouter, not a direct-provider contract. Published rates for route novita/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    novita/bf16 · bf16

    Declared parameters: frequency_penalty, max_tokens, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: sao10k/l3-lunaris-8b
    Protocol: openai
    Provider routing: Novita
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "sao10k/l3-lunaris-8b",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "Novita"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "sao10k/l3-lunaris-8b". Route requests to provider "Novita" by setting the `provider.order` field to ["Novita"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.05 / 1M input tokens, $0.05 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints.
    Price facts
    Input: $0.05 / 1M tokens
    Output: $0.05 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:44.773Z
    Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints
    Price Since Endpoints on OpenRouter
  • 19.Amazon: Nova Micro 1.0

    Amazon Bedrock via OpenRouter · amazon-bedrock

    Input / 1M
    $0.035
    Output / 1M
    $0.14
    Cache read / 1M
    —
    Context
    128K
    Your cost / mo
    $0.63
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by Amazon Bedrock through OpenRouter, not a direct-provider contract. Published rates for route amazon-bedrock; unknown quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    amazon-bedrock · unknown

    Declared parameters: max_tokens, stop, temperature, tools, top_k, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: amazon/nova-micro-v1
    Protocol: openai
    Provider routing: Amazon Bedrock
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "amazon/nova-micro-v1",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "Amazon Bedrock"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "amazon/nova-micro-v1". Route requests to provider "Amazon Bedrock" by setting the `provider.order` field to ["Amazon Bedrock"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.035 / 1M input tokens, $0.14 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/amazon/nova-micro-v1/endpoints.
    Price facts
    Input: $0.035 / 1M tokens
    Output: $0.14 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:17.502Z
    Source: https://openrouter.ai/api/v1/models/amazon/nova-micro-v1/endpoints
    Price Since Endpoints on OpenRouter
  • 20.Amazon: Nova Micro 1.0

    Amazon Bedrock via OpenRouter · amazon-bedrock/eu-west-1

    Input / 1M
    $0.035
    Output / 1M
    $0.14
    Cache read / 1M
    —
    Context
    128K
    Your cost / mo
    $0.63
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by Amazon Bedrock through OpenRouter, not a direct-provider contract. Published rates for route amazon-bedrock/eu-west-1; unknown quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    amazon-bedrock/eu-west-1 · unknown

    Declared parameters: max_tokens, stop, temperature, tools, top_k, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: amazon/nova-micro-v1
    Protocol: openai
    Provider routing: Amazon Bedrock
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "amazon/nova-micro-v1",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "Amazon Bedrock"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "amazon/nova-micro-v1". Route requests to provider "Amazon Bedrock" by setting the `provider.order` field to ["Amazon Bedrock"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.035 / 1M input tokens, $0.14 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/amazon/nova-micro-v1/endpoints.
    Price facts
    Input: $0.035 / 1M tokens
    Output: $0.14 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:17.502Z
    Source: https://openrouter.ai/api/v1/models/amazon/nova-micro-v1/endpoints
    Price Since Endpoints on OpenRouter
  • 21.OpenAI: gpt-oss-120b

    CoreWeave via OpenRouter · coreweave/fp4

    fp4
    Input / 1M
    $0.03
    Output / 1M
    $0.17
    Cache read / 1M
    $0.03
    Context
    131K
    Your cost / mo
    $0.64
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by CoreWeave through OpenRouter, not a direct-provider contract. Published rates for route coreweave/fp4; fp4 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    $0.03
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    coreweave/fp4 · fp4

    Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: openai/gpt-oss-120b
    Protocol: openai
    Provider routing: CoreWeave
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "openai/gpt-oss-120b",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "CoreWeave"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-120b". Route requests to provider "CoreWeave" by setting the `provider.order` field to ["CoreWeave"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.17 / 1M output tokens, $0.03 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints.
    Price facts
    Input: $0.03 / 1M tokens
    Output: $0.17 / 1M tokens
    Cache read: $0.03 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:16.695Z
    Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints
    Price Since Endpoints on OpenRouter
  • 22.OpenAI: GPT-5 Nano

    OpenAI via OpenRouter · openai/flex

    Input / 1M
    $0.025
    Output / 1M
    $0.20
    Cache read / 1M
    $0.0025
    Context
    400K
    Your cost / mo
    $0.65
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by OpenAI through OpenRouter, not a direct-provider contract. Published rates for route openai/flex; unknown quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    $0.0025
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    openai/flex · unknown

    Declared parameters: include_reasoning, max_tokens, reasoning, reasoning_effort, response_format, seed, structured_outputs, tool_choice, tools

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: openai/gpt-5-nano
    Protocol: openai
    Provider routing: OpenAI
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "openai/gpt-5-nano",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "OpenAI"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-5-nano". Route requests to provider "OpenAI" by setting the `provider.order` field to ["OpenAI"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.025 / 1M input tokens, $0.20 / 1M output tokens, $0.0025 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-5-nano/endpoints.
    Price facts
    Input: $0.025 / 1M tokens
    Output: $0.20 / 1M tokens
    Cache read: $0.0025 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:33.354Z
    Source: https://openrouter.ai/api/v1/models/openai/gpt-5-nano/endpoints
    Price Since Endpoints on OpenRouter
  • 23.Meta: Llama 3.1 8B Instruct

    Groq via OpenRouter · groq

    Input / 1M
    $0.05
    Output / 1M
    $0.08
    Cache read / 1M
    $0.025
    Context
    131K
    Your cost / mo
    $0.66
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by Groq through OpenRouter, not a direct-provider contract. Published rates for route groq; unknown quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    $0.025
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    groq · unknown

    Declared parameters: max_tokens, response_format, seed, stop, temperature, tool_choice, tools, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: meta-llama/llama-3.1-8b-instruct
    Protocol: openai
    Provider routing: Groq
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "meta-llama/llama-3.1-8b-instruct",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "Groq"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "meta-llama/llama-3.1-8b-instruct". Route requests to provider "Groq" by setting the `provider.order` field to ["Groq"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.05 / 1M input tokens, $0.08 / 1M output tokens, $0.025 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints.
    Price facts
    Input: $0.05 / 1M tokens
    Output: $0.08 / 1M tokens
    Cache read: $0.025 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:26.014Z
    Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints
    Price Since Endpoints on OpenRouter
  • 24.Mistral: Mistral Small 3

    DeepInfra via OpenRouter · deepinfra/fp8

    fp8
    Input / 1M
    $0.05
    Output / 1M
    $0.08
    Cache read / 1M
    —
    Context
    33K
    Your cost / mo
    $0.66
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    deepinfra/fp8 · fp8

    Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: mistralai/mistral-small-24b-instruct-2501
    Protocol: openai
    Provider routing: DeepInfra
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "mistralai/mistral-small-24b-instruct-2501",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "DeepInfra"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "mistralai/mistral-small-24b-instruct-2501". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.05 / 1M input tokens, $0.08 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/mistralai/mistral-small-24b-instruct-2501/endpoints.
    Price facts
    Input: $0.05 / 1M tokens
    Output: $0.08 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:28.946Z
    Source: https://openrouter.ai/api/v1/models/mistralai/mistral-small-24b-instruct-2501/endpoints
    Price Since Endpoints on OpenRouter
  • 25.OpenAI: gpt-oss-120b

    DekaLLM via OpenRouter · dekallm/bf16

    bf16
    Input / 1M
    $0.03
    Output / 1M
    $0.18
    Cache read / 1M
    —
    Context
    131K
    Your cost / mo
    $0.66
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by DekaLLM through OpenRouter, not a direct-provider contract. Published rates for route dekallm/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    dekallm/bf16 · bf16

    Declared parameters: frequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_logprobs, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: openai/gpt-oss-120b
    Protocol: openai
    Provider routing: DekaLLM
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "openai/gpt-oss-120b",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "DekaLLM"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-120b". Route requests to provider "DekaLLM" by setting the `provider.order` field to ["DekaLLM"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.18 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints.
    Price facts
    Input: $0.03 / 1M tokens
    Output: $0.18 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:16.695Z
    Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints
    Price Since Endpoints on OpenRouter
  • 26.Meta: Llama 3.2 1B Instruct

    Cloudflare via OpenRouter · cloudflare

    Input / 1M
    $0.027
    Output / 1M
    $0.201
    Cache read / 1M
    —
    Context
    60K
    Your cost / mo
    $0.672
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by Cloudflare through OpenRouter, not a direct-provider contract. Published rates for route cloudflare; unknown quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    cloudflare · unknown

    Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, seed, stop, temperature, top_k, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: meta-llama/llama-3.2-1b-instruct
    Protocol: openai
    Provider routing: Cloudflare
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "meta-llama/llama-3.2-1b-instruct",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "Cloudflare"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "meta-llama/llama-3.2-1b-instruct". Route requests to provider "Cloudflare" by setting the `provider.order` field to ["Cloudflare"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.027 / 1M input tokens, $0.201 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.2-1b-instruct/endpoints.
    Price facts
    Input: $0.027 / 1M tokens
    Output: $0.201 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:26.127Z
    Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.2-1b-instruct/endpoints
    Price Since Endpoints on OpenRouter
  • 27.Cohere: Command R7B (12-2024)

    Cohere via OpenRouter · cohere

    Input / 1M
    $0.0375
    Output / 1M
    $0.15
    Cache read / 1M
    —
    Context
    128K
    Your cost / mo
    $0.675
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by Cohere through OpenRouter, not a direct-provider contract. Published rates for route cohere; unknown quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    cohere · unknown

    Declared parameters: frequency_penalty, max_tokens, presence_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: cohere/command-r7b-12-2024
    Protocol: openai
    Provider routing: Cohere
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "cohere/command-r7b-12-2024",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "Cohere"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "cohere/command-r7b-12-2024". Route requests to provider "Cohere" by setting the `provider.order` field to ["Cohere"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.0375 / 1M input tokens, $0.15 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/cohere/command-r7b-12-2024/endpoints.
    Price facts
    Input: $0.0375 / 1M tokens
    Output: $0.15 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:21.103Z
    Source: https://openrouter.ai/api/v1/models/cohere/command-r7b-12-2024/endpoints
    Price Since Endpoints on OpenRouter
  • 28.Google: Gemma 3 4B

    DeepInfra via OpenRouter · deepinfra/bf16

    bf16
    Input / 1M
    $0.05
    Output / 1M
    $0.10
    Cache read / 1M
    —
    Context
    131K
    Your cost / mo
    $0.70
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    deepinfra/bf16 · bf16

    Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: google/gemma-3-4b-it
    Protocol: openai
    Provider routing: DeepInfra
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "google/gemma-3-4b-it",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "DeepInfra"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "google/gemma-3-4b-it". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.05 / 1M input tokens, $0.10 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/google/gemma-3-4b-it/endpoints.
    Price facts
    Input: $0.05 / 1M tokens
    Output: $0.10 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:24.464Z
    Source: https://openrouter.ai/api/v1/models/google/gemma-3-4b-it/endpoints
    Price Since Endpoints on OpenRouter
  • 29.OpenAI: gpt-oss-20b

    Novita via OpenRouter · novita/fp4

    fp4
    Input / 1M
    $0.04
    Output / 1M
    $0.15
    Cache read / 1M
    —
    Context
    131K
    Your cost / mo
    $0.70
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by Novita through OpenRouter, not a direct-provider contract. Published rates for route novita/fp4; fp4 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    Not published
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    novita/fp4 · fp4

    Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_logprobs, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: openai/gpt-oss-20b
    Protocol: openai
    Provider routing: Novita
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "openai/gpt-oss-20b",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "Novita"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "Novita" by setting the `provider.order` field to ["Novita"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.04 / 1M input tokens, $0.15 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints.
    Price facts
    Input: $0.04 / 1M tokens
    Output: $0.15 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:16.420Z
    Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints
    Price Since Endpoints on OpenRouter
  • 30.OpenAI: gpt-oss-120b

    AkashML via OpenRouter · akashml/bf16

    bf16
    Input / 1M
    $0.033
    Output / 1M
    $0.187
    Cache read / 1M
    $0.033
    Context
    131K
    Your cost / mo
    $0.704
    vs cheapest
    —
    Uptime 1d
    n/a
    Billing conditions

    Served by AkashML through OpenRouter, not a direct-provider contract. Published rates for route akashml/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.

    Cache read / 1M
    $0.033
    Cache write / 1M
    Not published
    Per request
    Not published
    Route
    akashml/bf16 · bf16

    Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p

    Use this endpoint
    Endpoint
    Base URL: https://openrouter.ai/api/v1
    Model: openai/gpt-oss-120b
    Protocol: openai
    Provider routing: AkashML
    API key env: OPENROUTER_API_KEY
    Docs: https://openrouter.ai/docs
    curl
    curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
      "model": "openai/gpt-oss-120b",
      "messages": [
        {
          "role": "user",
          "content": "Hello"
        }
      ],
      "provider": {
        "order": [
          "AkashML"
        ],
        "allow_fallbacks": false
      }
    }'
    Prompt for your coding agent
    Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-120b". Route requests to provider "AkashML" by setting the `provider.order` field to ["AkashML"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.033 / 1M input tokens, $0.187 / 1M output tokens, $0.033 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints.
    Price facts
    Input: $0.033 / 1M tokens
    Output: $0.187 / 1M tokens
    Cache read: $0.033 / 1M tokens
    Currency: USD
    Regions: global
    Unit source: api
    Checked: 2026-09-29T18:01:16.695Z
    Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints
    Price Since Endpoints on OpenRouter

Range = your assumptions’ low and high ends; retries from 1-day uptime.

Prices as of 2026-09-29 · tracked since 2026-09-29 · 8 checks. 4 model checks failed; earlier prices keep their timestamps. History & source status

Tracked models
521
Exact model IDs, compared per provider
Inference providers
83
+ 159 direct quotes verified by hand
Active endpoint quotes
1348
54 unavailable kept for reference
Latest price change
2026-09-29
4 model checks failed

How we compare

What “cheapest” means here

The lowest estimate among the quotes currently listed in one currency, for the volumes and tier you set. Standard, batch, off-peak, free and promotional pricing never share a ranking, and currencies are never converted. Model quality, rate limits and regional access are separate questions.

Every price keeps its source and currency

Provider endpoints keep paired input and output rates, route tags and quantization. Direct quotes keep the currency and region their provider bills in. Snapshots record when we observed a price change, not when a provider made it effective.

What is not covered yet

Most official direct price pages, regional and context-tier pricing, embeddings, audio, image and video billing, free-tier allowances and measured latency. Unknown values are shown as unknown and never counted as zero.

Models by providerProvider coveragePrice history & source statusSnapshot JSON