Model API prices & token cost calculator
Find a model
→ 1043 quotes match
Real cost
Compare below
Estimate, not a bill: token and request charges only, in the quote’s own currency. Excludes tools, storage, search, taxes and platform fees. Different models tokenize the same task differently.
Shortlist 0/3
Add quotes from the results with “+ Shortlist”, or let us pick three.
Today
Paired input/output rates per provider endpoint for tracked models. Purchased through OpenRouter; endpoints need review after 24 hours.
| # | Model · provider | Input / 1M | Output / 1M | Cache read / 1M | Context | Your cost / mo | vs cheapest | Uptime 1d | Price since · source | Shortlist |
|---|---|---|---|---|---|---|---|---|---|---|
| 1 | 1.Mistral: Mistral Nemo DekaLLM via OpenRouter · dekallm/fp8 fp8Billing conditionsServed by DekaLLM through OpenRouter, not a direct-provider contract. Published rates for route dekallm/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, logit_bias, logprobs, max_tokens, presence_penalty, response_format, seed, stop, structured_outputs, temperature, top_logprobs, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: mistralai/mistral-nemo Protocol: openai Provider routing: DekaLLM API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "mistralai/mistral-nemo",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"DekaLLM"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "mistralai/mistral-nemo". Route requests to provider "DekaLLM" by setting the `provider.order` field to ["DekaLLM"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.018 / 1M input tokens, $0.03 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints. Price facts Input: $0.018 / 1M tokens Output: $0.03 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:28.816Z Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints | $0.018 | $0.03 | — | 131K | $0.24 | — | n/a | Since | |
| 2 | 2.Mistral: Mistral Nemo DeepInfra via OpenRouter · deepinfra/fp8 fp8Billing conditionsServed by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: mistralai/mistral-nemo Protocol: openai Provider routing: DeepInfra API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "mistralai/mistral-nemo",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"DeepInfra"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "mistralai/mistral-nemo". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.019 / 1M input tokens, $0.03 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints. Price facts Input: $0.019 / 1M tokens Output: $0.03 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:28.816Z Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints | $0.019 | $0.03 | — | 131K | $0.25 | — | n/a | Since | |
| 3 | 3.Meta: Llama 3.1 8B Instruct DeepInfra via OpenRouter · deepinfra/fp8 fp8Billing conditionsServed by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, response_format, seed, stop, temperature, top_k, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: meta-llama/llama-3.1-8b-instruct Protocol: openai Provider routing: DeepInfra API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "meta-llama/llama-3.1-8b-instruct",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"DeepInfra"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "meta-llama/llama-3.1-8b-instruct". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.02 / 1M input tokens, $0.04 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints. Price facts Input: $0.02 / 1M tokens Output: $0.04 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:26.014Z Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints | $0.02 | $0.04 | — | 131K | $0.28 | — | n/a | Since | |
| 4 | 4.Meta: Llama 3.1 8B Instruct Novita via OpenRouter · novita/fp8 fp8Billing conditionsServed by Novita through OpenRouter, not a direct-provider contract. Published rates for route novita/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, logprobs, max_tokens, presence_penalty, repetition_penalty, response_format, seed, stop, temperature, top_k, top_logprobs, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: meta-llama/llama-3.1-8b-instruct Protocol: openai Provider routing: Novita API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "meta-llama/llama-3.1-8b-instruct",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"Novita"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "meta-llama/llama-3.1-8b-instruct". Route requests to provider "Novita" by setting the `provider.order` field to ["Novita"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.02 / 1M input tokens, $0.05 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints. Price facts Input: $0.02 / 1M tokens Output: $0.05 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:26.014Z Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints | $0.02 | $0.05 | — | 16K | $0.30 | — | n/a | Since | |
| 5 | 5.Mistral: Mistral Nemo Parasail via OpenRouter · parasail/fp8 fp8Billing conditionsServed by Parasail through OpenRouter, not a direct-provider contract. Published rates for route parasail/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, logit_bias, logprobs, max_tokens, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_logprobs, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: mistralai/mistral-nemo Protocol: openai Provider routing: Parasail API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "mistralai/mistral-nemo",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"Parasail"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "mistralai/mistral-nemo". Route requests to provider "Parasail" by setting the `provider.order` field to ["Parasail"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.03 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints. Price facts Input: $0.03 / 1M tokens Output: $0.03 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:28.816Z Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints | $0.03 | $0.03 | — | 131K | $0.36 | — | n/a | Since | |
| 6 | 6.OpenAI: gpt-oss-20b Darkbloom via OpenRouter · darkbloom/fp8 fp8Billing conditionsServed by Darkbloom through OpenRouter, not a direct-provider contract. Published rates for route darkbloom/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-20b Protocol: openai Provider routing: Darkbloom API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "openai/gpt-oss-20b",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"Darkbloom"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "Darkbloom" by setting the `provider.order` field to ["Darkbloom"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.018 / 1M input tokens, $0.09 / 1M output tokens, $0.009 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints. Price facts Input: $0.018 / 1M tokens Output: $0.09 / 1M tokens Cache read: $0.009 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.420Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints | $0.018 | $0.09 | $0.009 | 131K | $0.36 | — | n/a | Since | |
| 7 | 7.IBM: Granite 4.0 Micro Cloudflare via OpenRouter · cloudflare Billing conditionsServed by Cloudflare through OpenRouter, not a direct-provider contract. Published rates for route cloudflare; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, logit_bias, logprobs, max_tokens, min_p, presence_penalty, repetition_penalty, response_format, seed, stop, temperature, top_k, top_logprobs, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: ibm-granite/granite-4.0-h-micro Protocol: openai Provider routing: Cloudflare API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "ibm-granite/granite-4.0-h-micro",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"Cloudflare"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "ibm-granite/granite-4.0-h-micro". Route requests to provider "Cloudflare" by setting the `provider.order` field to ["Cloudflare"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.017 / 1M input tokens, $0.112 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/ibm-granite/granite-4.0-h-micro/endpoints. Price facts Input: $0.017 / 1M tokens Output: $0.112 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:24.978Z Source: https://openrouter.ai/api/v1/models/ibm-granite/granite-4.0-h-micro/endpoints | $0.017 | $0.112 | — | 131K | $0.394 | — | n/a | Since | |
| 8 | 8.OpenAI: gpt-oss-20b AkashML via OpenRouter · akashml/fp4 fp4Billing conditionsServed by AkashML through OpenRouter, not a direct-provider contract. Published rates for route akashml/fp4; fp4 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-20b Protocol: openai Provider routing: AkashML API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "openai/gpt-oss-20b",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"AkashML"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "AkashML" by setting the `provider.order` field to ["AkashML"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.02 / 1M input tokens, $0.10 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints. Price facts Input: $0.02 / 1M tokens Output: $0.10 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.420Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints | $0.02 | $0.10 | — | 131K | $0.40 | — | n/a | Since | |
| 9 | 9.Nex AGI: Nex-N2.5-Mini Nex AGI via OpenRouter · nex-agi/bf16 bf16Billing conditionsServed by Nex AGI through OpenRouter, not a direct-provider contract. Published rates for route nex-agi/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: include_reasoning, logprobs, max_tokens, reasoning, reasoning_effort, structured_outputs, temperature, top_k, top_logprobs, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: nex-agi/nex-n2.5-mini Protocol: openai Provider routing: Nex AGI API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "nex-agi/nex-n2.5-mini",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"Nex AGI"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "nex-agi/nex-n2.5-mini". Route requests to provider "Nex AGI" by setting the `provider.order` field to ["Nex AGI"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.025 / 1M input tokens, $0.10 / 1M output tokens, $0.0025 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/nex-agi/nex-n2.5-mini/endpoints. Price facts Input: $0.025 / 1M tokens Output: $0.10 / 1M tokens Cache read: $0.0025 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:30.211Z Source: https://openrouter.ai/api/v1/models/nex-agi/nex-n2.5-mini/endpoints | $0.025 | $0.10 | $0.0025 | 262K | $0.45 | — | n/a | Since | |
| 10 | 10.Sao10K: Llama 3 8B Lunaris DeepInfra via OpenRouter · deepinfra/turbo fp8Billing conditionsServed by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/turbo; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, seed, stop, temperature, top_k, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: sao10k/l3-lunaris-8b Protocol: openai Provider routing: DeepInfra API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "sao10k/l3-lunaris-8b",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"DeepInfra"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "sao10k/l3-lunaris-8b". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.04 / 1M input tokens, $0.05 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints. Price facts Input: $0.04 / 1M tokens Output: $0.05 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:44.773Z Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints | $0.04 | $0.05 | — | 8K | $0.50 | — | n/a | Since | |
| 11 | 11.Sao10K: Llama 3 8B Lunaris Parasail via OpenRouter · parasail/bf16 bf16Billing conditionsServed by Parasail through OpenRouter, not a direct-provider contract. Published rates for route parasail/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, logit_bias, logprobs, max_tokens, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_logprobs, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: sao10k/l3-lunaris-8b Protocol: openai Provider routing: Parasail API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "sao10k/l3-lunaris-8b",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"Parasail"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "sao10k/l3-lunaris-8b". Route requests to provider "Parasail" by setting the `provider.order` field to ["Parasail"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.04 / 1M input tokens, $0.05 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints. Price facts Input: $0.04 / 1M tokens Output: $0.05 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:44.773Z Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints | $0.04 | $0.05 | — | 8K | $0.50 | — | n/a | Since | |
| 12 | 12.OpenAI: gpt-oss-20b CoreWeave via OpenRouter · coreweave/fp4 fp4Billing conditionsServed by CoreWeave through OpenRouter, not a direct-provider contract. Published rates for route coreweave/fp4; fp4 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-20b Protocol: openai Provider routing: CoreWeave API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "openai/gpt-oss-20b",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"CoreWeave"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "CoreWeave" by setting the `provider.order` field to ["CoreWeave"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.13 / 1M output tokens, $0.03 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints. Price facts Input: $0.03 / 1M tokens Output: $0.13 / 1M tokens Cache read: $0.03 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.420Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints | $0.03 | $0.13 | $0.03 | 131K | $0.56 | — | n/a | Since | |
| 13 | 13.Qwen: Qwen3.7 Flash Alibaba via OpenRouter · alibaba Billing conditionsServed by Alibaba through OpenRouter, not a direct-provider contract. Published rates for route alibaba; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, response_format, seed, temperature, tool_choice, tools, top_logprobs, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: qwen/qwen3.7-flash Protocol: openai Provider routing: Alibaba API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "qwen/qwen3.7-flash",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"Alibaba"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "qwen/qwen3.7-flash". Route requests to provider "Alibaba" by setting the `provider.order` field to ["Alibaba"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.13 / 1M output tokens, $0.006 / 1M cache read tokens, $0.038 / 1M cache write tokens. Source: https://openrouter.ai/api/v1/models/qwen/qwen3.7-flash/endpoints. Price facts Input: $0.03 / 1M tokens Output: $0.13 / 1M tokens Cache read: $0.006 / 1M tokens Cache write: $0.038 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:43.500Z Source: https://openrouter.ai/api/v1/models/qwen/qwen3.7-flash/endpoints | $0.03 | $0.13 | $0.006 | 1M | $0.56 | — | n/a | Since | |
| 14 | 14.OpenAI: gpt-oss-20b DekaLLM via OpenRouter · dekallm/bf16 bf16Billing conditionsServed by DekaLLM through OpenRouter, not a direct-provider contract. Published rates for route dekallm/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_logprobs, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-20b Protocol: openai Provider routing: DekaLLM API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "openai/gpt-oss-20b",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"DekaLLM"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "DekaLLM" by setting the `provider.order` field to ["DekaLLM"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.029 / 1M input tokens, $0.14 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints. Price facts Input: $0.029 / 1M tokens Output: $0.14 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.420Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints | $0.029 | $0.14 | — | 131K | $0.57 | — | n/a | Since | |
| 15 | 15.OpenAI: gpt-oss-20b DeepInfra via OpenRouter · deepinfra/bf16 bf16Billing conditionsServed by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, include_reasoning, logit_bias, max_tokens, min_p, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-20b Protocol: openai Provider routing: DeepInfra API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "openai/gpt-oss-20b",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"DeepInfra"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.14 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints. Price facts Input: $0.03 / 1M tokens Output: $0.14 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.420Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints | $0.03 | $0.14 | — | 131K | $0.58 | — | n/a | Since | |
| 16 | 16.Inference.net: Schematron V2 Turbo InferenceNet via OpenRouter · inference-net Billing conditionsServed by InferenceNet through OpenRouter, not a direct-provider contract. Published rates for route inference-net; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: inference-net/schematron-v2-turbo Protocol: openai Provider routing: InferenceNet API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "inference-net/schematron-v2-turbo",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"InferenceNet"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "inference-net/schematron-v2-turbo". Route requests to provider "InferenceNet" by setting the `provider.order` field to ["InferenceNet"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.15 / 1M output tokens, $0.03 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/inference-net/schematron-v2-turbo/endpoints. Price facts Input: $0.03 / 1M tokens Output: $0.15 / 1M tokens Cache read: $0.03 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:25.607Z Source: https://openrouter.ai/api/v1/models/inference-net/schematron-v2-turbo/endpoints | $0.03 | $0.15 | $0.03 | 128K | $0.60 | — | n/a | Since | |
| 17 | 17.OpenAI: gpt-oss-20b Parasail via OpenRouter · parasail/fp4 fp4Billing conditionsServed by Parasail through OpenRouter, not a direct-provider contract. Published rates for route parasail/fp4; fp4 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-20b Protocol: openai Provider routing: Parasail API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "openai/gpt-oss-20b",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"Parasail"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "Parasail" by setting the `provider.order` field to ["Parasail"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.15 / 1M output tokens, $0.02 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints. Price facts Input: $0.03 / 1M tokens Output: $0.15 / 1M tokens Cache read: $0.02 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.420Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints | $0.03 | $0.15 | $0.02 | 131K | $0.60 | — | n/a | Since | |
| 18 | 18.Sao10K: Llama 3 8B Lunaris Novita via OpenRouter · novita/bf16 bf16Billing conditionsServed by Novita through OpenRouter, not a direct-provider contract. Published rates for route novita/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, max_tokens, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: sao10k/l3-lunaris-8b Protocol: openai Provider routing: Novita API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "sao10k/l3-lunaris-8b",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"Novita"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "sao10k/l3-lunaris-8b". Route requests to provider "Novita" by setting the `provider.order` field to ["Novita"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.05 / 1M input tokens, $0.05 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints. Price facts Input: $0.05 / 1M tokens Output: $0.05 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:44.773Z Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints | $0.05 | $0.05 | — | 8K | $0.60 | — | n/a | Since | |
| 19 | 19.Amazon: Nova Micro 1.0 Amazon Bedrock via OpenRouter · amazon-bedrock Billing conditionsServed by Amazon Bedrock through OpenRouter, not a direct-provider contract. Published rates for route amazon-bedrock; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: max_tokens, stop, temperature, tools, top_k, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: amazon/nova-micro-v1 Protocol: openai Provider routing: Amazon Bedrock API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "amazon/nova-micro-v1",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"Amazon Bedrock"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "amazon/nova-micro-v1". Route requests to provider "Amazon Bedrock" by setting the `provider.order` field to ["Amazon Bedrock"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.035 / 1M input tokens, $0.14 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/amazon/nova-micro-v1/endpoints. Price facts Input: $0.035 / 1M tokens Output: $0.14 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:17.502Z Source: https://openrouter.ai/api/v1/models/amazon/nova-micro-v1/endpoints | $0.035 | $0.14 | — | 128K | $0.63 | — | n/a | Since | |
| 20 | 20.Amazon: Nova Micro 1.0 Amazon Bedrock via OpenRouter · amazon-bedrock/eu-west-1 Billing conditionsServed by Amazon Bedrock through OpenRouter, not a direct-provider contract. Published rates for route amazon-bedrock/eu-west-1; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: max_tokens, stop, temperature, tools, top_k, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: amazon/nova-micro-v1 Protocol: openai Provider routing: Amazon Bedrock API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "amazon/nova-micro-v1",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"Amazon Bedrock"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "amazon/nova-micro-v1". Route requests to provider "Amazon Bedrock" by setting the `provider.order` field to ["Amazon Bedrock"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.035 / 1M input tokens, $0.14 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/amazon/nova-micro-v1/endpoints. Price facts Input: $0.035 / 1M tokens Output: $0.14 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:17.502Z Source: https://openrouter.ai/api/v1/models/amazon/nova-micro-v1/endpoints | $0.035 | $0.14 | — | 128K | $0.63 | — | n/a | Since | |
| 21 | 21.OpenAI: gpt-oss-120b CoreWeave via OpenRouter · coreweave/fp4 fp4Billing conditionsServed by CoreWeave through OpenRouter, not a direct-provider contract. Published rates for route coreweave/fp4; fp4 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-120b Protocol: openai Provider routing: CoreWeave API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "openai/gpt-oss-120b",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"CoreWeave"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-120b". Route requests to provider "CoreWeave" by setting the `provider.order` field to ["CoreWeave"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.17 / 1M output tokens, $0.03 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints. Price facts Input: $0.03 / 1M tokens Output: $0.17 / 1M tokens Cache read: $0.03 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.695Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints | $0.03 | $0.17 | $0.03 | 131K | $0.64 | — | n/a | Since | |
| 22 | 22.OpenAI: GPT-5 Nano OpenAI via OpenRouter · openai/flex Billing conditionsServed by OpenAI through OpenRouter, not a direct-provider contract. Published rates for route openai/flex; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: include_reasoning, max_tokens, reasoning, reasoning_effort, response_format, seed, structured_outputs, tool_choice, tools Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: openai/gpt-5-nano Protocol: openai Provider routing: OpenAI API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "openai/gpt-5-nano",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"OpenAI"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-5-nano". Route requests to provider "OpenAI" by setting the `provider.order` field to ["OpenAI"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.025 / 1M input tokens, $0.20 / 1M output tokens, $0.0025 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-5-nano/endpoints. Price facts Input: $0.025 / 1M tokens Output: $0.20 / 1M tokens Cache read: $0.0025 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:33.354Z Source: https://openrouter.ai/api/v1/models/openai/gpt-5-nano/endpoints | $0.025 | $0.20 | $0.0025 | 400K | $0.65 | — | n/a | Since | |
| 23 | 23.Meta: Llama 3.1 8B Instruct Groq via OpenRouter · groq Billing conditionsServed by Groq through OpenRouter, not a direct-provider contract. Published rates for route groq; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: max_tokens, response_format, seed, stop, temperature, tool_choice, tools, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: meta-llama/llama-3.1-8b-instruct Protocol: openai Provider routing: Groq API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "meta-llama/llama-3.1-8b-instruct",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"Groq"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "meta-llama/llama-3.1-8b-instruct". Route requests to provider "Groq" by setting the `provider.order` field to ["Groq"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.05 / 1M input tokens, $0.08 / 1M output tokens, $0.025 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints. Price facts Input: $0.05 / 1M tokens Output: $0.08 / 1M tokens Cache read: $0.025 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:26.014Z Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints | $0.05 | $0.08 | $0.025 | 131K | $0.66 | — | n/a | Since | |
| 24 | 24.Mistral: Mistral Small 3 DeepInfra via OpenRouter · deepinfra/fp8 fp8Billing conditionsServed by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: mistralai/mistral-small-24b-instruct-2501 Protocol: openai Provider routing: DeepInfra API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "mistralai/mistral-small-24b-instruct-2501",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"DeepInfra"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "mistralai/mistral-small-24b-instruct-2501". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.05 / 1M input tokens, $0.08 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/mistralai/mistral-small-24b-instruct-2501/endpoints. Price facts Input: $0.05 / 1M tokens Output: $0.08 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:28.946Z Source: https://openrouter.ai/api/v1/models/mistralai/mistral-small-24b-instruct-2501/endpoints | $0.05 | $0.08 | — | 33K | $0.66 | — | n/a | Since | |
| 25 | 25.OpenAI: gpt-oss-120b DekaLLM via OpenRouter · dekallm/bf16 bf16Billing conditionsServed by DekaLLM through OpenRouter, not a direct-provider contract. Published rates for route dekallm/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_logprobs, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-120b Protocol: openai Provider routing: DekaLLM API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "openai/gpt-oss-120b",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"DekaLLM"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-120b". Route requests to provider "DekaLLM" by setting the `provider.order` field to ["DekaLLM"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.18 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints. Price facts Input: $0.03 / 1M tokens Output: $0.18 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.695Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints | $0.03 | $0.18 | — | 131K | $0.66 | — | n/a | Since | |
| 26 | 26.Meta: Llama 3.2 1B Instruct Cloudflare via OpenRouter · cloudflare Billing conditionsServed by Cloudflare through OpenRouter, not a direct-provider contract. Published rates for route cloudflare; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, seed, stop, temperature, top_k, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: meta-llama/llama-3.2-1b-instruct Protocol: openai Provider routing: Cloudflare API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "meta-llama/llama-3.2-1b-instruct",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"Cloudflare"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "meta-llama/llama-3.2-1b-instruct". Route requests to provider "Cloudflare" by setting the `provider.order` field to ["Cloudflare"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.027 / 1M input tokens, $0.201 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.2-1b-instruct/endpoints. Price facts Input: $0.027 / 1M tokens Output: $0.201 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:26.127Z Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.2-1b-instruct/endpoints | $0.027 | $0.201 | — | 60K | $0.672 | — | n/a | Since | |
| 27 | 27.Cohere: Command R7B (12-2024) Cohere via OpenRouter · cohere Billing conditionsServed by Cohere through OpenRouter, not a direct-provider contract. Published rates for route cohere; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, max_tokens, presence_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: cohere/command-r7b-12-2024 Protocol: openai Provider routing: Cohere API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "cohere/command-r7b-12-2024",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"Cohere"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "cohere/command-r7b-12-2024". Route requests to provider "Cohere" by setting the `provider.order` field to ["Cohere"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.0375 / 1M input tokens, $0.15 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/cohere/command-r7b-12-2024/endpoints. Price facts Input: $0.0375 / 1M tokens Output: $0.15 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:21.103Z Source: https://openrouter.ai/api/v1/models/cohere/command-r7b-12-2024/endpoints | $0.0375 | $0.15 | — | 128K | $0.675 | — | n/a | Since | |
| 28 | 28.Google: Gemma 3 4B DeepInfra via OpenRouter · deepinfra/bf16 bf16Billing conditionsServed by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: google/gemma-3-4b-it Protocol: openai Provider routing: DeepInfra API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "google/gemma-3-4b-it",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"DeepInfra"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "google/gemma-3-4b-it". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.05 / 1M input tokens, $0.10 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/google/gemma-3-4b-it/endpoints. Price facts Input: $0.05 / 1M tokens Output: $0.10 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:24.464Z Source: https://openrouter.ai/api/v1/models/google/gemma-3-4b-it/endpoints | $0.05 | $0.10 | — | 131K | $0.70 | — | n/a | Since | |
| 29 | 29.OpenAI: gpt-oss-20b Novita via OpenRouter · novita/fp4 fp4Billing conditionsServed by Novita through OpenRouter, not a direct-provider contract. Published rates for route novita/fp4; fp4 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_logprobs, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-20b Protocol: openai Provider routing: Novita API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "openai/gpt-oss-20b",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"Novita"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "Novita" by setting the `provider.order` field to ["Novita"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.04 / 1M input tokens, $0.15 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints. Price facts Input: $0.04 / 1M tokens Output: $0.15 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.420Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints | $0.04 | $0.15 | — | 131K | $0.70 | — | n/a | Since | |
| 30 | 30.OpenAI: gpt-oss-120b AkashML via OpenRouter · akashml/bf16 bf16Billing conditionsServed by AkashML through OpenRouter, not a direct-provider contract. Published rates for route akashml/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.
Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p Use this endpointEndpoint Base URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-120b Protocol: openai Provider routing: AkashML API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs curl curl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{
"model": "openai/gpt-oss-120b",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"provider": {
"order": [
"AkashML"
],
"allow_fallbacks": false
}
}'Prompt for your coding agent Use the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-120b". Route requests to provider "AkashML" by setting the `provider.order` field to ["AkashML"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.033 / 1M input tokens, $0.187 / 1M output tokens, $0.033 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints. Price facts Input: $0.033 / 1M tokens Output: $0.187 / 1M tokens Cache read: $0.033 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.695Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints | $0.033 | $0.187 | $0.033 | 131K | $0.704 | — | n/a | Since |
1.Mistral: Mistral Nemo
DekaLLM via OpenRouter · dekallm/fp8
fp8- Input / 1M
- $0.018
- Output / 1M
- $0.03
- Cache read / 1M
- —
- Context
- 131K
- Your cost / mo
- $0.24
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by DekaLLM through OpenRouter, not a direct-provider contract. Published rates for route dekallm/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- dekallm/fp8 · fp8
Declared parameters: frequency_penalty, logit_bias, logprobs, max_tokens, presence_penalty, response_format, seed, stop, structured_outputs, temperature, top_logprobs, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: mistralai/mistral-nemo Protocol: openai Provider routing: DekaLLM API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "mistralai/mistral-nemo", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "DekaLLM" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "mistralai/mistral-nemo". Route requests to provider "DekaLLM" by setting the `provider.order` field to ["DekaLLM"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.018 / 1M input tokens, $0.03 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints.
Price factsInput: $0.018 / 1M tokens Output: $0.03 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:28.816Z Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints
Price Since Endpoints on OpenRouter2.Mistral: Mistral Nemo
DeepInfra via OpenRouter · deepinfra/fp8
fp8- Input / 1M
- $0.019
- Output / 1M
- $0.03
- Cache read / 1M
- —
- Context
- 131K
- Your cost / mo
- $0.25
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- deepinfra/fp8 · fp8
Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: mistralai/mistral-nemo Protocol: openai Provider routing: DeepInfra API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "mistralai/mistral-nemo", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "DeepInfra" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "mistralai/mistral-nemo". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.019 / 1M input tokens, $0.03 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints.
Price factsInput: $0.019 / 1M tokens Output: $0.03 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:28.816Z Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints
Price Since Endpoints on OpenRouter3.Meta: Llama 3.1 8B Instruct
DeepInfra via OpenRouter · deepinfra/fp8
fp8- Input / 1M
- $0.02
- Output / 1M
- $0.04
- Cache read / 1M
- —
- Context
- 131K
- Your cost / mo
- $0.28
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- deepinfra/fp8 · fp8
Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, response_format, seed, stop, temperature, top_k, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: meta-llama/llama-3.1-8b-instruct Protocol: openai Provider routing: DeepInfra API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "meta-llama/llama-3.1-8b-instruct", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "DeepInfra" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "meta-llama/llama-3.1-8b-instruct". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.02 / 1M input tokens, $0.04 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints.
Price factsInput: $0.02 / 1M tokens Output: $0.04 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:26.014Z Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints
Price Since Endpoints on OpenRouter4.Meta: Llama 3.1 8B Instruct
Novita via OpenRouter · novita/fp8
fp8- Input / 1M
- $0.02
- Output / 1M
- $0.05
- Cache read / 1M
- —
- Context
- 16K
- Your cost / mo
- $0.30
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by Novita through OpenRouter, not a direct-provider contract. Published rates for route novita/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- novita/fp8 · fp8
Declared parameters: frequency_penalty, logprobs, max_tokens, presence_penalty, repetition_penalty, response_format, seed, stop, temperature, top_k, top_logprobs, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: meta-llama/llama-3.1-8b-instruct Protocol: openai Provider routing: Novita API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "meta-llama/llama-3.1-8b-instruct", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "Novita" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "meta-llama/llama-3.1-8b-instruct". Route requests to provider "Novita" by setting the `provider.order` field to ["Novita"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.02 / 1M input tokens, $0.05 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints.
Price factsInput: $0.02 / 1M tokens Output: $0.05 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:26.014Z Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints
Price Since Endpoints on OpenRouter5.Mistral: Mistral Nemo
Parasail via OpenRouter · parasail/fp8
fp8- Input / 1M
- $0.03
- Output / 1M
- $0.03
- Cache read / 1M
- —
- Context
- 131K
- Your cost / mo
- $0.36
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by Parasail through OpenRouter, not a direct-provider contract. Published rates for route parasail/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- parasail/fp8 · fp8
Declared parameters: frequency_penalty, logit_bias, logprobs, max_tokens, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_logprobs, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: mistralai/mistral-nemo Protocol: openai Provider routing: Parasail API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "mistralai/mistral-nemo", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "Parasail" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "mistralai/mistral-nemo". Route requests to provider "Parasail" by setting the `provider.order` field to ["Parasail"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.03 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints.
Price factsInput: $0.03 / 1M tokens Output: $0.03 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:28.816Z Source: https://openrouter.ai/api/v1/models/mistralai/mistral-nemo/endpoints
Price Since Endpoints on OpenRouter6.OpenAI: gpt-oss-20b
Darkbloom via OpenRouter · darkbloom/fp8
fp8- Input / 1M
- $0.018
- Output / 1M
- $0.09
- Cache read / 1M
- $0.009
- Context
- 131K
- Your cost / mo
- $0.36
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by Darkbloom through OpenRouter, not a direct-provider contract. Published rates for route darkbloom/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- $0.009
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- darkbloom/fp8 · fp8
Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-20b Protocol: openai Provider routing: Darkbloom API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "openai/gpt-oss-20b", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "Darkbloom" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "Darkbloom" by setting the `provider.order` field to ["Darkbloom"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.018 / 1M input tokens, $0.09 / 1M output tokens, $0.009 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints.
Price factsInput: $0.018 / 1M tokens Output: $0.09 / 1M tokens Cache read: $0.009 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.420Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints
Price Since Endpoints on OpenRouter7.IBM: Granite 4.0 Micro
Cloudflare via OpenRouter · cloudflare
- Input / 1M
- $0.017
- Output / 1M
- $0.112
- Cache read / 1M
- —
- Context
- 131K
- Your cost / mo
- $0.394
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by Cloudflare through OpenRouter, not a direct-provider contract. Published rates for route cloudflare; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- cloudflare · unknown
Declared parameters: frequency_penalty, logit_bias, logprobs, max_tokens, min_p, presence_penalty, repetition_penalty, response_format, seed, stop, temperature, top_k, top_logprobs, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: ibm-granite/granite-4.0-h-micro Protocol: openai Provider routing: Cloudflare API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "ibm-granite/granite-4.0-h-micro", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "Cloudflare" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "ibm-granite/granite-4.0-h-micro". Route requests to provider "Cloudflare" by setting the `provider.order` field to ["Cloudflare"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.017 / 1M input tokens, $0.112 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/ibm-granite/granite-4.0-h-micro/endpoints.
Price factsInput: $0.017 / 1M tokens Output: $0.112 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:24.978Z Source: https://openrouter.ai/api/v1/models/ibm-granite/granite-4.0-h-micro/endpoints
Price Since Endpoints on OpenRouter8.OpenAI: gpt-oss-20b
AkashML via OpenRouter · akashml/fp4
fp4- Input / 1M
- $0.02
- Output / 1M
- $0.10
- Cache read / 1M
- —
- Context
- 131K
- Your cost / mo
- $0.40
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by AkashML through OpenRouter, not a direct-provider contract. Published rates for route akashml/fp4; fp4 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- akashml/fp4 · fp4
Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-20b Protocol: openai Provider routing: AkashML API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "openai/gpt-oss-20b", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "AkashML" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "AkashML" by setting the `provider.order` field to ["AkashML"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.02 / 1M input tokens, $0.10 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints.
Price factsInput: $0.02 / 1M tokens Output: $0.10 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.420Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints
Price Since Endpoints on OpenRouter9.Nex AGI: Nex-N2.5-Mini
Nex AGI via OpenRouter · nex-agi/bf16
bf16- Input / 1M
- $0.025
- Output / 1M
- $0.10
- Cache read / 1M
- $0.0025
- Context
- 262K
- Your cost / mo
- $0.45
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by Nex AGI through OpenRouter, not a direct-provider contract. Published rates for route nex-agi/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- $0.0025
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- nex-agi/bf16 · bf16
Declared parameters: include_reasoning, logprobs, max_tokens, reasoning, reasoning_effort, structured_outputs, temperature, top_k, top_logprobs, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: nex-agi/nex-n2.5-mini Protocol: openai Provider routing: Nex AGI API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "nex-agi/nex-n2.5-mini", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "Nex AGI" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "nex-agi/nex-n2.5-mini". Route requests to provider "Nex AGI" by setting the `provider.order` field to ["Nex AGI"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.025 / 1M input tokens, $0.10 / 1M output tokens, $0.0025 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/nex-agi/nex-n2.5-mini/endpoints.
Price factsInput: $0.025 / 1M tokens Output: $0.10 / 1M tokens Cache read: $0.0025 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:30.211Z Source: https://openrouter.ai/api/v1/models/nex-agi/nex-n2.5-mini/endpoints
Price Since Endpoints on OpenRouter10.Sao10K: Llama 3 8B Lunaris
DeepInfra via OpenRouter · deepinfra/turbo
fp8- Input / 1M
- $0.04
- Output / 1M
- $0.05
- Cache read / 1M
- —
- Context
- 8K
- Your cost / mo
- $0.50
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/turbo; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- deepinfra/turbo · fp8
Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, seed, stop, temperature, top_k, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: sao10k/l3-lunaris-8b Protocol: openai Provider routing: DeepInfra API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "sao10k/l3-lunaris-8b", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "DeepInfra" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "sao10k/l3-lunaris-8b". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.04 / 1M input tokens, $0.05 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints.
Price factsInput: $0.04 / 1M tokens Output: $0.05 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:44.773Z Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints
Price Since Endpoints on OpenRouter11.Sao10K: Llama 3 8B Lunaris
Parasail via OpenRouter · parasail/bf16
bf16- Input / 1M
- $0.04
- Output / 1M
- $0.05
- Cache read / 1M
- —
- Context
- 8K
- Your cost / mo
- $0.50
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by Parasail through OpenRouter, not a direct-provider contract. Published rates for route parasail/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- parasail/bf16 · bf16
Declared parameters: frequency_penalty, logit_bias, logprobs, max_tokens, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_logprobs, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: sao10k/l3-lunaris-8b Protocol: openai Provider routing: Parasail API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "sao10k/l3-lunaris-8b", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "Parasail" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "sao10k/l3-lunaris-8b". Route requests to provider "Parasail" by setting the `provider.order` field to ["Parasail"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.04 / 1M input tokens, $0.05 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints.
Price factsInput: $0.04 / 1M tokens Output: $0.05 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:44.773Z Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints
Price Since Endpoints on OpenRouter12.OpenAI: gpt-oss-20b
CoreWeave via OpenRouter · coreweave/fp4
fp4- Input / 1M
- $0.03
- Output / 1M
- $0.13
- Cache read / 1M
- $0.03
- Context
- 131K
- Your cost / mo
- $0.56
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by CoreWeave through OpenRouter, not a direct-provider contract. Published rates for route coreweave/fp4; fp4 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- $0.03
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- coreweave/fp4 · fp4
Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-20b Protocol: openai Provider routing: CoreWeave API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "openai/gpt-oss-20b", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "CoreWeave" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "CoreWeave" by setting the `provider.order` field to ["CoreWeave"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.13 / 1M output tokens, $0.03 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints.
Price factsInput: $0.03 / 1M tokens Output: $0.13 / 1M tokens Cache read: $0.03 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.420Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints
Price Since Endpoints on OpenRouter13.Qwen: Qwen3.7 Flash
Alibaba via OpenRouter · alibaba
- Input / 1M
- $0.03
- Output / 1M
- $0.13
- Cache read / 1M
- $0.006
- Context
- 1M
- Your cost / mo
- $0.56
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by Alibaba through OpenRouter, not a direct-provider contract. Published rates for route alibaba; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- $0.006
- Cache write / 1M
- $0.038
- Per request
- Not published
- Route
- alibaba · unknown
Declared parameters: include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, response_format, seed, temperature, tool_choice, tools, top_logprobs, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: qwen/qwen3.7-flash Protocol: openai Provider routing: Alibaba API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "qwen/qwen3.7-flash", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "Alibaba" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "qwen/qwen3.7-flash". Route requests to provider "Alibaba" by setting the `provider.order` field to ["Alibaba"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.13 / 1M output tokens, $0.006 / 1M cache read tokens, $0.038 / 1M cache write tokens. Source: https://openrouter.ai/api/v1/models/qwen/qwen3.7-flash/endpoints.
Price factsInput: $0.03 / 1M tokens Output: $0.13 / 1M tokens Cache read: $0.006 / 1M tokens Cache write: $0.038 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:43.500Z Source: https://openrouter.ai/api/v1/models/qwen/qwen3.7-flash/endpoints
Price Since Endpoints on OpenRouter14.OpenAI: gpt-oss-20b
DekaLLM via OpenRouter · dekallm/bf16
bf16- Input / 1M
- $0.029
- Output / 1M
- $0.14
- Cache read / 1M
- —
- Context
- 131K
- Your cost / mo
- $0.57
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by DekaLLM through OpenRouter, not a direct-provider contract. Published rates for route dekallm/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- dekallm/bf16 · bf16
Declared parameters: frequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_logprobs, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-20b Protocol: openai Provider routing: DekaLLM API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "openai/gpt-oss-20b", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "DekaLLM" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "DekaLLM" by setting the `provider.order` field to ["DekaLLM"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.029 / 1M input tokens, $0.14 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints.
Price factsInput: $0.029 / 1M tokens Output: $0.14 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.420Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints
Price Since Endpoints on OpenRouter15.OpenAI: gpt-oss-20b
DeepInfra via OpenRouter · deepinfra/bf16
bf16- Input / 1M
- $0.03
- Output / 1M
- $0.14
- Cache read / 1M
- —
- Context
- 131K
- Your cost / mo
- $0.58
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- deepinfra/bf16 · bf16
Declared parameters: frequency_penalty, include_reasoning, logit_bias, max_tokens, min_p, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-20b Protocol: openai Provider routing: DeepInfra API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "openai/gpt-oss-20b", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "DeepInfra" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.14 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints.
Price factsInput: $0.03 / 1M tokens Output: $0.14 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.420Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints
Price Since Endpoints on OpenRouter16.Inference.net: Schematron V2 Turbo
InferenceNet via OpenRouter · inference-net
- Input / 1M
- $0.03
- Output / 1M
- $0.15
- Cache read / 1M
- $0.03
- Context
- 128K
- Your cost / mo
- $0.60
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by InferenceNet through OpenRouter, not a direct-provider contract. Published rates for route inference-net; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- $0.03
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- inference-net · unknown
Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: inference-net/schematron-v2-turbo Protocol: openai Provider routing: InferenceNet API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "inference-net/schematron-v2-turbo", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "InferenceNet" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "inference-net/schematron-v2-turbo". Route requests to provider "InferenceNet" by setting the `provider.order` field to ["InferenceNet"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.15 / 1M output tokens, $0.03 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/inference-net/schematron-v2-turbo/endpoints.
Price factsInput: $0.03 / 1M tokens Output: $0.15 / 1M tokens Cache read: $0.03 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:25.607Z Source: https://openrouter.ai/api/v1/models/inference-net/schematron-v2-turbo/endpoints
Price Since Endpoints on OpenRouter17.OpenAI: gpt-oss-20b
Parasail via OpenRouter · parasail/fp4
fp4- Input / 1M
- $0.03
- Output / 1M
- $0.15
- Cache read / 1M
- $0.02
- Context
- 131K
- Your cost / mo
- $0.60
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by Parasail through OpenRouter, not a direct-provider contract. Published rates for route parasail/fp4; fp4 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- $0.02
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- parasail/fp4 · fp4
Declared parameters: frequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-20b Protocol: openai Provider routing: Parasail API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "openai/gpt-oss-20b", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "Parasail" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "Parasail" by setting the `provider.order` field to ["Parasail"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.15 / 1M output tokens, $0.02 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints.
Price factsInput: $0.03 / 1M tokens Output: $0.15 / 1M tokens Cache read: $0.02 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.420Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints
Price Since Endpoints on OpenRouter18.Sao10K: Llama 3 8B Lunaris
Novita via OpenRouter · novita/bf16
bf16- Input / 1M
- $0.05
- Output / 1M
- $0.05
- Cache read / 1M
- —
- Context
- 8K
- Your cost / mo
- $0.60
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by Novita through OpenRouter, not a direct-provider contract. Published rates for route novita/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- novita/bf16 · bf16
Declared parameters: frequency_penalty, max_tokens, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: sao10k/l3-lunaris-8b Protocol: openai Provider routing: Novita API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "sao10k/l3-lunaris-8b", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "Novita" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "sao10k/l3-lunaris-8b". Route requests to provider "Novita" by setting the `provider.order` field to ["Novita"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.05 / 1M input tokens, $0.05 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints.
Price factsInput: $0.05 / 1M tokens Output: $0.05 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:44.773Z Source: https://openrouter.ai/api/v1/models/sao10k/l3-lunaris-8b/endpoints
Price Since Endpoints on OpenRouter19.Amazon: Nova Micro 1.0
Amazon Bedrock via OpenRouter · amazon-bedrock
- Input / 1M
- $0.035
- Output / 1M
- $0.14
- Cache read / 1M
- —
- Context
- 128K
- Your cost / mo
- $0.63
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by Amazon Bedrock through OpenRouter, not a direct-provider contract. Published rates for route amazon-bedrock; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- amazon-bedrock · unknown
Declared parameters: max_tokens, stop, temperature, tools, top_k, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: amazon/nova-micro-v1 Protocol: openai Provider routing: Amazon Bedrock API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "amazon/nova-micro-v1", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "Amazon Bedrock" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "amazon/nova-micro-v1". Route requests to provider "Amazon Bedrock" by setting the `provider.order` field to ["Amazon Bedrock"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.035 / 1M input tokens, $0.14 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/amazon/nova-micro-v1/endpoints.
Price factsInput: $0.035 / 1M tokens Output: $0.14 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:17.502Z Source: https://openrouter.ai/api/v1/models/amazon/nova-micro-v1/endpoints
Price Since Endpoints on OpenRouter20.Amazon: Nova Micro 1.0
Amazon Bedrock via OpenRouter · amazon-bedrock/eu-west-1
- Input / 1M
- $0.035
- Output / 1M
- $0.14
- Cache read / 1M
- —
- Context
- 128K
- Your cost / mo
- $0.63
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by Amazon Bedrock through OpenRouter, not a direct-provider contract. Published rates for route amazon-bedrock/eu-west-1; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- amazon-bedrock/eu-west-1 · unknown
Declared parameters: max_tokens, stop, temperature, tools, top_k, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: amazon/nova-micro-v1 Protocol: openai Provider routing: Amazon Bedrock API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "amazon/nova-micro-v1", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "Amazon Bedrock" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "amazon/nova-micro-v1". Route requests to provider "Amazon Bedrock" by setting the `provider.order` field to ["Amazon Bedrock"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.035 / 1M input tokens, $0.14 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/amazon/nova-micro-v1/endpoints.
Price factsInput: $0.035 / 1M tokens Output: $0.14 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:17.502Z Source: https://openrouter.ai/api/v1/models/amazon/nova-micro-v1/endpoints
Price Since Endpoints on OpenRouter21.OpenAI: gpt-oss-120b
CoreWeave via OpenRouter · coreweave/fp4
fp4- Input / 1M
- $0.03
- Output / 1M
- $0.17
- Cache read / 1M
- $0.03
- Context
- 131K
- Your cost / mo
- $0.64
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by CoreWeave through OpenRouter, not a direct-provider contract. Published rates for route coreweave/fp4; fp4 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- $0.03
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- coreweave/fp4 · fp4
Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-120b Protocol: openai Provider routing: CoreWeave API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "openai/gpt-oss-120b", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "CoreWeave" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-120b". Route requests to provider "CoreWeave" by setting the `provider.order` field to ["CoreWeave"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.17 / 1M output tokens, $0.03 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints.
Price factsInput: $0.03 / 1M tokens Output: $0.17 / 1M tokens Cache read: $0.03 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.695Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints
Price Since Endpoints on OpenRouter22.OpenAI: GPT-5 Nano
OpenAI via OpenRouter · openai/flex
- Input / 1M
- $0.025
- Output / 1M
- $0.20
- Cache read / 1M
- $0.0025
- Context
- 400K
- Your cost / mo
- $0.65
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by OpenAI through OpenRouter, not a direct-provider contract. Published rates for route openai/flex; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- $0.0025
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- openai/flex · unknown
Declared parameters: include_reasoning, max_tokens, reasoning, reasoning_effort, response_format, seed, structured_outputs, tool_choice, tools
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: openai/gpt-5-nano Protocol: openai Provider routing: OpenAI API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "openai/gpt-5-nano", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "OpenAI" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-5-nano". Route requests to provider "OpenAI" by setting the `provider.order` field to ["OpenAI"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.025 / 1M input tokens, $0.20 / 1M output tokens, $0.0025 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-5-nano/endpoints.
Price factsInput: $0.025 / 1M tokens Output: $0.20 / 1M tokens Cache read: $0.0025 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:33.354Z Source: https://openrouter.ai/api/v1/models/openai/gpt-5-nano/endpoints
Price Since Endpoints on OpenRouter23.Meta: Llama 3.1 8B Instruct
Groq via OpenRouter · groq
- Input / 1M
- $0.05
- Output / 1M
- $0.08
- Cache read / 1M
- $0.025
- Context
- 131K
- Your cost / mo
- $0.66
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by Groq through OpenRouter, not a direct-provider contract. Published rates for route groq; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- $0.025
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- groq · unknown
Declared parameters: max_tokens, response_format, seed, stop, temperature, tool_choice, tools, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: meta-llama/llama-3.1-8b-instruct Protocol: openai Provider routing: Groq API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "meta-llama/llama-3.1-8b-instruct", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "Groq" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "meta-llama/llama-3.1-8b-instruct". Route requests to provider "Groq" by setting the `provider.order` field to ["Groq"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.05 / 1M input tokens, $0.08 / 1M output tokens, $0.025 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints.
Price factsInput: $0.05 / 1M tokens Output: $0.08 / 1M tokens Cache read: $0.025 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:26.014Z Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.1-8b-instruct/endpoints
Price Since Endpoints on OpenRouter24.Mistral: Mistral Small 3
DeepInfra via OpenRouter · deepinfra/fp8
fp8- Input / 1M
- $0.05
- Output / 1M
- $0.08
- Cache read / 1M
- —
- Context
- 33K
- Your cost / mo
- $0.66
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/fp8; fp8 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- deepinfra/fp8 · fp8
Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: mistralai/mistral-small-24b-instruct-2501 Protocol: openai Provider routing: DeepInfra API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "mistralai/mistral-small-24b-instruct-2501", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "DeepInfra" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "mistralai/mistral-small-24b-instruct-2501". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.05 / 1M input tokens, $0.08 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/mistralai/mistral-small-24b-instruct-2501/endpoints.
Price factsInput: $0.05 / 1M tokens Output: $0.08 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:28.946Z Source: https://openrouter.ai/api/v1/models/mistralai/mistral-small-24b-instruct-2501/endpoints
Price Since Endpoints on OpenRouter25.OpenAI: gpt-oss-120b
DekaLLM via OpenRouter · dekallm/bf16
bf16- Input / 1M
- $0.03
- Output / 1M
- $0.18
- Cache read / 1M
- —
- Context
- 131K
- Your cost / mo
- $0.66
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by DekaLLM through OpenRouter, not a direct-provider contract. Published rates for route dekallm/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- dekallm/bf16 · bf16
Declared parameters: frequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_logprobs, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-120b Protocol: openai Provider routing: DekaLLM API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "openai/gpt-oss-120b", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "DekaLLM" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-120b". Route requests to provider "DekaLLM" by setting the `provider.order` field to ["DekaLLM"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.03 / 1M input tokens, $0.18 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints.
Price factsInput: $0.03 / 1M tokens Output: $0.18 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.695Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints
Price Since Endpoints on OpenRouter26.Meta: Llama 3.2 1B Instruct
Cloudflare via OpenRouter · cloudflare
- Input / 1M
- $0.027
- Output / 1M
- $0.201
- Cache read / 1M
- —
- Context
- 60K
- Your cost / mo
- $0.672
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by Cloudflare through OpenRouter, not a direct-provider contract. Published rates for route cloudflare; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- cloudflare · unknown
Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, seed, stop, temperature, top_k, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: meta-llama/llama-3.2-1b-instruct Protocol: openai Provider routing: Cloudflare API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "meta-llama/llama-3.2-1b-instruct", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "Cloudflare" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "meta-llama/llama-3.2-1b-instruct". Route requests to provider "Cloudflare" by setting the `provider.order` field to ["Cloudflare"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.027 / 1M input tokens, $0.201 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.2-1b-instruct/endpoints.
Price factsInput: $0.027 / 1M tokens Output: $0.201 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:26.127Z Source: https://openrouter.ai/api/v1/models/meta-llama/llama-3.2-1b-instruct/endpoints
Price Since Endpoints on OpenRouter27.Cohere: Command R7B (12-2024)
Cohere via OpenRouter · cohere
- Input / 1M
- $0.0375
- Output / 1M
- $0.15
- Cache read / 1M
- —
- Context
- 128K
- Your cost / mo
- $0.675
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by Cohere through OpenRouter, not a direct-provider contract. Published rates for route cohere; unknown quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- cohere · unknown
Declared parameters: frequency_penalty, max_tokens, presence_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: cohere/command-r7b-12-2024 Protocol: openai Provider routing: Cohere API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "cohere/command-r7b-12-2024", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "Cohere" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "cohere/command-r7b-12-2024". Route requests to provider "Cohere" by setting the `provider.order` field to ["Cohere"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.0375 / 1M input tokens, $0.15 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/cohere/command-r7b-12-2024/endpoints.
Price factsInput: $0.0375 / 1M tokens Output: $0.15 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:21.103Z Source: https://openrouter.ai/api/v1/models/cohere/command-r7b-12-2024/endpoints
Price Since Endpoints on OpenRouter28.Google: Gemma 3 4B
DeepInfra via OpenRouter · deepinfra/bf16
bf16- Input / 1M
- $0.05
- Output / 1M
- $0.10
- Cache read / 1M
- —
- Context
- 131K
- Your cost / mo
- $0.70
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by DeepInfra through OpenRouter, not a direct-provider contract. Published rates for route deepinfra/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- deepinfra/bf16 · bf16
Declared parameters: frequency_penalty, logit_bias, max_tokens, min_p, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: google/gemma-3-4b-it Protocol: openai Provider routing: DeepInfra API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "google/gemma-3-4b-it", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "DeepInfra" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "google/gemma-3-4b-it". Route requests to provider "DeepInfra" by setting the `provider.order` field to ["DeepInfra"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.05 / 1M input tokens, $0.10 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/google/gemma-3-4b-it/endpoints.
Price factsInput: $0.05 / 1M tokens Output: $0.10 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:24.464Z Source: https://openrouter.ai/api/v1/models/google/gemma-3-4b-it/endpoints
Price Since Endpoints on OpenRouter29.OpenAI: gpt-oss-20b
Novita via OpenRouter · novita/fp4
fp4- Input / 1M
- $0.04
- Output / 1M
- $0.15
- Cache read / 1M
- —
- Context
- 131K
- Your cost / mo
- $0.70
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by Novita through OpenRouter, not a direct-provider contract. Published rates for route novita/fp4; fp4 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- Not published
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- novita/fp4 · fp4
Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_k, top_logprobs, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-20b Protocol: openai Provider routing: Novita API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "openai/gpt-oss-20b", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "Novita" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-20b". Route requests to provider "Novita" by setting the `provider.order` field to ["Novita"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.04 / 1M input tokens, $0.15 / 1M output tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints.
Price factsInput: $0.04 / 1M tokens Output: $0.15 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.420Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-20b/endpoints
Price Since Endpoints on OpenRouter30.OpenAI: gpt-oss-120b
AkashML via OpenRouter · akashml/bf16
bf16- Input / 1M
- $0.033
- Output / 1M
- $0.187
- Cache read / 1M
- $0.033
- Context
- 131K
- Your cost / mo
- $0.704
- vs cheapest
- —
- Uptime 1d
- n/a
Billing conditions
Served by AkashML through OpenRouter, not a direct-provider contract. Published rates for route akashml/bf16; bf16 quantization. Other fees, tier conditions and regional access must be checked before use.
- Cache read / 1M
- $0.033
- Cache write / 1M
- Not published
- Per request
- Not published
- Route
- akashml/bf16 · bf16
Declared parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p
Use this endpoint
EndpointBase URL: https://openrouter.ai/api/v1 Model: openai/gpt-oss-120b Protocol: openai Provider routing: AkashML API key env: OPENROUTER_API_KEY Docs: https://openrouter.ai/docs
curlcurl "https://openrouter.ai/api/v1/chat/completions" -H "Authorization: Bearer $OPENROUTER_API_KEY" -H "Content-Type: application/json" -d '{ "model": "openai/gpt-oss-120b", "messages": [ { "role": "user", "content": "Hello" } ], "provider": { "order": [ "AkashML" ], "allow_fallbacks": false } }'Prompt for your coding agentUse the OpenAI-compatible API at https://openrouter.ai/api/v1 with model "openai/gpt-oss-120b". Route requests to provider "AkashML" by setting the `provider.order` field to ["AkashML"] with `allow_fallbacks: false`. Read the API key from the OPENROUTER_API_KEY environment variable; never hardcode it. Pricing observed on 2026-09-29: $0.033 / 1M input tokens, $0.187 / 1M output tokens, $0.033 / 1M cache read tokens. Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints.
Price factsInput: $0.033 / 1M tokens Output: $0.187 / 1M tokens Cache read: $0.033 / 1M tokens Currency: USD Regions: global Unit source: api Checked: 2026-09-29T18:01:16.695Z Source: https://openrouter.ai/api/v1/models/openai/gpt-oss-120b/endpoints
Price Since Endpoints on OpenRouter
Range = your assumptions’ low and high ends; retries from 1-day uptime.
Prices as of 2026-09-29 · tracked since 2026-09-29 · 8 checks. 4 model checks failed; earlier prices keep their timestamps. History & source status
How we compare
What “cheapest” means here
The lowest estimate among the quotes currently listed in one currency, for the volumes and tier you set. Standard, batch, off-peak, free and promotional pricing never share a ranking, and currencies are never converted. Model quality, rate limits and regional access are separate questions.
Every price keeps its source and currency
Provider endpoints keep paired input and output rates, route tags and quantization. Direct quotes keep the currency and region their provider bills in. Snapshots record when we observed a price change, not when a provider made it effective.
What is not covered yet
Most official direct price pages, regional and context-tier pricing, embeddings, audio, image and video billing, free-tier allowances and measured latency. Unknown values are shown as unknown and never counted as zero.
Models by providerProvider coveragePrice history & source statusSnapshot JSON