Z.ai
GLM-5.3-Flash
CONTEXT 1.05M
MAX OUTPUT 131.07K
Text Image 1 gateway
A multimodal Z.ai model for coding, visual understanding and tool-using agents. Its sparse architecture targets efficient inference with a long context window.
Details & access +
Limits above: NVIDIA. Gateway differences are listed below.
Thinking levels low ยท high ยท max
Default effort max
NVIDIA
z-ai/glm-5.3-flash
API trial terms, quotas and model licenses apply.
base nvidia/z-ai/glm-5.3-flash
Sources
Try this model โ
Moonshot AI
Kimi K3
CONTEXT 1.05M
MAX OUTPUT 131.07K
Text Image Video 1 gateway
Moonshot AI's multimodal reasoning model for long-running agent tasks. It accepts text and images and offers adjustable thinking effort.
Details & access +
Limits above: NVIDIA. Gateway differences are listed below.
Thinking levels low ยท high ยท max
Thinking controls Can be toggled
NVIDIA
moonshotai/kimi-k3
API trial terms, quotas and model licenses apply.
base nvidia/moonshotai/kimi-k3
Sources
Try this model โ
Meituan
LongCat 2.5 Preview
26 Sep 2026
๐ OpenCode Zen
29 Sep 2026
๐ฅ Nous Portal
CONTEXT 1.05M
MAX OUTPUT 131.07K
โ Reasoning
Text Image 2 gateways
Meituan's multimodal coding and agent model, with image understanding and long-context reasoning. Its thinking mode can be switched on or off; named effort levels are not documented.
Details & access +
Limits above: Nous Portal. Gateway differences are listed below.
Thinking controls Can be toggled
Reviewed as the same advertised Meituan LongCat 2.5 Preview model across Nous Portal and OpenCode Zen, not LongCat 2.0. Gateway limits and access conditions remain separate.
OpenCode Zen
longcat-2.5-preview-free
No passing direct API route in the latest scan.
base opencode/longcat-2.5-preview-free
Nous Portal
meituan/longcat-2.5-preview:free
Check the portal for current free routes and limits.
base nous/meituan/longcat-2.5-preview:free
Sources
Try this model โ
Anonymous ยท stealth preview
Space Bunny Alpha
26 Sep 2026
๐ Kilo๐ OpenCode Zen๐ OpenRouter
CONTEXT 1M
MAX OUTPUT 524.29K
โ Reasoning
Text Image Video 3 gateways
Stealth model ๐ฅท
An anonymous preview model advertised for coding, reasoning and multimodal input. Its developer and underlying model identity have not been disclosed.
Details & access +
Limits above: Kilo. Gateway differences are listed below.
Curated Space Bunny alias family. The creator is undisclosed; matching aliases do not verify identical weights across gateways.
Kilo
stealth/space-bunny-alpha
Free availability and upstream data policies vary.
base kilo/stealth/space-bunny-alpha
OpenCode Zen
space-bunny-free
Context here 1,048,576
Thinking levels here: low ยท medium ยท high ยท xhigh ยท max
base opencode/space-bunny-free
OpenRouter
stealth/space-bunny-alpha
Thinking levels here: max ยท xhigh ยท high ยท medium ยท low
Free-model quotas and model-specific terms apply.
base openrouter/stealth/space-bunny-alpha
Sources
Try this model โ
SDAIA
ALLaM-2-7b
CONTEXT 4.1K
MAX OUTPUT 4.1K
Text 1 gateway
SDAIA's 7-billion-parameter instruction-tuned model for Arabic and English. Trained from scratch with staged English and Arabic-English pretraining, it supports bilingual conversations, text generation and summarization.
Details & access +
Limits above: Groq. Gateway differences are listed below.
Groq
allam-2-7b
A free developer tier with per-model limits.
Sources
Try this model โ
Kilo
Auto Free
CONTEXT 256K
MAX OUTPUT 32.77K
โ Reasoning
Text 1 gateway
Kilo's Auto Free router distributes requests across available free models on OpenRouter, with the pool updated as availability changes. It is not a fixed model; upstream providers may log prompts and outputs, so do not submit confidential data.
Details & access +
Limits above: Kilo. Gateway differences are listed below.
Kilo
kilo-auto/free
Free availability and upstream data policies vary.
Sources
Try this model โ
Cohere
North Mini Code
โ Reasoning
Text 2 gateways
Cohere's compact mixture-of-experts model focused on agentic coding. It targets code changes and tool-driven software development.
Details & access +
Limits above: Kilo. Gateway differences are listed below.
Kilo
cohere/north-mini-code:free
Free availability and upstream data policies vary.
base kilo/cohere/north-mini-code:free
OpenRouter
cohere/north-mini-code:free
Free-model quotas and model-specific terms apply.
fast openrouter/cohere/north-mini-code:free-fastthink openrouter/cohere/north-mini-code:free-think
Sources
Try this model โ
Dots Studio
Dots3-Note Preview
CONTEXT 512K
MAX OUTPUT 460.8K
โ Reasoning
Text Image 2 gateways
A preview of Dots Studio's lighter Dots 3 mixture-of-experts model. It supports long-context work and configurable reasoning.
Details & access +
Limits above: Kilo. Gateway differences are listed below.
Kilo
dots-studio/dots-3-note-preview:free
Free availability and upstream data policies vary.
base kilo/dots-studio/dots-3-note-preview:free
OpenRouter
dots-studio/dots-3-note-preview:free
Free-model quotas and model-specific terms apply.
fast openrouter/dots-studio/dots-3-note-preview:free-fastthink openrouter/dots-studio/dots-3-note-preview:free-think
Sources
Try this model โ
Google DeepMind
Lyria 3 Pro Preview
CONTEXT 1.05M
MAX OUTPUT 65.54K
Text Image 2 gateways
Google's music-generation model for full-length songs, not a conventional chat LLM. Zero token prices in the catalog do not mean song generation is free.
Details & access +
Music generation is billed per song/clip despite zero token prices in the catalog. This is not an unconditionally free chat model.
Limits above: OpenRouter. Gateway differences are listed below.
Output text, audio
Kilo
google/lyria-3-pro-preview
Free availability and upstream data policies vary.
No passing direct API route in the latest scan.
base kilo/google/lyria-3-pro-preview
OpenRouter
google/lyria-3-pro-preview
Free-model quotas and model-specific terms apply.
base openrouter/google/lyria-3-pro-preview
Sources
Try this model โ
OpenAI
GPT OSS 120B
CONTEXT 131.07K
MAX OUTPUT 65.54K
Text 1 gateway
OpenAI's larger open-weight reasoning model, served here by Groq. It offers adjustable reasoning effort for text and tool-based tasks.
Details & access +
Limits above: Groq. Gateway differences are listed below.
Thinking levels low ยท medium ยท high
Groq
openai/gpt-oss-120b
A free developer tier with per-model limits.
base groq/openai/gpt-oss-120b
Sources
Try this model โ
OpenAI
GPT OSS 20B
CONTEXT 131.07K
MAX OUTPUT 65.54K
Text 1 gateway
The smaller open-weight GPT-OSS reasoning model, served here by Groq. It supports low, medium and high reasoning effort.
Details & access +
Limits above: Groq. Gateway differences are listed below.
Thinking levels low ยท medium ยท high
Groq
openai/gpt-oss-20b
A free developer tier with per-model limits.
base groq/openai/gpt-oss-20b
Sources
Try this model โ
InclusionAI
Ling 3.0 Flash Sante
CONTEXT 262.14K
MAX OUTPUT 32.77K
โ Reasoning
Text 3 gateways
A health- and medicine-focused Ling 3.0 Flash model from InclusionAI. Domain specialization does not make its responses medical advice.
Details & access +
Limits above: Kilo. Gateway differences are listed below.
Kilo
inclusionai/ling-3.0-flash-sante:free
Free availability and upstream data policies vary.
base kilo/inclusionai/ling-3.0-flash-sante:free
Nous Portal
inclusionai/ling-3.0-flash-sante:free
Check the portal for current free routes and limits.
base nous/inclusionai/ling-3.0-flash-sante:free
OpenRouter
inclusionai/ling-3.0-flash-sante:free
Free-model quotas and model-specific terms apply.
fast openrouter/inclusionai/ling-3.0-flash-sante:free-fastthink openrouter/inclusionai/ling-3.0-flash-sante:free-think
Sources
Try this model โ
Liquid AI
LFM2.5-2.6B
CONTEXT 65.54K
MAX OUTPUT 8.19K
โ Reasoning
Text 2 gateways
Liquid AI's compact reasoning model for extraction, retrieval and agent workflows. Its model guidance does not position it as an agentic coding specialist.
Details & access +
Limits above: Kilo. Gateway differences are listed below.
Kilo
liquid/lfm-2.5-2.6b:free
Free availability and upstream data policies vary.
base kilo/liquid/lfm-2.5-2.6b:free
OpenRouter
liquid/lfm-2.5-2.6b:free
Free-model quotas and model-specific terms apply.
base openrouter/liquid/lfm-2.5-2.6b:free
Sources
Try this model โ
NVIDIA
Nemotron 3 Nano Omni
CONTEXT 256K
MAX OUTPUT 65.54K
โ Reasoning
Text Audio Image Video 2 gateways
A multimodal NVIDIA model that can work with text, images, audio and video. Designed for perception and context processing in agent workflows.
Details & access +
Limits above: Kilo. Gateway differences are listed below.
Kilo
nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free
Free availability and upstream data policies vary.
base kilo/nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free
OpenRouter
nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free
Free-model quotas and model-specific terms apply.
fast openrouter/nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free-fastthink openrouter/nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free-think
Sources
Try this model โ
NVIDIA
Nemotron 3 Super
CONTEXT 262.14K
MAX OUTPUT 235.93K
โ Reasoning
Text 2 gateways
NVIDIA's hybrid mixture-of-experts reasoning model for multi-agent workflows. Only a subset of its total parameters is active per token.
Details & access +
Limits above: Kilo. Gateway differences are listed below.
Kilo
nvidia/nemotron-3-super-120b-a12b:free
Free availability and upstream data policies vary.
base kilo/nvidia/nemotron-3-super-120b-a12b:free
OpenRouter
nvidia/nemotron-3-super-120b-a12b:free
Thinking levels here: medium ยท low
Free-model quotas and model-specific terms apply.
fast openrouter/nvidia/nemotron-3-super-120b-a12b:free-fastthink openrouter/nvidia/nemotron-3-super-120b-a12b:free-think
Sources
Try this model โ
NVIDIA
Nemotron 3 Ultra
CONTEXT 1M
MAX OUTPUT 65.54K
โ Reasoning
Text 3 gateways
NVIDIA's larger Nemotron 3 model for reasoning and agent orchestration. It combines a long context window with a sparse hybrid architecture.
Details & access +
Limits above: Kilo. Gateway differences are listed below.
Kilo
nvidia/nemotron-3-ultra-550b-a55b:free
Free availability and upstream data policies vary.
base kilo/nvidia/nemotron-3-ultra-550b-a55b:free
OpenRouter
nvidia/nemotron-3-ultra-550b-a55b:free
Thinking levels here: high ยท medium
Free-model quotas and model-specific terms apply.
fast openrouter/nvidia/nemotron-3-ultra-550b-a55b:free-fastthink openrouter/nvidia/nemotron-3-ultra-550b-a55b:free-think
OpenCode Zen 22 Sep 2026
nemotron-3-ultra-free
Max output here 128,000
No passing direct API route in the latest scan.
base opencode/nemotron-3-ultra-free
Sources
Try this model โ
NVIDIA
Nemotron 3.5 Content Safety
CONTEXT 128K
MAX OUTPUT 8.19K
โ Reasoning
Text Image 2 gateways
A compact multimodal safety model for evaluating prompts and model responses. This is a guardrail specialist, not a general-purpose chat assistant.
Details & access +
Limits above: Kilo. Gateway differences are listed below.
Kilo
nvidia/nemotron-3.5-content-safety:free
Free availability and upstream data policies vary.
base kilo/nvidia/nemotron-3.5-content-safety:free
OpenRouter
nvidia/nemotron-3.5-content-safety:free
Free-model quotas and model-specific terms apply.
fast openrouter/nvidia/nemotron-3.5-content-safety:free-fastthink openrouter/nvidia/nemotron-3.5-content-safety:free-think
Sources
Try this model โ
NVIDIA
Nemotron 3.5 Lightning
CONTEXT 1M
MAX OUTPUT 65.54K
โ Reasoning
Text 3 gateways
NVIDIA's lightweight mixture-of-experts model for responsive agent workflows. It balances a small active parameter count with long-context processing.
Details & access +
Limits above: OpenRouter. Gateway differences are listed below.
Thinking controls Can be toggled
Kilo
nvidia/nemotron-3.5-lightning:free
Free availability and upstream data policies vary.
No passing direct API route in the latest scan.
base kilo/nvidia/nemotron-3.5-lightning:free
OpenRouter
nvidia/nemotron-3.5-lightning:free
Free-model quotas and model-specific terms apply.
fast openrouter/nvidia/nemotron-3.5-lightning:free-fastthink openrouter/nvidia/nemotron-3.5-lightning:free-think
OpenCode Zen 22 Sep 2026
nemotron-3.5-lightning-free
Context here 262,144
Max output here 262,144
No passing direct API route in the latest scan.
base opencode/nemotron-3.5-lightning-free
Sources
Try this model โ
OpenRouter
Free Models Router
โ Reasoning
Text Image 2 gateways
A router that selects from OpenRouter's available free models. The actual model and provider can change between requests.
Details & access +
Limits above: Kilo. Gateway differences are listed below.
Kilo
openrouter/free
Free availability and upstream data policies vary.
OpenRouter
openrouter/free
Free-model quotas and model-specific terms apply.
fast openrouter/openrouter/free-fastthink openrouter/openrouter/free-think
Sources
Try this model โ
Poolside
Laguna S 2.1
CONTEXT 262.14K
MAX OUTPUT 32.77K
โ Reasoning
Text 3 gateways
Poolside's coding-agent model for tool-driven software engineering. Gateway-specific context and output limits are listed separately.
Details & access +
Limits above: Kilo. Gateway differences are listed below.
Kilo
poolside/laguna-s-2.1:free
Free availability and upstream data policies vary.
base kilo/poolside/laguna-s-2.1:free
Nous Portal
poolside/laguna-s-2.1:free
Max output here 131,072
Check the portal for current free routes and limits.
base nous/poolside/laguna-s-2.1:free
OpenRouter
poolside/laguna-s-2.1:free
Free-model quotas and model-specific terms apply.
fast openrouter/poolside/laguna-s-2.1:free-fastthink openrouter/poolside/laguna-s-2.1:free-think
Sources
Try this model โ
Poolside
Laguna XS 2.1
CONTEXT 262.14K
MAX OUTPUT 32.77K
โ Reasoning
Text 3 gateways
A smaller coding-agent model in Poolside's Laguna family. It is designed for software development workflows with a relatively small active parameter count.
Details & access +
Limits above: Kilo. Gateway differences are listed below.
Kilo
poolside/laguna-xs-2.1:free
Free availability and upstream data policies vary.
base kilo/poolside/laguna-xs-2.1:free
Nous Portal
poolside/laguna-xs-2.1:free
Check the portal for current free routes and limits.
base nous/poolside/laguna-xs-2.1:free
OpenRouter
poolside/laguna-xs-2.1:free
Free-model quotas and model-specific terms apply.
fast openrouter/poolside/laguna-xs-2.1:free-fastthink openrouter/poolside/laguna-xs-2.1:free-think
Sources
Try this model โ
Alibaba
Qwen3.8 27B
CONTEXT 131.07K
MAX OUTPUT 16.38K
Text Image 3 gateways
A dense Qwen vision-language model for coding, visual analysis and agent tasks. Thinking controls and context limits vary by gateway.
Details & access +
Limits above: Groq. Gateway differences are listed below.
Thinking levels none ยท default ยท low ยท medium ยท high
Thinking controls Can be toggled
Groq
qwen/qwen3.8-27b
A free developer tier with per-model limits.
base groq/qwen/qwen3.8-27b
Kilo
qwen/qwen3.8-27b:free
Context here 262,144
Max output here 235,929
Reasoning here โ
Thinking levels here: Not documented
Free availability and upstream data policies vary.
No passing direct API route in the latest scan.
base kilo/qwen/qwen3.8-27b:free
OpenRouter
qwen/qwen3.8-27b:free
Context here 262,144
Max output here 235,929
Reasoning here โ
Thinking levels here: xhigh ยท medium ยท low
Free-model quotas and model-specific terms apply.
No passing direct API route in the latest scan.
fast openrouter/qwen/qwen3.8-27b:free-fastthink openrouter/qwen/qwen3.8-27b:free-think
Sources
Try this model โ
StepFun
Step 3.7 Flash
CONTEXT 262.14K
MAX OUTPUT 262.14K
โ Reasoning
Text Image 2 gateways
StepFun's multimodal mixture-of-experts model with image and video understanding. It combines reasoning and tool use in a relatively efficient architecture.
Details & access +
Limits above: Kilo. Gateway differences are listed below.
Kilo
stepfun/step-3.7-flash:free
Free availability and upstream data policies vary.
base kilo/stepfun/step-3.7-flash:free
Nous Portal
stepfun/step-3.7-flash:free
Max output here 32,768
Thinking levels here: high ยท medium ยท low
Check the portal for current free routes and limits.
No passing direct API route in the latest scan.
base nous/stepfun/step-3.7-flash:free
Sources
Try this model โ
Upstage
Solar Pro 4
CONTEXT 524.29K
MAX OUTPUT 131.07K
โ Reasoning
Text 1 gateway
Upstage's long-context model for document-heavy work and agent workflows. Its Nous listing exposes several named reasoning-effort levels.
Details & access +
Limits above: Nous Portal. Gateway differences are listed below.
Thinking levels max ยท xhigh ยท high ยท medium ยท low ยท minimal ยท none
Thinking controls Can be toggled
Default effort medium
Nous Portal
upstage/solar-pro4:free
Check the portal for current free routes and limits.
base nous/upstage/solar-pro4:free
Sources
Try this model โ