OpenRouter
Free model variants
One OpenAI-compatible API for a rotating pool of zero-price model variants.
Shared free-tier limits; model availability changes.
A live, free-first catalogue that joins model availability, official access links and benchmark evidence — so you can choose the best service for your actual task.
Live directory
Free access firstCompare free routes across official gateways, web apps and local runtimes. Scores are shown only when a live benchmark source has them; a dash means “not measured”, not “bad”.
Free rows
0
across live feeds
Direct access
0
API, chat or local
Scored
0
with an intelligence index
Visible
0
of 0 free rows
Free does not mean unlimited. Provider rate limits, sign-in requirements, retention policies and model availability can change. Check the linked official service before sending sensitive or production data. Prices shown for paid context rows are approximate input cost per one million tokens, not a full task quote.
More ways to use AI
These are first-party destinations, not anonymous “free API” aggregators. Quotas and eligibility change, so the directory links to the provider’s current terms instead of promising a fixed limit.
Free model variants
One OpenAI-compatible API for a rotating pool of zero-price model variants.
Shared free-tier limits; model availability changes.
Zen free models
Official Zen gateway used by OpenCode with a small, changing free catalogue.
Gateway limits and eligibility are provider-controlled.
Free gateway models
OpenAI-compatible gateway with a published free-model list.
Gateway limits can change without notice.
Kilo gateway free models
Coding gateway that publishes the models currently available through its free route.
Gateway and provider rate limits apply.
Cline provider gateway
Cline’s provider endpoint; its catalogue can include free model routes.
Availability and limits are controlled by Cline and upstream providers.
Google AI Studio / Gemini API
First-party Gemini playground and developer API with a changing limited free tier.
Model- and region-specific rate limits; confirm the current quota in AI Studio.
GroqCloud free tier
Very fast hosted open-weight models through an OpenAI-compatible API.
Per-model requests and token limits; see the live limits page.
Cerebras Inference
Fast hosted open-weight inference with a limited developer tier.
Quota varies by account and model.
OpenRouter’s live catalogue is useful when you want one API key and several free routes.
Benchmark context
A capability signal for ranking models — not a guarantee of free availability or task quality.
Agents
30%
Long-horizon, tool-using, multi-step professional work and business workflows.
Coding
20%
Terminal-driven software engineering and scientific code generation.
General
30%
Long-document reading, factual recall, and avoiding hallucinations.
Scientific Reasoning
20%
Expert-level knowledge, frontier reasoning, and graduate-level physics.
The index helps shortlist capability. The free-access links tell you where the model can actually be used.
Intelligence Index v4.3 blends 11 weighted benchmarks across 4 categories; the weights below sum to 100%. A model with no score is a new or specialised listing, not an automatic failure.
Treat the overall index as a starting point. Coding, tool use, vision, long-context reading and latency can disagree; use the dedicated sub-scores and modality filter for your workload.
Methodology
11 benchmarks feed the Intelligence Index; 10 related benchmarks are shown for task-specific context only.
| Benchmark | Category | Scale | Weight | What it tests |
|---|---|---|---|---|
| AA-Briefcase | Agents | Professional knowledge-work tasks | 15% | Agentic knowledge work across realistic professional tasks, including planning, tool use, judgement and delivery. |
| GDPval-AA v2 | Agents | 44 occupations · 9 GDP sectors | 10% | Real economically valuable knowledge work across occupations and sectors. |
| AutomationBench-AA | Agents | 657 business workflows | 5% | Multi-application business workflows with partial-credit objectives and guardrail checks, developed with Zapier. |
| Terminal-Bench v4.0 | Coding | 66 terminal tasks | 10% | Complex terminal work across software engineering, machine learning, science, operations, security, hardware and media. |
| SciCode | Coding | 80 problems · 338 subproblems | 10% | Challenging coding problems across 16 natural-science subfields, including mathematics, physics, chemistry and biology. |
| AA-Omniscience Accuracy | General | Economically relevant factual questions | 10% | Factual recall on economically relevant questions, measuring whether the model gives the correct answer. |
| AA-Omniscience Non-Hallucination | General | Economically relevant factual questions | 5% | The ability to avoid fabricating answers when the model lacks enough knowledge to answer reliably. |
| GDP.pdf | General | Professional document reasoning | 10% | Reasoning over professional documents and extracting the right evidence from long, structured files. |
| AA-LCR v1.1 | General | 10k–100k token documents | 5% | Extracting, reasoning about and synthesising information from long-form documents. |
| Humanity's Last Exam | Scientific Reasoning | ~2,500 questions | 10% | Expert-vetted, adversarially hard cross-discipline questions spanning advanced knowledge and reasoning. |
| CritPt | Scientific Reasoning | Research-level physics | 10% | Research-level physics reasoning problems with a private test set. |
| Benchmark | Category | Scale | What it tests |
|---|---|---|---|
| τ³-Banking | Agents | Multi-turn | Tool use and knowledge retrieval in simulated banking customer-support flows. It was replaced by AutomationBench-AA in v4.3. |
| APEX-Agents-AA | Agents | 452 tasks | Long-horizon multi-step professional tasks in banking, consulting and law: plan, act, check, recover and deliver. |
| ITBench-AA | Agents | 59 Kubernetes incidents | Diagnosing Kubernetes incidents from alerts, events, traces and topology. |
| BrowseComp | Agents | 1,266 questions | Multi-hop hard web research and evidence gathering. |
| SWE-bench | Coding | 2,294 tasks | Real-world software-engineering tasks derived from GitHub issues across popular Python repositories. |
| LiveCodeBench | Coding | Continuous | Fresh contest problems from LeetCode, AtCoder and Codeforces designed to reduce contamination. |
| GPQA Diamond | Scientific Reasoning | Expert MCQ set | Expert-written Google-proof questions in biology, chemistry and physics. |
| SimpleQA | General | Short factual questions | Short factual questions with stable ground-truth answers designed to expose hallucination. |
| MMMU / MMMU-Pro | General | 11.5K questions · 30+ disciplines | College-level questions requiring image and text reasoning. Multimodal context only. |
| MathVista | General | 6,141 examples | Math grounded in charts, diagrams and figures. Multimodal context only. |
Useful live catalogue views when you want one API key for several free routes. Always verify the current price marker before making a request.
Free · highest intelligence
Best first place to scan for general reasoning.
Free · highest coding
Start here for code generation and software work.
Free · highest agentic
Useful for tools, plans and multi-step tasks.
All free model routes
Check pricing and limits immediately before use.
Most popular overall
Useful context, but popularity is not a quality score.
Newest models
New models may appear before benchmark scores land.