MIT

awesome-freellm-apis

134+ free LLM APIs & AI API keys from 40+ providers. Google Gemini, NVIDIA NIM, Groq, OpenRouter & more. One-click setup for Claude Code, Cursor and Codex.

O

open-free-llm-api

Dernière activité 29 sept. 2026
open-free-llm-api/awesome-freellm-apis

3,4 k

étoiles

513

forks

5

issues ouvertes

free-llm-apifree-llm-modelsfree-llm-resourcesfree-modelsfreellmfreellmapi

Ce README est souvent en anglais.

awesome-free-llm-apis

506+ free LLM APIs from 30 providers — find, compare & configure free models in seconds.

🌐 Live at freellm.net — Browse models · Playground · Config generator · API keys

Provider Logos

🔄 Data refreshed daily from freellm.net — Last updated: 2026-10-11

🌐 English · 简体中文 · 繁體中文 · 日本語 · 한국어


Why This Exists

Finding a free LLM API shouldn't mean hunting through a dozen GitHub READMEs, signing up for five different platforms, or guessing which models still have a free tier.

This repo is a structured, machine-readable directory of every free LLM API — rate limits, context windows, one-click config snippets, and direct API key links. Updated daily.

Why this repo + freellm.net:

  • ✅ Always up-to-date — data refreshed daily via automated monitoring, not a 2-year-old static list
  • ✅ Credit card transparency — clearly shows which providers require a card, phone verification, or nothing at all
  • ✅ One-click configs — ready-to-copy snippets for Claude Code, Cursor, Codex, Aider, and 10+ more tools
  • ✅ Side-by-side comparison — compare context windows, rate limits, and modalities across providers instantly

How to Use — 3 Steps

  1. Pick a provider — see Provider Directory below. Start with Groq (no credit card, 30 RPM free).
  2. Get your API key — click any Get Key → link below, sign up (most need just an email), and copy your key. Takes < 1 minute.
  3. Plug it in — copy the base URL + model ID, paste into the Quick Start examples below.

Configuring a specific tool? Claude Code · Cursor · Codex · OpenHuman · OpenCode · OpenClaw — one-click configs at freellm.net/config/.

Quick Start — Use Any Free API in 30 Seconds

Never used an API before? Here's the simplest path: go to console.groq.com/keys, sign up with just an email (no credit card), copy your free key, and paste it into any example below. You'll be running in under a minute.

All providers below expose an OpenAI-compatible endpoint. Any tool that accepts baseURL + apiKey works — just swap the base URL and key.

Python (OpenAI SDK)

from openai import OpenAI

client = OpenAI(
    base_url="https://api.groq.com/openai/v1",  # free, no credit card
    api_key="GROQ_API_KEY",                     # get at console.groq.com/keys
)

response = client.chat.completions.create(
    model="llama-3.3-70b-versatile",            # see Best Models table below
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
# Groq free tier: 30 RPM, 14,400 RPD — generous for personal use

Codex CLI

export OPENAI_BASE_URL="https://api.groq.com/openai/v1"
export OPENAI_API_KEY="your-groq-key"          # get at console.groq.com/keys
codex --model "llama-3.3-70b-versatile"

Cursor

Settings → Models → Add Model
  Model name: llama-3.3-70b-versatile
  Base URL: https://api.groq.com/openai/v1
  API key: your-groq-key                       # get at console.groq.com/keys

Claude Code

# Claude Code needs an Anthropic-compatible API — use OpenRouter
export ANTHROPIC_BASE_URL="https://openrouter.ai/api"
export ANTHROPIC_AUTH_TOKEN="sk-or-v1-your-key"  # openrouter.ai/keys
export ANTHROPIC_API_KEY=""                       # must be empty
# Note: OpenRouter Anthropic models need $10 top-up (one-time)

Using Other Tools?

Most AI dev tools accept custom API endpoints — just point them at any provider above. Grab your free key, then:

More ready-to-copy configs at freellm.net/config/.

All providers, base URLs, and API key links are in the Quick Reference below.


Provider Directory

⚡ Permanent Free Tiers

These providers offer a permanently free tier — no credit card required for most.

Provider Free Models Credit Card? Max Context Modalities Get API Key
NVIDIA NIM 132 Phone verification 1M audio, embedding, image, pdf, reasoning, rerank, text, video, vision →
ModelScope 61 Registration 1M audio, image, reasoning, text, video, vision →
Cloudflare Workers AI 40 No 262K code, image, reasoning, text, video →
OpenCode Zen 34 Registration 1M audio, reasoning, vision →
LLM7.io 24 No 1M audio, code, image, pdf, reasoning, text, video, vision →
Kilo Code 22 No 1M audio, code, image, reasoning, text, video →
Google Gemini 19 No 1M audio, image, pdf, reasoning, text, video, vision →
Ollama Cloud 17 Registration 1M code, image, reasoning, text, video, vision →
Mistral AI 15 No 256K code, image, text →
OVHcloud AI Endpoints 15 Registration 262K audio, code, image, reasoning, text, video →
Groq 12 No 262K image, reasoning, text, video →
Cohere 12 No 256K image, text →
Aion Labs 11 Registration 131K text →
Hugging Face 9 No 131K code, image, text →
Z AI (Zhipu AI) 8 No 200K image, reasoning, text, video →
Cerebras 6 No 131K reasoning, text →
Agnes AI 5 Registration 256K image, reasoning, text, video, vision →
Alibaba Cloud Model Studio 5 Registration 1M code, image, text →
SiliconFlow 4 Registration 131K reasoning, text →
SambaNova 4 Registration 128K image, reasoning, text →
Cline 4 Registration 0 text →
xAI 3 Registration 2M text →
Chutes.ai 2 Registration 131K reasoning, text →
Glhf.chat 2 Registration 131K text →
Grok (xAI) 2 Registration 131K text →
AI21 Labs 2 Registration 256K text →
DeepSeek 2 Registration 128K text →
Nscale 2 Registration 128K text →
Nebius 1 Registration 128K text →

💰 Renewable Credits

Providers that periodically renew free credits.

Provider Free Models Credit Model Max Context Modalities Get API Key
OpenRouter 31 Free tier + $10 topup → 1K RPD 1M audio, code, decisions, embeddings, image, reasoning, rerank, speech, text, video →

Quick Reference — Base URLs & API Keys

Provider Base URL Get API Key Credit Card?
NVIDIA NIM https://integrate.api.nvidia.com/v1 Get Key → Phone verification
ModelScope https://api-inference.modelscope.cn/v1 Get Key → Registration
Cloudflare Workers AI https://api.cloudflare.com/client/v4/accounts/{account_id}/ai/run Get Key → No
OpenCode Zen https://opencode.ai/zen/v1 Get Key → Registration
OpenRouter https://openrouter.ai/api/v1 Get Key → Registration
LLM7.io https://api.llm7.io/v1 Get Key → No
Kilo Code https://api.kilo.ai/api/gateway Get Key → No
Google Gemini https://generativelanguage.googleapis.com/v1beta Get Key → No
Ollama Cloud https://ollama.com/api Get Key → Registration
Mistral AI https://api.mistral.ai/v1 Get Key → No
OVHcloud AI Endpoints https://oai.endpoints.kepler.ai.cloud.ovh.net/v1 Get Key → Registration
Groq https://api.groq.com/openai/v1 Get Key → No
Cohere https://api.cohere.com/v2 Get Key → No
Aion Labs https://api.aionlabs.ai/v1 Get Key → Registration
Hugging Face https://router.huggingface.co/v1 Get Key → No
Z AI (Zhipu AI) https://open.bigmodel.cn/api/paas/v4 Get Key → No
Cerebras https://api.cerebras.ai/v1 Get Key → No
Agnes AI https://apihub.agnes-ai.com/v1 Get Key → Registration
Alibaba Cloud Model Studio https://dashscope-intl.aliyuncs.com/compatible-mode/v1 Get Key → Registration
SiliconFlow https://api.siliconflow.cn/v1 Get Key → Registration
SambaNova https://api.sambanova.ai/v1 Get Key → Registration
Cline `` Get Key → Registration
xAI https://api.x.ai/v1 Get Key → Registration
Chutes.ai https://api.chutes.ai/v1 Get Key → Registration
Glhf.chat https://glhf.chat/api/openai/v1 Get Key → Registration
Grok (xAI) https://api.x.ai/v1 Get Key → Registration
AI21 Labs https://api.ai21.com/studio/v1 Get Key → Registration
DeepSeek https://api.deepseek.com/v1 Get Key → Registration
Nscale https://inference.api.nscale.com/v1 Get Key → Registration
Nebius https://api.studio.nebius.com/v1 Get Key → Registration

Best Free Models by Provider

Provider Best Free Model Model ID Max Context Rate Limit
NVIDIA NIM z-ai/glm-5.2 z-ai/glm-5.2 1M Up to 40 RPM
z-ai/glm-5.1 z-ai/glm-5.1 202K Up to 40 RPM
moonshotai/kimi-k2.6 moonshotai/kimi-k2.6 262K Up to 40 RPM
ModelScope MiniMax-M2.5-highspeed MiniMax/MiniMax-M2.5 204K See provider
Qwen/Qwen3.5-35B-A3B qwen-qwen3-5-35b-a3b 256K 2,000 RPD total; <=500 ..
Qwen/Qwen3.5-27B qwen-qwen3-5-27b 256K 2,000 RPD total; <=500 ..
Cloudflare Workers AI Mistral 7B @cf/mistral/mistral-7b-instruct-v0.1 32K See provider
Qwen 1.5 7B @cf/qwen/qwen1.5-7b-chat 32K See provider
@cf/meta/llama-3.3-70b-instruct-fp8-fast @cf/meta/llama-3.3-70b-instruct-fp8-fast 24K 10K neurons/day (shared)
OpenCode Zen big-pickle big-pickle 0
DeepSeek V4 Flash deepseek-v4-flash-free 1M
MiMo-V2.5 mimo-v2.5-free 1M
OpenRouter NVIDIA: Nemotron 3 Ultra (free) nvidia/nemotron-3-ultra-550b-a55b:free 1M See provider
Poolside: Laguna S 2.1 (free) poolside/laguna-s-2.1:free 262K See provider
NVIDIA: Nemotron 3.5 Lightning (free) nvidia/nemotron-3.5-lightning:free 1M See provider
LLM7.io gpt-oss:20b gpt-oss-20b 128K 60 RPM, 250 req/hr (fre..
mistral-Nemo-Instruct-2407 mistral-Nemo-Instruct-2407 128K 60 RPM, 250 req/hr (fre..
minimax-m2.7 minimax-m2-7 180K 60 RPM, 250 req/hr (fre..
Kilo Code nvidia/nemotron-3-ultra-550b-a55b:free nvidia/nemotron-3-ultra-550b-a55b:free 1M 200 req/hr
stepfun/step-3.7-flash:free stepfun/step-3.7-flash 262K 200 req/hr
nvidia/nemotron-3-super-120b-a12b:free nvidia/nemotron-3-super-120b-a12b:free 262K 200 req/hr
Google Gemini Gemini 3.8 Flash gemini-3.8-flash 1M —
Gemini 3.7 Flash gemini-3.7-flash 1M —
Gemini 3.6 Flash gemini-3.6-flash 1M 15 RPM, 1,500 RPD
Ollama Cloud deepseek-v4-pro deepseek-v4-pro 1M Session/weekly limits (..
deepseek-v4-flash deepseek-v4-flash 1M Session/weekly limits (..
minimax-m3 minimax-m3 512K Session/weekly limits (..
Mistral AI Mistral 7B open-mistral-7b 32K See provider
Mixtral 8x7B open-mixtral-8x7b 32K See provider
Mistral Medium 3.5 (128B) mistral-medium-3-5-128b 256K ~1 RPS, 500K TPM
OVHcloud AI Endpoints Qwen3.5-397B-A17B qwen3.5-397b-a17b 262K 2 RPM (anonymous)
Meta-Llama-3_3-70B-Instruct meta-llama-3_3-70b-instruct 131K 2 RPM (anonymous)
Qwen3.8-27B qwen3.8-27b 262K 2 RPM (anonymous)
Groq Moonshot Kimi K2 moonshotai/kimi-k2-instruct 131K See provider
Moonshot Kimi K2 0905 moonshotai/kimi-k2-instruct-0905 131K See provider
groq/compound groq/compound 131K 30 RPM, 250 RPD
Cohere Command A+ (218B) command-a-218b 128K 20 RPM
Command A (111B) command-a-111b 256K 20 RPM
Command R+ command-r 128K 20 RPM
Aion Labs aion-labs/aion-2.0 aion-labs-aion-2-0 128K 15 RPM, 20K TPD
aion-labs/aion-rp-llama-3.1-8b aion-labs-aion-rp-llama-3-1-8b 32K 15 RPM, 20K TPD
aion-labs/aion-3.0 aion-labs-aion-3-0 128K 15 RPM, 20K TPD
Hugging Face Meta-Llama-3.1-8B-Instruct meta-llama-3-1-8b-instruct 128K Credit-metered
gemma-3-4b-it google/gemma-3-4b-it 131K Credit-metered
phi-4 phi-4 16K Credit-metered
Z AI (Zhipu AI) GLM-4.7-Flash glm-4.7-flash 200K 1 concurrent request
GLM-4.5-Flash (retirement announced) glm-4-5-flash-retirement-announced 128K 1 concurrent request
GLM-4.6V-Flash glm-4.6v-flash 128K 1 concurrent request
Cerebras Llama 3.1 70B llama3.1-70b 131K See provider
zai-glm-4.7 (deprecated Aug 2026) zai-glm-4-7-deprecated-aug-2026 131K 5 RPM, 30K TPM, 1M TPD
zai-glm-4.7 zai-glm-4.7 128K 10 RPM, 100 RPD, 1M TPD
Agnes AI agnes-1.5-flash agnes-1.5-flash 256K 30 RPM
agnes-2.0-flash agnes-2.0-flash 256K 30 RPM
agnes-image-2.0-flash agnes-image-2.0-flash 4K 30 RPM (1K)
Alibaba Cloud Model Studio Qwen3-Max qwen3-max 128K Tiered by region
Qwen3-Plus qwen3-plus 1M Tiered by region
Qwen3-VL-Plus qwen3-vl-plus 128K Tiered by region
SiliconFlow Qwen/Qwen3-8B Qwen/Qwen3-8B 128K 1,000 RPM, 50,000 TPM
Abbreviation abbreviation 131K See provider
deepseek-ai/DeepSeek-R1-Distill-Qwen-7B deepseek-ai-deepseek-r1-distill-qwen-7b 131K 30 RPM, 60K TPM
SambaNova DeepSeek-V3.1 deepseek-v3-1 128K 20 RPM, 20 RPD, 200K TPD
DeepSeek-V3.2 (Preview) deepseek-v3-2-preview 128K 20 RPM, 20 RPD, 200K TPD
MiniMax-M2.7 minimax-m2-7 128K 20 RPM, 20 RPD, 200K TPD
Cline Muse Spark 1.3 Contributor cline-free/muse-spark-1.3-contributor 0 See provider
Mimo V2.6 Flash cline-free/mimo-v2.6-flash 0 See provider
Solar Mini 4 cline-free/solar-mini4 0 See provider
xAI grok-4.3 grok-4-3 1M Credit-based
grok-4.1-fast grok-4-1-fast 2M Credit-based
grok-3-mini grok-3-mini 131K Credit-based
Chutes.ai DeepSeek-R1 deepseek-ai/DeepSeek-R1 131K Community-powered, no h..
Llama 3.1 70B meta-llama/Meta-Llama-3.1-70B-Instruct 131K Community-powered, no h..
Glhf.chat Llama 3.1 70B meta-llama/Meta-Llama-3.1-70B-Instruct 131K Unlimited for free models
Mixtral 8x7B mistralai/Mixtral-8x7B-Instruct-v0.1 32K Unlimited for free models
Grok (xAI) Grok-2 grok-2 131K $25/month free credits,..
Grok-2 Mini grok-2-mini 131K $25/month free credits,..
AI21 Labs Jamba Large 1.7 jamba-large-1-7 256K 200 RPM, 10 RPS
Jamba Mini 2 jamba-mini-2 256K 200 RPM, 10 RPS
DeepSeek deepseek-chat (V3.2) deepseek-chat-v3-2 128K Dynamic
deepseek-reasoner (R1) deepseek-reasoner-r1 128K Dynamic
Nscale Llama-3.3-70B-Instruct llama-3-3-70b-instruct 128K Fair-use
DeepSeek-R1-Distill-Llama-70B deepseek-r1-distill-llama-70b 128K Fair-use
Nebius Qwen3-235B-A22B qwen3-235b-a22b 128K Tier-based

🖥️ Local / Self-Hosted (Unlimited, Private, Free Forever)

Tool Type Highlights
Ollama CLI + API 100+ models, GPU acceleration, OpenAI-compatible endpoint
LM Studio Desktop GUI Any GGUF model, built-in model browser, offline
llama.cpp C/C++ engine Runs any GGUF, minimal dependencies
GPT4All Desktop app CPU-only, no GPU required, open source
Jan.ai Desktop app Privacy-focused, 100% offline ChatGPT alternative
KoboldCpp Single executable Optimized for creative writing, GGUF

Top Free Models (by Weekly Usage)

Data from freellm.net, updated daily via API monitoring.

Model Provider Context Weekly Usage
NVIDIA: Nemotron 3 Ultra (free) OpenRouter 1M 6626B tokens
z-ai/glm-5.2 NVIDIA NIM 1M 2998B tokens
Poolside: Laguna S 2.1 (free) OpenRouter 262K 976B tokens
NVIDIA: Nemotron 3.5 Lightning (free) OpenRouter 1M 875B tokens
Poolside: Laguna M.1 (free) OpenRouter 262K 768B tokens
Dots Studio: Dots3-Note Preview (free) OpenRouter 512K 697B tokens
NVIDIA: Nemotron 3 Super (free) OpenRouter 262K 605B tokens
inclusionAI: Ling 3.1 Flash OpenRouter 262K 596B tokens
Apodex: Apodex 1.1 Mini (free) OpenRouter 262K 272B tokens
z-ai/glm-5.1 NVIDIA NIM 202K 158B tokens

Repository Structure

awesome-free-llm-apis/
├── README.md              ← Complete provider directory & code examples
├── code-examples/          ← Ready-to-use config snippets
│   ├── claude-code.md
│   ├── cursor.md
│   └── codex.md
└── LICENSE                 ← MIT

For the full structured dataset with 453 models and daily updates, visit freellm.net.


Contributing

We welcome contributions!

  • Add a missing free model — Open an issue or submit a PR
  • Fix inaccurate data — Rate limits change, providers graduate. PRs welcome
  • Add a config snippet — Have a working config for a tool we don't cover? Add it to code-examples/

Criteria for inclusion

A model belongs in this list if:

  1. The provider explicitly offers a free tier (not just a trial credit)
  2. The API is publicly accessible (no waitlist, closed beta, or reverse-engineering)
  3. For trial credits: clearly labeled and minimum $1 credit value

License

MIT © open-free-llm-api


Last updated: 2026-10-11

Projets similaires

List of Permanent Free LLM API (API Keys)

JavaScriptai-agentsanthropicawesome
Mmnfst
8,6 k étoiles856

免费大模型API,支持免费调用GPT、DeepSeek等主流模型,免费额度10000点,每日刷新!另付费价格最低官方1-2折!

apichatgptclaude
Cchatanywhere
43,4 k étoiles2,7 k

Directory of 34+ free LLM & AI APIs — permanent free tiers, trial credits, and no-card options. GPT-4o, Gemini, Claude, Llama, DeepSeek, Mistral & more. Synced daily from the live directory at free-llm.com.

JavaScriptchatgpt-alternativeclaude-4-6-opusclaude-api
Nnejib1
537 étoiles62