List of Permanent Free LLM API (API Keys)
-
Updated
Aug 21, 2026 - JavaScript
List of Permanent Free LLM API (API Keys)
Streamline your workflow with Lynkr, a CLI tool that acts as an HTTP proxy for efficient code interactions using Claude Code CLI.
Use your Claude and ChatGPT subscriptions in Cursor, Cline, Aider, Claude Code and the Agent SDK — at subscription pricing, not per-token API bills. One local Anthropic + OpenAI-compatible endpoint: either plan answers either wire shape, auto-failover when one hits its limit, multi-seat pooling, live Claude Code drift tracking.
Token Bank — the local LLM gateway that sits between your AI agents and every provider. Know where tokens go · Spend less with smart routing to Ollama, Groq, GitHub Models · Earn by sharing idle quota on a community P2P network. One-click onboarding for Cursor, Claude Code, Codex CLI, Gemini CLI — no agent changes. Full trace, seamless model swap
Open-source AI backend control plane for model routing, agent tools, OpenAI-compatible APIs, and private workspaces.
Local multi-model fusion gateway for Codex and OpenAI-compatible API relays
Self-hosted AI gateway & LLM proxy: zero-downtime failover, MOA Fusion, token-saving compression. One OpenAI-compatible endpoint for Claude Code, Cursor, Codex, Cline + 40+ providers.
List of Permanent Free LLM API (API Keys)
The coding agent that pays for the frontier model only when your tests prove it is needed. Verification-gated model cascade with an OpenAI-compatible gateway. Zero deps, MIT.
Codex Desktop model router: when your 5-hour quota hits 0%, type /dsk to continue the same thread on DeepSeek - with context, MCP tools, and shell intact. No restart, no new chat; /gpt switches back. Unofficial.
Smart model router for OpenClaw agents
Route the easy work your expensive Claude model keeps repeating down to haiku/sonnet — post-hoc session analysis, no realtime router, no extra LLM calls.
Self-hosted AI gateway. Accepts OpenAI, Anthropic, Gemini, Codex, Cursor, Kiro and Ollama request formats, translates each into whatever the upstream provider expects, and falls back across models and accounts. Chat, embeddings, audio, images, video, rerank and OCR, with a dashboard for credentials, quota and cost.
CLI-first OpenAI-compatible LLM router/proxy for coding agents that picks the cheapest fast-enough model and proves speed, savings, and prompt-free privacy.
Local OpenCode router for MiniMax, GLM, Claude, and Codex with adaptive routing and configurable tool policies
OpenAI-compatible multi-LLM ensemble proxy with Chairman fusion. Drop-in for any chat completions client.
FreeDSH — free-first multi-model routing, automatic fallback, safe updates and pt-BR support for DeepSeek Harness.
Budget-LLM pool router for DeepSeek Harness: ledger, health tracking, Thompson-sampling failover across free-tier models. 穷鬼路由器
Local model router for Codex Desktop: native GPT + DeepSeek, Kimi, Qwen and more, with multi-key quota rotation and free OpenCode Zen models.
To associate your repository with the llm-router topic, visit your repo's landing page and select "manage topics."