freellmapi: Route LLM Requests Through One API

Summary
FreeLLMAPI is a self-hosted TypeScript router that combines free LLM provider accounts and custom OpenAI-compatible endpoints behind one API. It selects models, tracks quotas, and retries across providers, mainly for personal experimentation and development.
At a glance
- Language
- TypeScript
- License
- MIT
- Stars
- 30.3k
- Forks
- 4.2k
- Added to OSRepos
- June 27, 2026
- Last analyzed
- October 3, 2026
Topics
Click on any tag to explore related repositories
Use at your own risk
OSRepos shares public repositories for knowledge and discovery only. Any installation, execution, configuration, or use of code from these repositories is the user's own responsibility. Always review the repository, source code, dependencies, licenses, and security implications before running or installing anything. OSRepos is not responsible for issues, damages, or losses resulting from third-party repositories.
Overview
FreeLLMAPI is a self-hosted gateway for sending requests to free LLM provider tiers and custom OpenAI-compatible endpoints through a single API. It aims to reduce the work of configuring separate provider integrations and handling their different limits and failures.
The router supports common LLM and media request types, offers a dashboard for managing provider keys and routing, and works with compatible clients and coding agents. It is intended for personal experimentation, not as a production inference service.
Key Features
- OpenAI-compatible API with additional Anthropic, Gemini, and optional Ollama-compatible interfaces.
- Routes across configured providers using model selection, quota tracking, key rotation, and automatic fallback on rate limits or server errors.
- Supports custom OpenAI-compatible endpoints, including local or remote services.
- Provides a dashboard for keys, routing chains, a request playground, and usage analytics.
- Encrypts provider keys in SQLite and decrypts them in memory for requests.
- Includes client setup helpers and an MCP server for compatible tools and agents.
- Can run locally or in Docker, with desktop installers also described in the README.
Use Cases
- Developers experimenting with free provider tiers who want to avoid wiring each provider into an application separately.
- Users of OpenAI-compatible SDKs or coding agents who want to switch among configured models through one base URL.
- Local-model users who want to route requests to their own OpenAI-compatible endpoint alongside hosted providers.
- Individuals prototyping chat, embeddings, image, audio, or video workflows while accepting free-tier limits and variability.
Project Facts
- Language: TypeScript
- License: MIT
- Stars: 30.3k
- Forks: 4.2k
- Topics: none listed
- Archived: no
Getting Started
With Docker installed, run the installer:
curl -fsSL https://freellmapi.co/install.sh | bash
Open http://localhost:3001, add provider keys, then use the unified API key and local /v1 endpoint. See the README for other installation options and client configuration.
Alternatives
- claude-code-proxy: Targets Anthropic-compatible clients and translates their requests to configured backends, rather than aggregating free provider accounts behind an OpenAI-compatible API.
Considerations
Free provider tiers can have changing quotas, variable latency, and limited availability; the README explicitly warns that this is for personal experimentation and learning, not production. You need provider keys to use upstream services, and their terms still apply when requests are routed through FreeLLMAPI. The optional live model catalog is a paid service, while the router itself is MIT-licensed.
Source repository
Open the original repository on GitHub.
26 counted GitHub visits
Related repositories
Similar repositories that may be relevant next.

OrcaReplay: Record and Replay AI Agent Runs
October 2, 2026
OrcaReplay records AI agent runs so developers can inspect what happened, replay interactions offline, and fork a run from a checkpoint onto another model. It is a TypeScript CLI for debugging and comparing agent behavior without modifying the agent.

openmake_llm: Coordinate AI Models, Agents, and Tools
October 1, 2026
OpenMake is a self-hosted AI workspace that coordinates chat models, agents, and tools in one place. It suits people who want multi-step AI workflows with local or open-weight models, while keeping infrastructure and model choices under their control.

zennotes: Manage Markdown Notes with Keyboard-First Tools
October 1, 2026
ZenNotes is a local-first Markdown notes app for people who prefer plain files, keyboard navigation, and Vim motions. It runs as an Electron desktop app or a self-hosted web app, with search, diagrams, CLI, and MCP integrations.

Kun: Local-First AI Agent Workspace
September 27, 2026
Kun is a local-first workspace for using AI agents across coding, writing, design, research, and automation. Its desktop GUI and terminal TUI share one runtime, keeping tasks and approvals connected.