freellmapi: Route LLM Requests Through One API

Summary
FreeLLMAPI is a self-hosted TypeScript router that combines free LLM provider accounts and custom OpenAI-compatible endpoints behind one API. It selects models, tracks quotas, and retries across providers, mainly for personal experimentation and development.
At a glance
- Language
- TypeScript
- License
- MIT
- Stars
- 30.3k
- Forks
- 4.2k
- Added to OSRepos
- June 27, 2026
- Last analyzed
- October 3, 2026
Topics
Click on any tag to explore related repositories
Use at your own risk
OSRepos shares public repositories for knowledge and discovery only. Any installation, execution, configuration, or use of code from these repositories is the user's own responsibility. Always review the repository, source code, dependencies, licenses, and security implications before running or installing anything. OSRepos is not responsible for issues, damages, or losses resulting from third-party repositories.
Overview
FreeLLMAPI is a self-hosted gateway for sending requests to free LLM provider tiers and custom OpenAI-compatible endpoints through a single API. It aims to reduce the work of configuring separate provider integrations and handling their different limits and failures.
The router supports common LLM and media request types, offers a dashboard for managing provider keys and routing, and works with compatible clients and coding agents. It is intended for personal experimentation, not as a production inference service.
Key Features
- OpenAI-compatible API with additional Anthropic, Gemini, and optional Ollama-compatible interfaces.
- Routes across configured providers using model selection, quota tracking, key rotation, and automatic fallback on rate limits or server errors.
- Supports custom OpenAI-compatible endpoints, including local or remote services.
- Provides a dashboard for keys, routing chains, a request playground, and usage analytics.
- Encrypts provider keys in SQLite and decrypts them in memory for requests.
- Includes client setup helpers and an MCP server for compatible tools and agents.
- Can run locally or in Docker, with desktop installers also described in the README.
Use Cases
- Developers experimenting with free provider tiers who want to avoid wiring each provider into an application separately.
- Users of OpenAI-compatible SDKs or coding agents who want to switch among configured models through one base URL.
- Local-model users who want to route requests to their own OpenAI-compatible endpoint alongside hosted providers.
- Individuals prototyping chat, embeddings, image, audio, or video workflows while accepting free-tier limits and variability.
Project Facts
- Language: TypeScript
- License: MIT
- Stars: 30.3k
- Forks: 4.2k
- Topics: none listed
- Archived: no
Getting Started
With Docker installed, run the installer:
curl -fsSL https://freellmapi.co/install.sh | bash
Open http://localhost:3001, add provider keys, then use the unified API key and local /v1 endpoint. See the README for other installation options and client configuration.
Alternatives
- claude-code-proxy: Targets Anthropic-compatible clients and translates their requests to configured backends, rather than aggregating free provider accounts behind an OpenAI-compatible API.
Considerations
Free provider tiers can have changing quotas, variable latency, and limited availability; the README explicitly warns that this is for personal experimentation and learning, not production. You need provider keys to use upstream services, and their terms still apply when requests are routed through FreeLLMAPI. The optional live model catalog is a paid service, while the router itself is MIT-licensed.
Source repository
Open the original repository on GitHub.
26 counted GitHub visits
Related repositories
Similar repositories that may be relevant next.

OrcaReplay: Time Travel for AI Agents, Debugging and Evaluation
October 2, 2026
OrcaReplay introduces "time travel" capabilities for AI agents, allowing developers to record, replay, fork, and debug any agent run with any model. It addresses the challenges of AI agent debugging by providing byte-for-byte reproducibility, offline analysis, and the ability to compare different models from specific checkpoints. This tool, built by the OrcaRouter.ai team, enhances observability and control over complex agent behaviors.

OpenMake LLM: Self-Hosted AI Workspace for Local and Open-Weight LLMs
October 1, 2026
OpenMake LLM is an open-source, self-hosted AI workspace for local and open-weight LLMs. It coordinates specialized models, autonomous agents, and tools for deep research and artifact generation. This platform supports vLLM, LiteLLM, and BYOK providers, offering a robust environment for managing AI workloads.

ZenNotes: Keyboard-First Markdown Notes with Vim, Diagrams, and MCP Integration
October 1, 2026
ZenNotes is a versatile, keyboard-first Markdown notes app designed for speed and flexibility. It stores notes as plain Markdown files, offering Vim-friendly editing, diagram support, and integration with MCP tools. Available as a desktop app (Electron) and a self-hosted web app, ZenNotes provides a powerful solution for organizing your thoughts.

lat.md: A Knowledge Graph for Your Codebase, Written in Markdown
September 26, 2026
lat.md is an innovative tool that transforms your codebase knowledge into an interconnected graph of markdown files. It helps both AI agents and human developers quickly understand project architecture, business logic, and design decisions. By integrating directly into your project, lat.md ensures documentation remains consistent and up-to-date.