lilbee: Run Local AI and Search Your Files

Summary
lilbee is a local AI model manager and search engine for files, code, and crawled websites. It offers cited answers through a terminal app, CLI, MCP server, REST API, and Python library.
At a glance
- Language
- Python
- License
- MIT
- Stars
- 62
- Forks
- 7
- Added to OSRepos
- October 6, 2026
- Last analyzed
- October 6, 2026
Topics
Click on any tag to explore related repositories
Use at your own risk
OSRepos shares public repositories for knowledge and discovery only. Any installation, execution, configuration, or use of code from these repositories is the user's own responsibility. Always review the repository, source code, dependencies, licenses, and security implications before running or installing anything. OSRepos is not responsible for issues, damages, or losses resulting from third-party repositories.
Overview
lilbee combines a local model runner with a search engine for your files, notes, code, and crawled web pages. It indexes content and uses models to answer questions with citations that point back to source files and lines.
It is aimed at people who want a self-hosted alternative to cloud-first search and retrieval, and at coding agents that need to look up project material. Models and indexes can stay on your machine; a cloud provider is optional. You can use lilbee through a terminal interface, CLI, MCP server, REST API, or Python library.
Key Features
- Manages chat, embedding, vision, and reranking models, with support for local GPU backends and multi-GPU placement.
- Indexes documents and code, with citations linking answers to their sources.
- Crawls websites and adds their pages to a searchable local library.
- Exposes search and indexing to MCP-aware coding agents.
- Provides a terminal UI, CLI, REST API, and Python library.
- Can use existing Ollama or LM Studio setups, as well as optional hosted models.
Use Cases
- Researchers can index papers, manuals, and notes, then ask questions that link back to source passages.
- Developers can index a codebase and its documentation so a coding agent can retrieve relevant code with file and line citations.
- Teams or individuals can crawl documentation sites for a local, searchable copy that remains useful offline.
- Privacy-conscious users can run models and search personal files locally without relying on a cloud provider.
Getting Started
The README recommends installing a bundled build, then running:
lilbee self-check
lilbee
See the README for installation options, hardware-specific builds, and configuration. The project site also links to tutorials and the REST API reference.
Alternatives
- SurfSense: SurfSense focuses on desktop document research and study materials, while lilbee also manages local models and searches code and crawled websites through multiple interfaces.
Considerations
- The project describes itself as active beta software. Its interfaces, command names, and on-disk formats may change between pre-releases.
- A GPU is not required, but model performance and which models fit depend on available hardware and memory.
- Bundled builds are large, and the first launch has a one-time unpacking step. Developer installs through pip or uv require choosing and installing the separate model-engine extra for the relevant hardware.
- Local operation is the default, but using a hosted model sends questions and relevant retrieved snippets to the provider you configure.
Found this useful?
Share it with someone who would like lilbee.
Source repository
Open the original repository on GitHub.
Related repositories
Similar repositories that may be relevant next.

seclab-taskflow-agent: Define AI Workflows in YAML
October 6, 2026
A Python framework and CLI for building multi-agent workflows from YAML, with MCP tools and configurable model backends. It is aimed at security research, code auditing, and other repeatable agent tasks.

AgentSec: Audit AI Agent Workflows for Security Risks
October 6, 2026
AgentSec statically analyzes AI agent workflows for excessive permissions and paths from untrusted input to dangerous capabilities. It is aimed at developers and security teams reviewing supported agent frameworks before deployment or as part of CI.

omnigent: Orchestrate AI Coding Agents Across Harnesses
October 5, 2026
Omnigent provides a shared orchestration layer for AI coding agents, with policies, sandboxing, and team collaboration. It suits developers who want to combine agent runtimes and access sessions across devices without tying workflows to one harness.

agentevals: Evaluate AI Agents from OpenTelemetry Traces
October 4, 2026
agentevals scores AI agent behavior from existing OpenTelemetry traces, without rerunning agents or making extra model calls. It suits teams building instrumented agents that need local evaluation, golden-set checks, or CI quality gates.