Open Source AI Projects
Discover 275 open source AI repositories from GitHub, each with an analysis of what it does, key features, use cases and alternatives. AI projects here are most often combined with LLM, Python and TypeScript. Last updated October 4, 2026.
275 repositories · updated October 4, 2026

docling: Parse Documents for AI Workflows
Docling converts documents from formats such as PDF, Office files, images, and audio into structured representations and exports. It suits developers building document-processing and generative AI pipelines that need format coverage, OCR, or local execution.

LocalAI: Run AI Models on Your Own Hardware
LocalAI serves language, vision, audio, and image models through compatible APIs on local or distributed hardware. It suits developers and teams who want a flexible self-hosted AI service without depending on a GPU.

logto: Build Authentication and Authorization for Apps
Logto is an identity platform for SaaS and AI applications, built around OIDC and OAuth 2.1. It provides sign-in flows, multi-tenancy, enterprise SSO, and authorization capabilities for teams building user-facing apps and APIs.

ImageToolbox: Edit, Convert, and Process Images on Android
ImageToolbox is a feature-rich Android app for photo editing, image conversion, OCR, PDF tasks, and other media workflows. It suits users who want many image utilities in one place, with distribution through Google Play, F-Droid, and GitHub releases.

read-frog: Learn Languages While Reading the Web
Read Frog is a browser extension for language learners who want translation and AI-assisted study tools alongside everyday web reading. It supports multiple AI providers, selection and subtitle translation, text-to-speech, and flashcards.

Waifu2x-Extension-GUI: Upscale Images and Video
A Windows desktop app for enlarging and denoising images, GIFs, and video, with AI-based video frame interpolation. It combines multiple processing engines and supports AMD, Nvidia, and Intel GPUs.

mcp-ui: Build Interactive Interfaces for MCP Tools
mcp-ui provides SDKs for attaching web interfaces to Model Context Protocol tools and rendering them in compatible hosts. Use it when an MCP server needs interactive, sandboxed UI alongside tool results.

KBLaM: Add Knowledge Bases to Language Models
KBLaM is a research implementation for giving transformer language models access to external knowledge through learned adapters and special knowledge tokens. It is aimed at researchers testing knowledge-grounded answers without a separate retrieval module.

vanna: Turn Natural-Language Questions into SQL Insights
Vanna helps teams build chat interfaces that translate natural-language questions into SQL and return streamed tables, charts, and summaries. It is aimed at applications that need database-backed answers with user-aware permissions and an embedded web UI.

ai-baby-monitor: Local AI Video Monitoring for Baby Safety
A self-hosted baby-monitoring tool that checks webcam or RTSP video against natural-language safety rules using a local video LLM. It sends a quiet beep when a rule appears to be broken, with a dashboard for viewing the stream and model logs.

mcp-server-cloudflare: Connect AI Clients to Cloudflare Services
A collection of domain-specific Model Context Protocol servers that let compatible AI clients inspect and work with Cloudflare services through guided tools. Use them when you want product-focused interactions rather than broad API access through code execution.

courses: Learn Claude API and Prompting Techniques
Anthropic’s courses repository offers hands-on learning materials for using Claude through its API. It suits developers learning prompting, evaluations, and tool use, especially when they want guided examples before building Claude-powered workflows.