Open Source Python Projects
Discover 606 open source Python repositories from GitHub, each with an analysis of what it does, key features, use cases and alternatives. Python projects here are most often combined with Library, Developer Tools and LLM. Last updated October 4, 2026.
606 repositories · updated October 4, 2026

RAGChecker: Diagnose Retrieval-Augmented Generation Systems
RAGChecker evaluates RAG pipelines with overall, retriever, and generator metrics. It is for developers and researchers who need to identify whether retrieval or generation is driving quality problems and guide targeted improvements.

rerankers: Use Diverse Reranking Models Through One Python API
rerankers provides a shared Python interface for reranking documents with cross-encoders, LLM-based methods, and hosted APIs. It suits developers building retrieval systems who want to compare or switch rerankers without adapting their application to each model's interface.

llm-compressor: Compress Models for vLLM Inference
LLM Compressor applies post-training quantization and related model transformations to prepare Hugging Face models for vLLM deployment. It suits teams seeking smaller or more inference-efficient checkpoints, with support for multiple formats and large-model workflows.

LightLLM: Serve Large Language Models with a Python Framework
LightLLM is a Python framework for LLM inference and serving, designed for scalable deployment and fast generation. It suits teams operating model-serving systems and researchers building on inference components.

torchchat: Run PyTorch LLMs on Desktop, Server, and Mobile
torchchat is a PyTorch codebase for running and interacting with language models locally through Python, native C++ runners, and mobile apps. It supports several execution and export paths, but is no longer under active development.

TensorRT-LLM: Optimize LLM Inference on NVIDIA GPUs
TensorRT-LLM is a Python framework and runtime for efficient LLM and visual-generation inference on NVIDIA GPUs. It suits teams deploying models on NVIDIA hardware that need optimized kernels and configurable single- or distributed-GPU execution.

docling: Convert Documents into Structured Content
Docling converts PDFs and many other document formats into structured representations and exports such as Markdown and JSON. It is suited to developers building document ingestion workflows for search, analytics, and generative AI, including local processing of sensitive files.

DataDreamer: Generate Synthetic Data and Train LLMs
DataDreamer is a Python library for building LLM workflows, generating synthetic datasets, and training or aligning models. It suits researchers and developers who want reproducible, resumable workflows across open-source and API-based models.

deepfabric: Generate and Evaluate Synthetic Training Data
DeepFabric generates domain-specific synthetic datasets for language-model training and agent evaluation. It combines topic planning, tool-use traces, validation, and evaluation in a Python library and CLI.

EasyInstruct: Generate, Select, and Prompt LLM Instructions
EasyInstruct is a Python framework for preparing instruction data and prompts for large language model research. It combines instruction generation and dataset selection tools with prompt and local-model execution modules.

LazyLLM: Build Multi-Agent LLM Applications
LazyLLM is a Python framework for assembling, testing, and deploying multi-agent LLM applications with low-code workflows. It suits developers who want to combine models, tools, and RAG components while iterating across local or online services.

ChatArena: Build Multi-Agent Language Game Environments
ChatArena is a Python framework for running language games with multiple LLM agents. It suits researchers and developers studying agent interaction, collaboration, and social behavior, but the project was deprecated in August 2025 and is no longer supported.