Open Source Machine Learning Projects

Discover 191 open source Machine Learning repositories from GitHub, each with an analysis of what it does, key features, use cases and alternatives. Machine Learning projects here are most often combined with Python, Deep Learning and LLM. Last updated October 4, 2026.

191 repositories · updated October 4, 2026

rag-from-scratch: Learn Retrieval-Augmented Generation Step by Step

rag-from-scratch: Learn Retrieval-Augmented Generation Step by Step

A Jupyter notebook series that teaches the building blocks of retrieval-augmented generation, from indexing and retrieval to generation. It is aimed at learners who want to understand RAG concepts through an educational progression rather than adopt a ready-made application.

RAGJupyter NotebookLLM
Added Apr 30, 2026 View details
StreamingKokoroJS: Generate Speech Locally in Your Browser

StreamingKokoroJS: Generate Speech Locally in Your Browser

StreamingKokoroJS turns text into speech in the browser with the Kokoro-82M model. It streams audio locally, with WebGPU acceleration when available and a WebAssembly fallback, without sending text to a server.

JavaScriptAIText To Speech
Added Apr 26, 2026 View details
chatterbox: Generate Speech from Text with Voice Cloning

chatterbox: Generate Speech from Text with Voice Cloning

Chatterbox is Resemble AI’s Python text-to-speech model family, with English and multilingual options, voice cloning, and expressive speech controls. It suits developers building voice experiences who can manage model inference and its compute requirements.

PythonAIMachine Learning
Added Apr 19, 2026 View details
Kimi-k1.5: Train Multimodal LLMs with Reinforcement Learning

Kimi-k1.5: Train Multimodal LLMs with Reinforcement Learning

Kimi k1.5 is a research project describing reinforcement learning methods for training long-context, multimodal language models. It is aimed at researchers studying LLM reasoning and training, rather than users looking for a ready-to-install model.

AILLMMachine Learning
Added Apr 17, 2026 View details
judges: Evaluate LLM Outputs with Reusable AI Judges

judges: Evaluate LLM Outputs with Reusable AI Judges

Databricks judges is a Python library for evaluating language model outputs with reusable LLM-based classifiers and graders. Use its research-backed judges, combine evaluations with a jury, or build a custom judge for your task.

PythonLLMEvaluation
Added Apr 14, 2026 View details
CompreFace: Run Self-Hosted Face Recognition APIs

CompreFace: Run Self-Hosted Face Recognition APIs

CompreFace is a Docker-based face analysis service with REST APIs for recognition, verification, detection, and related tasks. It suits teams that need to integrate face processing into applications while keeping deployment on their own infrastructure.

Computer VisionJavaDocker
Added Apr 12, 2026 View details
co-tracker: Track Points Across Video Frames

co-tracker: Track Points Across Video Frames

CoTracker tracks selected or grid-sampled points through video using a transformer-based model. It offers offline and online inference, pretrained checkpoints, and tools for evaluation and training.

Computer VisionMachine LearningDeep Learning
Added Apr 11, 2026 View details
pyAudioAnalysis: Extract Features and Analyze Audio in Python

pyAudioAnalysis: Extract Features and Analyze Audio in Python

pyAudioAnalysis is a Python library for extracting audio features and building classification, detection, and segmentation workflows. It suits researchers and developers who want an established toolkit for analyzing audio files with machine-learning methods.

PythonMachine LearningLibrary
Added Apr 6, 2026 View details
notebooks: Learn and Apply Computer Vision Models

notebooks: Learn and Apply Computer Vision Models

Roboflow notebooks is a hands-on tutorial collection for computer vision, covering model training, inference, detection, segmentation, and related tasks. Use it to explore techniques and run examples in hosted notebook environments.

Computer VisionMachine LearningDeep Learning
Added Apr 6, 2026 View details
Spark-TTS: Generate Speech and Clone Voices from Text

Spark-TTS: Generate Speech and Clone Voices from Text

Spark-TTS is a PyTorch inference project for bilingual text-to-speech and zero-shot voice cloning. It uses a Qwen2.5-based model to generate speech and supports adjustable voice characteristics.

PythonText To SpeechMachine Learning
Added Apr 5, 2026 View details
TRELLIS: Generate 3D Assets from Text or Images

TRELLIS: Generate 3D Assets from Text or Images

TRELLIS is a research model and toolkit for generating 3D assets from text or images. Its structured latent representation can produce meshes, 3D Gaussians, and radiance fields, but local use requires a compatible NVIDIA GPU and a substantial setup.

PythonGenerative AIComputer Vision
Added Apr 1, 2026 View details
AudioSep: Separate Sounds from Natural-Language Descriptions

AudioSep: Separate Sounds from Natural-Language Descriptions

AudioSep is a Python foundation model for separating sounds from audio based on natural-language descriptions. It supports open-domain tasks such as isolating events, instruments, or speech, with inference, fine-tuning, and evaluation workflows.

PythonMachine LearningDeep Learning
Added Mar 30, 2026 View details

Related topics

OS
OSRepos

Analysis and discovery of open source repositories. Find interesting projects and follow their updates.

Monitor your website with YourWebsiteScore

OSRepos shares public repositories for knowledge and discovery only. Any installation, execution, configuration, or use of third-party repository code is at your own risk. Always review source code, dependencies, licenses, and security implications before running anything.

© 2025 OSRepos. Built with Nuxt 3 and lots of ❤️