Open Source Machine Learning Projects
Discover 191 open source Machine Learning repositories from GitHub, each with an analysis of what it does, key features, use cases and alternatives. Machine Learning projects here are most often combined with Python, Deep Learning and LLM. Last updated October 4, 2026.
191 repositories · updated October 4, 2026

DragGAN: Edit GAN Images by Dragging Points
DragGAN is a research tool for interactive point-based editing of images generated by StyleGAN models. It lets users move image features by dragging points while the model updates the image, making controlled edits easier than conventional prompt-based workflows.

SyncTalk: Generate Synchronized Talking-Head Videos
SyncTalk is a CVPR 2024 system for generating talking-head videos from a person’s footage and audio. It targets synchronized lip movement, facial expression, and head pose, with workflows for training on a subject or running inference with provided models.

vggt: Reconstruct 3D Scenes from Images
VGGT is a feed-forward vision model that estimates cameras, depth, point maps, and point tracks from one or more images. It suits researchers and developers who need fast 3D scene reconstruction without a traditional multi-stage pipeline.

MuseTalk: Generate Audio-Synced Talking-Head Videos
MuseTalk creates lip-synced video from a source video or image and an audio clip using latent-space inpainting. It supports training and inference workflows, with real-time performance reported on a Tesla V100.

judgy: Estimate LLM Judge Success Rates with Bias Correction
judgy estimates a system’s true pass rate from human-labeled calibration data and LLM judge predictions. It corrects for judge errors and uses bootstrap resampling to produce a confidence interval, making it useful when evaluating larger unlabeled datasets.

LAM: Create Animatable 3D Gaussian Avatars from One Image
LAM reconstructs a 3D Gaussian head avatar from a single image and supports animation and rendering across devices. It is aimed at developers building digital humans, especially interactive avatars, and requires model assets and a compatible compute environment for local use.

whisper-web: Transcribe Speech in Your Browser
Whisper Web uses Transformers.js to run speech recognition in a browser. It suits people who want to try transcription without setting up a separate server, though browser support and performance may vary.

DataScienceInteractivePython: Learn Data Science Through Dashboards
Interactive Python dashboards let learners explore statistics, machine learning, and geostatistics by changing inputs and observing results. Designed for students and practitioners who benefit from hands-on experimentation.

multiresolution-time-series-transformer: Forecast Time Series at Multiple Scales
A PyTorch forecasting model that processes time series at several temporal resolutions, then fuses the representations to predict future values. It suits experimentation with multi-scale forecasting, with adaptations from the paper that users should account for.

PPS-Ctrl: Translate Colonoscopy Images for Depth Estimation
PPS-Ctrl explores controllable sim-to-real translation for colonoscopy images, using per-pixel shading maps to guide Stable Diffusion and ControlNet. The repository currently provides partial pseudocode rather than a complete, ready-to-run implementation.

subwiz: Discover Subdomains with a Lightweight GPT Model
subwiz uses a small transformer model to predict candidate subdomains from known subdomains, then can check whether predictions resolve. It is aimed at security teams and researchers who want an additional discovery step after passive enumeration.

ManimML: Animate Machine Learning Concepts with Manim
ManimML provides reusable Manim components for visualizing neural networks and common machine learning operations. It suits educators and creators who want to build explanatory animations in Python rather than implement each visualization from scratch.