Open Source Generative AI Projects
Discover 80 open source Generative AI repositories from GitHub, each with an analysis of what it does, key features, use cases and alternatives. Generative AI projects here are most often combined with Python, Machine Learning and AI. Last updated October 4, 2026.
80 repositories · updated October 4, 2026

youtube-summarizer: Summarize YouTube Videos and Playlists
A Flask web app that turns YouTube video and playlist transcripts into AI-generated summaries, with optional audio playback. It suits individuals or small groups who want a self-hosted way to review video content.

StreamDiffusion: Generate Images in Real Time with Diffusion
StreamDiffusion adapts diffusion pipelines for interactive image generation, with support for text-to-image and image-to-image workflows. It targets developers building responsive GPU-powered demos and applications.

DragGAN: Edit GAN Images by Dragging Points
DragGAN is a research tool for interactive point-based editing of images generated by StyleGAN models. It lets users move image features by dragging points while the model updates the image, making controlled edits easier than conventional prompt-based workflows.

SyncTalk: Generate Synchronized Talking-Head Videos
SyncTalk is a CVPR 2024 system for generating talking-head videos from a person’s footage and audio. It targets synchronized lip movement, facial expression, and head pose, with workflows for training on a subject or running inference with provided models.

MuseTalk: Generate Audio-Synced Talking-Head Videos
MuseTalk creates lip-synced video from a source video or image and an audio clip using latent-space inpainting. It supports training and inference workflows, with real-time performance reported on a Tesla V100.

LAM: Create Animatable 3D Gaussian Avatars from One Image
LAM reconstructs a 3D Gaussian head avatar from a single image and supports animation and rendering across devices. It is aimed at developers building digital humans, especially interactive avatars, and requires model assets and a compatible compute environment for local use.

GenerativeAICourse: Learn Generative AI Through Notebook Labs
A notebook-based course introducing generative AI and practical AI engineering, from LLM fundamentals to chatbots, RAG, agents, and MCP. It is aimed at learners who want guided explanations and hands-on Python exercises.

open-notebooklm: Turn PDFs Into Podcast Audio
Open NotebookLM turns a PDF into an AI-generated podcast dialogue and MP3. It suits readers who want an audio-style overview of a document, and requires a Fireworks API key to run.

PPS-Ctrl: Translate Colonoscopy Images for Depth Estimation
PPS-Ctrl explores controllable sim-to-real translation for colonoscopy images, using per-pixel shading maps to guide Stable Diffusion and ControlNet. The repository currently provides partial pseudocode rather than a complete, ready-to-run implementation.

logocreator: Generate Logos and Brand Kits with AI
LogoCreator is a web app for generating and editing logos with AI, then exporting assets for a brand kit. It suits people who need a quick starting point for a new brand and want to run the app themselves with a Together AI API key.

GLM-4.5: Run Agentic Reasoning and Coding Models
GLM-4.5 is a family of open-weight mixture-of-experts models for reasoning, coding, and tool-using agents. It includes large and more compact variants, with deployment and fine-tuning guidance for teams with substantial GPU resources.

txtinstruct: Build Instruction-Tuned Models from Your Data
txtinstruct is a Python framework for creating instruction-following datasets and training instruction-tuned models. It is intended for people who want greater control over dataset licensing or to incorporate their own data.