Open Source Generative AI Projects

Discover 80 open source Generative AI repositories from GitHub, each with an analysis of what it does, key features, use cases and alternatives. Generative AI projects here are most often combined with Python, Machine Learning and AI. Last updated October 4, 2026.

80 repositories · updated October 4, 2026

youtube-summarizer: Summarize YouTube Videos and Playlists

youtube-summarizer: Summarize YouTube Videos and Playlists

A Flask web app that turns YouTube video and playlist transcripts into AI-generated summaries, with optional audio playback. It suits individuals or small groups who want a self-hosted way to review video content.

PythonAIGenerative AI
Added Dec 17, 2025 View details
StreamDiffusion: Generate Images in Real Time with Diffusion

StreamDiffusion: Generate Images in Real Time with Diffusion

StreamDiffusion adapts diffusion pipelines for interactive image generation, with support for text-to-image and image-to-image workflows. It targets developers building responsive GPU-powered demos and applications.

PythonMachine LearningGenerative AI
Added Dec 13, 2025 View details
DragGAN: Edit GAN Images by Dragging Points

DragGAN: Edit GAN Images by Dragging Points

DragGAN is a research tool for interactive point-based editing of images generated by StyleGAN models. It lets users move image features by dragging points while the model updates the image, making controlled edits easier than conventional prompt-based workflows.

PythonGenerative AIMachine Learning
Added Dec 12, 2025 View details
SyncTalk: Generate Synchronized Talking-Head Videos

SyncTalk: Generate Synchronized Talking-Head Videos

SyncTalk is a CVPR 2024 system for generating talking-head videos from a person’s footage and audio. It targets synchronized lip movement, facial expression, and head pose, with workflows for training on a subject or running inference with provided models.

PythonComputer VisionDeep Learning
Added Dec 11, 2025 View details
MuseTalk: Generate Audio-Synced Talking-Head Videos

MuseTalk: Generate Audio-Synced Talking-Head Videos

MuseTalk creates lip-synced video from a source video or image and an audio clip using latent-space inpainting. It supports training and inference workflows, with real-time performance reported on a Tesla V100.

PythonAIMachine Learning
Added Dec 9, 2025 View details
LAM: Create Animatable 3D Gaussian Avatars from One Image

LAM: Create Animatable 3D Gaussian Avatars from One Image

LAM reconstructs a 3D Gaussian head avatar from a single image and supports animation and rendering across devices. It is aimed at developers building digital humans, especially interactive avatars, and requires model assets and a compatible compute environment for local use.

PythonAIMachine Learning
Added Dec 7, 2025 View details
GenerativeAICourse: Learn Generative AI Through Notebook Labs

GenerativeAICourse: Learn Generative AI Through Notebook Labs

A notebook-based course introducing generative AI and practical AI engineering, from LLM fundamentals to chatbots, RAG, agents, and MCP. It is aimed at learners who want guided explanations and hands-on Python exercises.

Generative AIEducationJupyter Notebook
Added Dec 4, 2025 View details
open-notebooklm: Turn PDFs Into Podcast Audio

open-notebooklm: Turn PDFs Into Podcast Audio

Open NotebookLM turns a PDF into an AI-generated podcast dialogue and MP3. It suits readers who want an audio-style overview of a document, and requires a Fireworks API key to run.

PythonAIPDF
Added Nov 29, 2025 View details
PPS-Ctrl: Translate Colonoscopy Images for Depth Estimation

PPS-Ctrl: Translate Colonoscopy Images for Depth Estimation

PPS-Ctrl explores controllable sim-to-real translation for colonoscopy images, using per-pixel shading maps to guide Stable Diffusion and ControlNet. The repository currently provides partial pseudocode rather than a complete, ready-to-run implementation.

PythonMachine LearningComputer Vision
Added Nov 28, 2025 View details
logocreator: Generate Logos and Brand Kits with AI

logocreator: Generate Logos and Brand Kits with AI

LogoCreator is a web app for generating and editing logos with AI, then exporting assets for a brand kit. It suits people who need a quick starting point for a new brand and want to run the app themselves with a Together AI API key.

TypeScriptAIGenerative AI
Added Nov 27, 2025 View details
GLM-4.5: Run Agentic Reasoning and Coding Models

GLM-4.5: Run Agentic Reasoning and Coding Models

GLM-4.5 is a family of open-weight mixture-of-experts models for reasoning, coding, and tool-using agents. It includes large and more compact variants, with deployment and fine-tuning guidance for teams with substantial GPU resources.

PythonLLMAI Agents
Added Nov 24, 2025 View details
txtinstruct: Build Instruction-Tuned Models from Your Data

txtinstruct: Build Instruction-Tuned Models from Your Data

txtinstruct is a Python framework for creating instruction-following datasets and training instruction-tuned models. It is intended for people who want greater control over dataset licensing or to incorporate their own data.

PythonMachine LearningLLM
Added Nov 23, 2025 View details

Related topics

OS
OSRepos

Analysis and discovery of open source repositories. Find interesting projects and follow their updates.

Monitor your website with YourWebsiteScore

OSRepos shares public repositories for knowledge and discovery only. Any installation, execution, configuration, or use of third-party repository code is at your own risk. Always review source code, dependencies, licenses, and security implications before running anything.

© 2025 OSRepos. Built with Nuxt 3 and lots of ❤️