Open Source Generative AI Projects
Discover 80 open source Generative AI repositories from GitHub, each with an analysis of what it does, key features, use cases and alternatives. Generative AI projects here are most often combined with Python, Machine Learning and AI. Last updated October 4, 2026.
80 repositories · updated October 4, 2026

KBLaM: Add Knowledge Bases to Language Models
KBLaM is a research implementation for giving transformer language models access to external knowledge through learned adapters and special knowledge tokens. It is aimed at researchers testing knowledge-grounded answers without a separate retrieval module.

logo-ai: Generate Custom Logos with AI
LogoAI is a web app for creating custom logos from user-selected styles, colors, sizes, and quality settings. It suits individuals and small teams who want to explore logo concepts quickly, with generation powered by Nebius AI.

theatre: Create and Edit Motion Graphics for the Web
Theatre.js is a TypeScript animation library with visual tools for creating detailed motion graphics on the web. Use it to animate 3D scenes, HTML and SVG, or other JavaScript values, either programmatically or through its editor.

poml: Structure and Render Prompts for Language Models
POML is a markup language and toolkit for building structured, reusable prompts for large language models. It combines templating, data components, styling, and development tools for teams managing prompts in code.

lmql: Program LLMs with Python-Like Constraints
LMQL combines Python-style control flow with language-model queries and constraints on generated text. It suits developers building structured, model-driven workflows who need more control than prompt templates provide.

PartCrafter: Generate Structured 3D Meshes from Images
PartCrafter generates part-separated 3D objects and scenes from a single RGB image using compositional latent diffusion. It suits researchers and developers exploring image-to-3D generation who have access to a CUDA-enabled GPU.

csm: Generate Conversational Speech from Text and Audio
CSM is Sesame’s speech-generation model, producing audio from text and optional conversation context. It suits developers building voice experiences who can run large models on a CUDA-compatible GPU and provide the required Hugging Face checkpoints.
HunyuanVideo-Avatar: Create Audio-Driven Character Videos
HunyuanVideo-Avatar generates dynamic, emotion-controllable videos of one or more characters from avatar images and audio. It is aimed at creators and researchers who need expressive talking-avatar or dialogue video generation and have access to compatible NVIDIA GPU hardware.

clarity-upscaler: Enhance and Upscale Images with AI
Clarity-Upscaler is a Python image-to-image project for increasing image resolution and enhancing details with Stable Diffusion workflows. It suits users comfortable with Cog or image-generation tools who want a configurable alternative to hosted upscaling services.

TextMachina: Build Datasets for Machine-Generated Text Tasks
TextMachina is a Python framework for generating and exploring datasets for machine-generated text detection, attribution, and boundary tasks. It helps researchers and developers combine text sources, language models, and configurable generation pipelines while checking for dataset quality and bias.

anse: Chat with Multiple AI Models in One Interface
Anse is a TypeScript web app for chatting with AI services and generating images through a unified interface. It suits people who want configurable model providers, locally stored chat sessions, and deployment options for their own instance.

CineScale: Generate High-Resolution Video Without Fine-Tuning
CineScale is an inference framework for generating high-resolution video with pretrained diffusion models, without fine-tuning. It targets researchers and practitioners who want to upscale generation beyond a model’s training resolution, including 4K workflows.