Open Source Computer Vision Projects

Discover 67 open source Computer Vision repositories from GitHub, each with an analysis of what it does, key features, use cases and alternatives. Computer Vision projects here are most often combined with Python, Machine Learning and Deep Learning. Last updated October 3, 2026.

67 repositories · updated October 3, 2026

LivePortrait: Animate Portraits from Images and Video

LivePortrait: Animate Portraits from Images and Video

LivePortrait is a PyTorch project for animating human and animal portraits using a driving video or motion template. It supports portrait video editing and offers command-line inference and a Gradio interface.

PythonAIMachine Learning
Added Oct 12, 2025 View details
Ollama-OCR: Extract Text from Images and PDFs with Vision Models

Ollama-OCR: Extract Text from Images and PDFs with Vision Models

Ollama-OCR uses vision-language models served by Ollama to extract text and structured content from images and PDFs. It offers a Python package for single or batch processing and a Streamlit interface for interactive use.

PythonAIComputer Vision
Added Oct 12, 2025 View details
Leffa: Generate Controllable Person Images

Leffa: Generate Controllable Person Images

Leffa is a diffusion-based framework for virtual try-on and pose transfer that aims to preserve fine-grained details from reference images. It is suited to researchers and developers building or evaluating controllable person-image generation systems.

PythonMachine LearningDeep Learning
Added Oct 12, 2025 View details
smolvlm-realtime-webcam: Analyze Webcam Video with SmolVLM

smolvlm-realtime-webcam: Analyze Webcam Video with SmolVLM

A browser-based demo sends webcam frames to a llama.cpp server running SmolVLM 500M for real-time visual analysis. It suits developers exploring local vision models and webcam workflows, with performance depending on the model server and hardware.

HTMLAIComputer Vision
Added Oct 11, 2025 View details
DeepFaceLive: Real-Time Face Swapping for Video

DeepFaceLive: Real-Time Face Swapping for Video

DeepFaceLive is a Windows desktop app for swapping faces from a webcam or video in real time, and for animating a still face image. It targets streamers and video-call users with a compatible GPU.

PythonMachine LearningComputer Vision
Added Oct 11, 2025 View details
jscanify: Detect and Straighten Paper in Images

jscanify: Detect and Straighten Paper in Images

jscanify is a JavaScript document-scanning library for detecting paper in images and correcting perspective. It uses OpenCV.js and supports browser and Node.js workflows, including live camera input.

JavaScriptComputer VisionDocument Scanner
Added Oct 11, 2025 View details
face-cropper: Crop and Align Faces in Images

face-cropper: Crop and Align Faces in Images

A Python tool that detects faces and crops them using MediaPipe FaceDetection and FaceMesh. It can optionally remove background pixels and correct face roll, making it useful for image preprocessing and lightweight real-time applications.

PythonComputer VisionMachine Learning
Added Oct 11, 2025 View details
Previous Page 6 Next

Related topics

OS
OSRepos

Analysis and discovery of open source repositories. Find interesting projects and follow their updates.

Monitor your website with YourWebsiteScore

OSRepos shares public repositories for knowledge and discovery only. Any installation, execution, configuration, or use of third-party repository code is at your own risk. Always review source code, dependencies, licenses, and security implications before running anything.

© 2025 OSRepos. Built with Nuxt 3 and lots of ❤️