Trending open-source projects · updated daily

AI Image GitHub Trending

The image board collects GitHub projects that work with pictures: generation and diffusion models, computer vision, upscaling, background removal, OCR, annotation and the plumbing around them. Entries come from GitHub topic pages for image generation, stable diffusion and computer vision, ranked by total stars, so the list mixes well-known model repositories with newer tools that are still gaining ground. Every card shows the language, star and fork counts, and opens a detail page with the project description, license, last push date and README, plus related image projects drawn from the same board. If you are choosing a library to build on, check the last push date on the detail page before you commit — the board is about popularity, not maintenance.

Trending AI Image Projects

1
2

Comfy-Org / ComfyUI

The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.

Python★ 133,352⑂ 0
View Comfy-Org/ComfyUI
3

opencv / opencv

Open Source Computer Vision Library

C++★ 90,849⑂ 0
View opencv/opencv
4

Developer-Y / cs-video-courses

List of Computer Science courses with video lectures.

★ 83,507⑂ 0
View Developer-Y/cs-video-courses
5

d2l-ai / d2l-zh

《动手学深度学习》:面向中文读者、能运行、可讨论。中英文版被70多个国家的500多所大学用于教学。

Python★ 80,708⑂ 0
View d2l-ai/d2l-zh
6

unslothai / unsloth

Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.

Python★ 76,204⑂ 0
View unslothai/unsloth
7

microsoft / AI-For-Beginners

12 Weeks, 24 Lessons, AI for All!

Jupyter Notebook★ 68,520⑂ 0
View microsoft/AI-For-Beginners
8

ultralytics / ultralytics

Ultralytics YOLO27, YOLO26, YOLO11, YOLOv8 — object detection, instance segmentation, semantic segmentation, image classification, pose estimation, object tracking

Python★ 61,632⑂ 0
View ultralytics/ultralytics
9

calesthio / OpenMontage

World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent skill and production-knowledge files.

Python★ 59,289⑂ 0
View calesthio/OpenMontage
10

ultralytics / yolov5

Ultralytics YOLOv5 in PyTorch for object detection, instance segmentation, classification, training, and export.

Python★ 58,015⑂ 0
View ultralytics/yolov5
11

rohitg00 / ai-engineering-from-scratch

Learn it. Build it. Ship it for others.

Python★ 54,696⑂ 0
View rohitg00/ai-engineering-from-scratch
12

roboflow / supervision

We write your reusable computer vision tools. 💜

Python★ 50,230⑂ 0
View roboflow/supervision
13

mudler / LocalAI

LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.

Go★ 49,126⑂ 0
View mudler/LocalAI
14

khoj-ai / khoj

Your AI second brain. Self-hostable. Get answers from the web or your docs. Build custom agents, schedule automations, do deep research.

Python★ 37,349⑂ 0
View khoj-ai/khoj
15

google-ai-edge / mediapipe

Cross-platform, customizable ML solutions for live and streaming media.

C++★ 36,965⑂ 0
View google-ai-edge/mediapipe
16
17

huggingface / diffusers

🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.

Python★ 34,525⑂ 0
View huggingface/diffusers
18

CMU-Perceptual-Computing-Lab / openpose

OpenPose: Real-time multi-person keypoint detection library for body, face, hands, and foot estimation

C++★ 34,447⑂ 0
View CMU-Perceptual-Computing-Lab/openpose
19

eugeneyan / applied-ml

📚 Papers & tech blogs by companies sharing their work on data science & machine learning in production.

★ 30,128⑂ 0
View eugeneyan/applied-ml
20

d2l-ai / d2l-en

Interactive deep learning book with multi-framework code, math, and discussions. Adopted at 500 universities from 70 countries including Stanford, MIT, Harvard, and Cambridge.

Python★ 29,612⑂ 0
View d2l-ai/d2l-en
21

HumanSignal / label-studio

Label Studio is a multi-type data labeling and annotation tool with standardized output format

TypeScript★ 28,269⑂ 0
View HumanSignal/label-studio
22

invoke-ai / InvokeAI

Invoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest AI-driven technologies.

Python★ 28,224⑂ 0
View invoke-ai/InvokeAI
23

ATH-MaaS / Pixelle-Video

🚀 AI 全自动短视频引擎 | AI Fully Automated Short Video Engine

Python★ 28,129⑂ 0
View ATH-MaaS/Pixelle-Video
24

op7418 / guizang-ppt-skill

AI-agent Skill for generating polished HTML slide decks: editorial magazine and Swiss layouts, image prompts, social covers, and a WebGL/low-power presentation runtime.

HTML★ 26,360⑂ 0
View op7418/guizang-ppt-skill
25

lucidrains / vit-pytorch

Implementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in Pytorch

Python★ 25,506⑂ 0
View lucidrains/vit-pytorch
26

junyanz / pytorch-CycleGAN-and-pix2pix

Image-to-Image Translation in PyTorch

Python★ 25,245⑂ 0
View junyanz/pytorch-CycleGAN-and-pix2pix
27

graphdeco-inria / gaussian-splatting

Original reference implementation of "3D Gaussian Splatting for Real-Time Radiance Field Rendering"

Python★ 23,867⑂ 0
View graphdeco-inria/gaussian-splatting
28

AccumulateMore / CV

✅(已完结)超级全面的 深度学习 笔记【土堆 Pytorch】【李沐 动手学深度学习】【吴恩达 深度学习】【大飞 大模型Agent】

Jupyter Notebook★ 23,710⑂ 0
View AccumulateMore/CV
29

Sanster / IOPaint

Image inpainting tool powered by SOTA AI Model. Remove any unwanted object, defect, people from your pictures or erase and replace(powered by stable diffusion) any thing on your pictures.

Python★ 23,329⑂ 0
View Sanster/IOPaint
30

MaaAssistantArknights / MaaAssistantArknights

《明日方舟》小助手,全日常一键长草!| A one-click tool for the daily tasks of Arknights, supporting all clients.

C++★ 23,280⑂ 0
View MaaAssistantArknights/MaaAssistantArknights
31

spmallick / learnopencv

Learn OpenCV : C++ and Python Examples

Jupyter Notebook★ 23,151⑂ 0
View spmallick/learnopencv
32

amusi / CVPR2026-Papers-with-Code

CVPR 2026 论文和开源项目合集

★ 22,834⑂ 0
View amusi/CVPR2026-Papers-with-Code
33

AlexeyAB / darknet

YOLOv4 / Scaled-YOLOv4 / YOLO - Neural Networks for Object Detection (Windows and Linux version of Darknet )

C★ 22,146⑂ 0
View AlexeyAB/darknet
34

huggingface / datasets

🤗 The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools

Python★ 21,973⑂ 0
View huggingface/datasets
35

screenpipe / screenpipe

YC (S26) | Open Computer History | Record your screen continuously locally and provide context to your agents (Claude, Codex, Openclaw, Hermes, Runner...)

Rust★ 21,588⑂ 0
View screenpipe/screenpipe
36

ShusenTang / Dive-into-DL-PyTorch

本项目将《动手学深度学习》(Dive into Deep Learning)原书中的MXNet实现改为PyTorch实现。

Jupyter Notebook★ 19,499⑂ 0
View ShusenTang/Dive-into-DL-PyTorch
37

microsoft / AirSim

Open source simulator for autonomous vehicles built on Unreal Engine / Unity, from Microsoft AI & Research

C++★ 18,480⑂ 0
View microsoft/AirSim
38

pytorch / vision

Datasets, Transforms and Models specific to Computer Vision

Python★ 17,911⑂ 0
View pytorch/vision
39

NVlabs / instant-ngp

Instant neural graphics primitives: lightning fast NeRF and more

Cuda★ 17,550⑂ 0
View NVlabs/instant-ngp
40
41

cvat-ai / cvat

Computer Vision Annotation Tool (CVAT) is a leading platform for building high-quality visual datasets for vision AI.

Python★ 16,722⑂ 0
View cvat-ai/cvat
42

bharathgs / Awesome-pytorch-list

A comprehensive list of pytorch related content on github,such as different models,implementations,helper libraries,tutorials etc.

★ 16,668⑂ 0
View bharathgs/Awesome-pytorch-list
43

wkentaro / labelme

Image annotation with Python. Supports polygon, rectangle, circle, line, point, and AI-assisted annotation.

Python★ 16,171⑂ 0
View wkentaro/labelme
44

camenduru / stable-diffusion-webui-colab

stable diffusion webui colab

Jupyter Notebook★ 15,912⑂ 0
View camenduru/stable-diffusion-webui-colab
45

HBAI-Ltd / Toonflow-app

Toonflow 是开源一站式 AI 短剧创作工具,将小说、剧本快速转化为动画短剧。集成 AI 编剧、智能分镜、角色与视频生成,跨平台桌面端轻量部署,助力创作者低成本批量产出视觉内容。Toonflow is an open-source AI tool that turns stories and scripts into animated short dramas.

TypeScript★ 15,642⑂ 0
View HBAI-Ltd/Toonflow-app
46

virgili0 / Virgilio

Your new Mentor for Data Science E-Learning.

Jupyter Notebook★ 14,983⑂ 0
View virgili0/Virgilio
47

NVIDIA / DeepLearningExamples

State-of-the-Art Deep Learning scripts organized by models - easy to train and deploy with reproducible accuracy and performance on enterprise-grade infrastructure.

Jupyter Notebook★ 14,843⑂ 0
View NVIDIA/DeepLearningExamples
48

davisking / dlib

A toolkit for making real world machine learning and data analysis applications in C++

C++★ 14,439⑂ 0
View davisking/dlib
49

carla-simulator / carla

Open-source simulator for autonomous driving research.

C++★ 14,401⑂ 0
View carla-simulator/carla
50

davidsandberg / facenet

Face recognition using Tensorflow

Python★ 14,349⑂ 0
View davidsandberg/facenet
51

mlfoundations / open_clip

An open source implementation of CLIP.

Python★ 14,144⑂ 0
View mlfoundations/open_clip
52

vercel / satori

Enlightened library to convert HTML and CSS to SVG

TypeScript★ 13,945⑂ 0
View vercel/satori
53

divamgupta / diffusionbee-stable-diffusion-ui

Diffusion Bee is the easiest way to run Stable Diffusion locally on your M1 Mac. Comes with a one-click installer. No dependencies or technical knowledge needed.

JavaScript★ 13,586⑂ 0
View divamgupta/diffusionbee-stable-diffusion-ui
54

YouMind-OpenLab / awesome-nano-banana-pro-prompts

🍌 World's largest Nano Banana Pro prompt library — 10,000+ curated prompts with preview images, 16 languages. Google Gemini AI image generation. Free & open source.

TypeScript★ 13,426⑂ 0
View YouMind-OpenLab/awesome-nano-banana-pro-prompts
55

jacobgil / pytorch-grad-cam

Advanced AI Explainability for computer vision. Support for CNNs, Vision Transformers, Classification, Object detection, Segmentation, Image similarity and more.

Python★ 12,971⑂ 0
View jacobgil/pytorch-grad-cam
56

alicevision / Meshroom

Node-based Visual Programming Toolbox

Python★ 12,962⑂ 0
View alicevision/Meshroom
57

kmario23 / deep-learning-drizzle

Drench yourself in Deep Learning, Reinforcement Learning, Machine Learning, Computer Vision, and NLP by learning from these exciting lectures!!

HTML★ 12,944⑂ 0
View kmario23/deep-learning-drizzle
58

junyanz / CycleGAN

Software that can generate photos from paintings, turn horses into zebras, perform style transfer, and more.

Lua★ 12,871⑂ 0
View junyanz/CycleGAN
59

zalandoresearch / fashion-mnist

A MNIST-like fashion product database. Benchmark 👇

Python★ 12,826⑂ 0
View zalandoresearch/fashion-mnist
60

colmap / colmap

COLMAP - Structure-from-Motion and Multi-View Stereo

C++★ 12,721⑂ 0
View colmap/colmap
61

extreme-assistant / CVPR2024-Paper-Code-Interpretation

cvpr2024/cvpr2023/cvpr2022/cvpr2021/cvpr2020/cvpr2019/cvpr2018/cvpr2017 论文/代码/解读/直播合集,极市团队整理

★ 12,460⑂ 0
View extreme-assistant/CVPR2024-Paper-Code-Interpretation
62

nerfstudio-project / nerfstudio

A collaboration friendly studio for NeRFs

Python★ 11,998⑂ 0
View nerfstudio-project/nerfstudio
63

hua1995116 / awesome-ai-painting

AI绘画资料合集(包含国内外可使用平台、使用教程、参数教程、部署教程、业界新闻等等) Stable diffusion、AnimateDiff、Stable Cascade 、Stable SDXL Turbo

★ 11,785⑂ 0
View hua1995116/awesome-ai-painting
64

ludwig-ai / ludwig

Low-code framework for building custom LLMs, neural networks, and other AI models

Python★ 11,756⑂ 0
View ludwig-ai/ludwig
65

qubvel-org / segmentation_models.pytorch

Semantic segmentation models with 500+ pretrained convolutional and transformer-based backbones.

Python★ 11,734⑂ 0
View qubvel-org/segmentation_models.pytorch
66

helloianneo / ian-xiaohei-illustrations

中文小黑怪诞正文配图生成 Skill | 16:9 白底手绘 | 少量红橙蓝批注 | Codex Skill

★ 11,690⑂ 0
View helloianneo/ian-xiaohei-illustrations
67

rerun-io / rerun

Visualize, query, and stream to train on multimodal robotics data.

Rust★ 11,452⑂ 0
View rerun-io/rerun
68

kornia / kornia

🐍 Geometric Computer Vision Library for Spatial AI

Python★ 11,357⑂ 0
View kornia/kornia
69

PointCloudLibrary / pcl

Point Cloud Library (PCL)

C++★ 11,119⑂ 0
View PointCloudLibrary/pcl
70

voxel51 / fiftyone

Refine high-quality datasets and visual AI models

TypeScript★ 11,089⑂ 0
View voxel51/fiftyone
71

openvinotoolkit / openvino

OpenVINO™ is an open source toolkit for optimizing and deploying AI inference

C++★ 10,857⑂ 0
View openvinotoolkit/openvino
72

Baiyuetribe / paper2gui

Convert AI papers to GUI,Make it easy and convenient for everyone to use artificial intelligence technology。让每个人都简单方便的使用前沿人工智能技术

Jupyter Notebook★ 10,704⑂ 0
View Baiyuetribe/paper2gui
73

phillipi / pix2pix

Image-to-image translation with conditional adversarial nets

Lua★ 10,658⑂ 0
View phillipi/pix2pix
74

autogluon / autogluon

Fast and Accurate ML in 3 Lines of Code

Python★ 10,658⑂ 0
View autogluon/autogluon
75

ultralytics / yolov3

PyTorch implementation of YOLOv3, YOLOv3-SPP, and YOLOv3-tiny for real-time object detection with training, validation, inference, and multi-format export.

Python★ 10,606⑂ 0
View ultralytics/yolov3
76

Acly / krita-ai-diffusion

Streamlined interface for generating images with AI in Krita. Inpaint and outpaint with optional text prompt, no tweaking required.

Python★ 10,590⑂ 0
View Acly/krita-ai-diffusion
77

esimov / caire

Content aware image resize library

Go★ 10,466⑂ 0
View esimov/caire
78

CVHub520 / X-AnyLabeling

X-AnyLabeling: A lightweight, efficient, and unified cross-platform desktop application for annotating text, image, video, and multimodal data

Python★ 10,440⑂ 0
View CVHub520/X-AnyLabeling
79

openframeworks / openFrameworks

openFrameworks is a community-developed cross platform toolkit for creative coding in C++.

C++★ 10,426⑂ 0
View openframeworks/openFrameworks
80

advimman / lama

🦙 LaMa Image Inpainting, Resolution-robust Large Mask Inpainting with Fourier Convolutions, WACV 2022

Jupyter Notebook★ 10,260⑂ 0
View advimman/lama
81

YouMind-OpenLab / awesome-gpt-image-2

🚀 World's largest GPT Image 2 prompt library, updated daily — 2000+ curated prompts with preview images, 16 languages.

TypeScript★ 9,885⑂ 0
View YouMind-OpenLab/awesome-gpt-image-2
82

microsoft / computervision-recipes

Best Practices, code samples, and documentation for Computer Vision.

Jupyter Notebook★ 9,880⑂ 0
View microsoft/computervision-recipes
83

xuebinqin / U-2-Net

The code for our newly accepted paper in Pattern Recognition 2020: "U^2-Net: Going Deeper with Nested U-Structure for Salient Object Detection."

Python★ 9,859⑂ 0
View xuebinqin/U-2-Net
84

alexjc / neural-doodle

Turn your two-bit doodles into fine artworks with deep neural networks, generate seamless textures from photos, transfer style from one image to another, perform example-based upscaling, but wait...

Python★ 9,850⑂ 0
View alexjc/neural-doodle
85

roboflow / notebooks

A collection of tutorials on state-of-the-art computer vision models and techniques. Explore everything from foundational architectures like ResNet to cutting-edge models like RF-DETR, YOLO11, SAM 3

Jupyter Notebook★ 9,672⑂ 0
View roboflow/notebooks
86

PeterL1n / RobustVideoMatting

Robust Video Matting in PyTorch, TensorFlow, TensorFlow.js, ONNX, CoreML!

Python★ 9,524⑂ 0
View PeterL1n/RobustVideoMatting
87

roboflow / rf-detr

RF-DETR is a real-time object detection and segmentation model architecture developed by Roboflow, SOTA on COCO, designed for fine-tuning. [ICLR 2026]

Python★ 9,487⑂ 0
View roboflow/rf-detr
88

activeloopai / deeplake

Deeplake is AI Data Runtime for Agents. It provides serverless postgres with a multimodal datalake, enabling scalable retrieval and training.

C++★ 9,235⑂ 0
View activeloopai/deeplake
89

Stability-AI / StableStudio

Community interface for generative AI

TypeScript★ 9,046⑂ 0
View Stability-AI/StableStudio
90
91

dusty-nv / jetson-inference

Hello AI World guide to deploying deep-learning inference networks and deep vision primitives with TensorRT and NVIDIA Jetson.

C++★ 8,992⑂ 0
View dusty-nv/jetson-inference
92

amusi / Deep-Learning-Interview-Book

深度学习面试宝典(含数学、机器学习、深度学习、计算机视觉、自然语言处理和SLAM等方向)

★ 8,915⑂ 0
View amusi/Deep-Learning-Interview-Book
93

ashawkey / stable-dreamfusion

Text-to-3D & Image-to-3D & Mesh Exportation with NeRF + Diffusion.

Python★ 8,857⑂ 0
View ashawkey/stable-dreamfusion
94

LykosAI / StabilityMatrix

Multi-Platform Package Manager for Stable Diffusion

C#★ 8,784⑂ 0
View LykosAI/StabilityMatrix
95

MITDeepLearning / introtodeeplearning

Lab Materials for MIT 6.S191: Introduction to Deep Learning

Jupyter Notebook★ 8,783⑂ 0
View MITDeepLearning/introtodeeplearning
96

FoundationVision / VAR

[NeurIPS 2024 Best Paper Award][GPT beats diffusion🔥] [scaling laws in visual generation📈] Official impl. of "Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction".

Jupyter Notebook★ 8,732⑂ 0
View FoundationVision/VAR
97

bytedeco / javacv

Java interface to OpenCV, FFmpeg, and more

Java★ 8,337⑂ 0
View bytedeco/javacv
98

exadel-inc / CompreFace

Leading free and open-source face recognition system

Java★ 8,310⑂ 0
View exadel-inc/CompreFace
99

carson-katri / dream-textures

Stable Diffusion built-in to Blender

Python★ 8,202⑂ 0
View carson-katri/dream-textures
100

jamez-bondos / awesome-gpt4o-images

Awesome curated collection of images and prompts generated by GPT-4o and gpt-image-1. Explore AI generated visuals created with ChatGPT and Sora

JavaScript★ 8,146⑂ 0
View jamez-bondos/awesome-gpt4o-images

Computer Vision GitHub Projects

Image Generation Repositories

Popular Image Processing Tools

More GitHub Trending Projects

1

alibaba / open-code-review

Go★ 28,800⑂ 2,060▲ 2,756 stars
2

debpalash / VoiceStudio

Python★ 31,078⑂ 3,708▲ 2,072 stars
3

JustVugg / colibri

C★ 33,938⑂ 3,550▲ 2,026 stars
4

cloudflare / security-audit-skill

JavaScript★ 5,154⑂ 316▲ 1,434 stars
5

tt-a1i / archify

JavaScript★ 63,676⑂ 4,224▲ 1,373 stars
6

abue-ammar / tinycast

Swift★ 4,787⑂ 228▲ 1,076 stars

Related GitHub Trending Lists