Trending open-source projects ยท updated daily

Image Trending

๐Ÿ–ผ AI art, image generation, computer vision and image tooling โ€” sorted by stars, 24 projects.

Source: GitHub topic pages ยท 24 projects indexed

1

AUTOMATIC1111 / stable-diffusion-webui

Stable Diffusion web UI

Pythonโ˜… 164,941โ‘‚ 0
โ†’
2

Comfy-Org / ComfyUI

The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.

Pythonโ˜… 133,234โ‘‚ 0
โ†’
3

opencv / opencv

Open Source Computer Vision Library

C++โ˜… 90,839โ‘‚ 0
โ†’
4

Developer-Y / cs-video-courses

List of Computer Science courses with video lectures.

โ˜… 83,501โ‘‚ 0
โ†’
5

d2l-ai / d2l-zh

ใ€ŠๅŠจๆ‰‹ๅญฆๆทฑๅบฆๅญฆไน ใ€‹๏ผš้ขๅ‘ไธญๆ–‡่ฏป่€…ใ€่ƒฝ่ฟ่กŒใ€ๅฏ่ฎจ่ฎบใ€‚ไธญ่‹ฑๆ–‡็‰ˆ่ขซ70ๅคšไธชๅ›ฝๅฎถ็š„500ๅคšๆ‰€ๅคงๅญฆ็”จไบŽๆ•™ๅญฆใ€‚

Pythonโ˜… 80,678โ‘‚ 0
โ†’
6

unslothai / unsloth

Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.

Pythonโ˜… 76,179โ‘‚ 0
โ†’
7

microsoft / AI-For-Beginners

12 Weeks, 24 Lessons, AI for All!

Jupyter Notebookโ˜… 68,502โ‘‚ 0
โ†’
8

ultralytics / ultralytics

Ultralytics YOLO27, YOLO26, YOLO11, YOLOv8 โ€” object detection, instance segmentation, semantic segmentation, image classification, pose estimation, object tracking

Pythonโ˜… 61,610โ‘‚ 0
โ†’
9

calesthio / OpenMontage

World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent skill and production-knowledge files. Turn your AI coding assistant into a full video p

Pythonโ˜… 59,198โ‘‚ 0
โ†’
10

ultralytics / yolov5

Ultralytics YOLOv5 in PyTorch for object detection, instance segmentation, classification, training, and export.

Pythonโ˜… 58,012โ‘‚ 0
โ†’
11

rohitg00 / ai-engineering-from-scratch

Learn it. Build it. Ship it for others.

Pythonโ˜… 54,619โ‘‚ 0
โ†’
12

roboflow / supervision

We write your reusable computer vision tools. ๐Ÿ’œ

Pythonโ˜… 50,108โ‘‚ 0
โ†’
13

mudler / LocalAI

LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.

Goโ˜… 49,114โ‘‚ 0
โ†’
14

khoj-ai / khoj

Your AI second brain. Self-hostable. Get answers from the web or your docs. Build custom agents, schedule automations, do deep research. Turn any online or local LLM into your personal, autonomous AI

Pythonโ˜… 37,337โ‘‚ 0
โ†’
15

google-ai-edge / mediapipe

Cross-platform, customizable ML solutions for live and streaming media.

C++โ˜… 36,951โ‘‚ 0
โ†’
16

ashishpatel26 / 500-AI-Machine-learning-Deep-learning-Computer-vision-NLP-Projects-with-code

500 AI Machine learning Deep learning Computer vision NLP Projects with code

โ˜… 36,854โ‘‚ 0
โ†’
17

huggingface / diffusers

๐Ÿค— Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.

Pythonโ˜… 34,517โ‘‚ 0
โ†’
18

CMU-Perceptual-Computing-Lab / openpose

OpenPose: Real-time multi-person keypoint detection library for body, face, hands, and foot estimation

C++โ˜… 34,445โ‘‚ 0
โ†’
19

eugeneyan / applied-ml

๐Ÿ“š Papers & tech blogs by companies sharing their work on data science & machine learning in production.

โ˜… 30,128โ‘‚ 0
โ†’
20

d2l-ai / d2l-en

Interactive deep learning book with multi-framework code, math, and discussions. Adopted at 500 universities from 70 countries including Stanford, MIT, Harvard, and Cambridge.

Pythonโ˜… 29,608โ‘‚ 0
โ†’
21

HumanSignal / label-studio

Label Studio is a multi-type data labeling and annotation tool with standardized output format

TypeScriptโ˜… 28,267โ‘‚ 0
โ†’
22

invoke-ai / InvokeAI

Invoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest AI-driven technologies. The sol

Pythonโ˜… 28,219โ‘‚ 0
โ†’
23

ATH-MaaS / Pixelle-Video

๐Ÿš€ AI ๅ…จ่‡ชๅŠจ็Ÿญ่ง†้ข‘ๅผ•ๆ“Ž | AI Fully Automated Short Video Engine

Pythonโ˜… 28,120โ‘‚ 0
โ†’
24

op7418 / guizang-ppt-skill

AI-agent Skill for generating polished HTML slide decks: editorial magazine and Swiss layouts, image prompts, social covers, and a WebGL/low-power presentation runtime.

HTMLโ˜… 26,314โ‘‚ 0
โ†’