AUTOMATIC1111 / stable-diffusion-webui
Stable Diffusion web UI
View AUTOMATIC1111/stable-diffusion-webuiThe image board collects GitHub projects that work with pictures: generation and diffusion models, computer vision, upscaling, background removal, OCR, annotation and the plumbing around them. Entries come from GitHub topic pages for image generation, stable diffusion and computer vision, ranked by total stars, so the list mixes well-known model repositories with newer tools that are still gaining ground. Every card shows the language, star and fork counts, and opens a detail page with the project description, license, last push date and README, plus related image projects drawn from the same board. If you are choosing a library to build on, check the last push date on the detail page before you commit — the board is about popularity, not maintenance.
Stable Diffusion web UI
View AUTOMATIC1111/stable-diffusion-webuiThe most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
View Comfy-Org/ComfyUIList of Computer Science courses with video lectures.
View Developer-Y/cs-video-coursesLocal UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.
View unslothai/unslothUltralytics YOLO27, YOLO26, YOLO11, YOLOv8 — object detection, instance segmentation, semantic segmentation, image classification, pose estimation, object tracking
View ultralytics/ultralyticsWorld's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent skill and production-knowledge files.
View calesthio/OpenMontageUltralytics YOLOv5 in PyTorch for object detection, instance segmentation, classification, training, and export.
View ultralytics/yolov5Learn it. Build it. Ship it for others.
View rohitg00/ai-engineering-from-scratchLocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.
View mudler/LocalAIYour AI second brain. Self-hostable. Get answers from the web or your docs. Build custom agents, schedule automations, do deep research.
View khoj-ai/khojCross-platform, customizable ML solutions for live and streaming media.
View google-ai-edge/mediapipe500 AI Machine learning Deep learning Computer vision NLP Projects with code
View ashishpatel26/500-AI-Machine-learning-Deep-learning-Computer-vision-NLP-Projects-with-code🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
View huggingface/diffusersOpenPose: Real-time multi-person keypoint detection library for body, face, hands, and foot estimation
View CMU-Perceptual-Computing-Lab/openpose📚 Papers & tech blogs by companies sharing their work on data science & machine learning in production.
View eugeneyan/applied-mlInteractive deep learning book with multi-framework code, math, and discussions. Adopted at 500 universities from 70 countries including Stanford, MIT, Harvard, and Cambridge.
View d2l-ai/d2l-enLabel Studio is a multi-type data labeling and annotation tool with standardized output format
View HumanSignal/label-studioInvoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest AI-driven technologies.
View invoke-ai/InvokeAI🚀 AI 全自动短视频引擎 | AI Fully Automated Short Video Engine
View ATH-MaaS/Pixelle-VideoAI-agent Skill for generating polished HTML slide decks: editorial magazine and Swiss layouts, image prompts, social covers, and a WebGL/low-power presentation runtime.
View op7418/guizang-ppt-skillImplementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in Pytorch
View lucidrains/vit-pytorchImage-to-Image Translation in PyTorch
View junyanz/pytorch-CycleGAN-and-pix2pixOriginal reference implementation of "3D Gaussian Splatting for Real-Time Radiance Field Rendering"
View graphdeco-inria/gaussian-splatting✅(已完结)超级全面的 深度学习 笔记【土堆 Pytorch】【李沐 动手学深度学习】【吴恩达 深度学习】【大飞 大模型Agent】
View AccumulateMore/CVImage inpainting tool powered by SOTA AI Model. Remove any unwanted object, defect, people from your pictures or erase and replace(powered by stable diffusion) any thing on your pictures.
View Sanster/IOPaint《明日方舟》小助手,全日常一键长草!| A one-click tool for the daily tasks of Arknights, supporting all clients.
View MaaAssistantArknights/MaaAssistantArknightsYOLOv4 / Scaled-YOLOv4 / YOLO - Neural Networks for Object Detection (Windows and Linux version of Darknet )
View AlexeyAB/darknet🤗 The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools
View huggingface/datasetsYC (S26) | Open Computer History | Record your screen continuously locally and provide context to your agents (Claude, Codex, Openclaw, Hermes, Runner...)
View screenpipe/screenpipe本项目将《动手学深度学习》(Dive into Deep Learning)原书中的MXNet实现改为PyTorch实现。
View ShusenTang/Dive-into-DL-PyTorchOpen source simulator for autonomous vehicles built on Unreal Engine / Unity, from Microsoft AI & Research
View microsoft/AirSimInstant neural graphics primitives: lightning fast NeRF and more
View NVlabs/instant-ngpGPT-Image-2 API and Prompts
View EvoLinkAI/awesome-gpt-image-2-API-and-PromptsComputer Vision Annotation Tool (CVAT) is a leading platform for building high-quality visual datasets for vision AI.
View cvat-ai/cvatA comprehensive list of pytorch related content on github,such as different models,implementations,helper libraries,tutorials etc.
View bharathgs/Awesome-pytorch-listImage annotation with Python. Supports polygon, rectangle, circle, line, point, and AI-assisted annotation.
View wkentaro/labelmestable diffusion webui colab
View camenduru/stable-diffusion-webui-colabToonflow 是开源一站式 AI 短剧创作工具,将小说、剧本快速转化为动画短剧。集成 AI 编剧、智能分镜、角色与视频生成,跨平台桌面端轻量部署,助力创作者低成本批量产出视觉内容。Toonflow is an open-source AI tool that turns stories and scripts into animated short dramas.
View HBAI-Ltd/Toonflow-appState-of-the-Art Deep Learning scripts organized by models - easy to train and deploy with reproducible accuracy and performance on enterprise-grade infrastructure.
View NVIDIA/DeepLearningExamplesA toolkit for making real world machine learning and data analysis applications in C++
View davisking/dlibOpen-source simulator for autonomous driving research.
View carla-simulator/carlaDiffusion Bee is the easiest way to run Stable Diffusion locally on your M1 Mac. Comes with a one-click installer. No dependencies or technical knowledge needed.
View divamgupta/diffusionbee-stable-diffusion-ui🍌 World's largest Nano Banana Pro prompt library — 10,000+ curated prompts with preview images, 16 languages. Google Gemini AI image generation. Free & open source.
View YouMind-OpenLab/awesome-nano-banana-pro-promptsAdvanced AI Explainability for computer vision. Support for CNNs, Vision Transformers, Classification, Object detection, Segmentation, Image similarity and more.
View jacobgil/pytorch-grad-camDrench yourself in Deep Learning, Reinforcement Learning, Machine Learning, Computer Vision, and NLP by learning from these exciting lectures!!
View kmario23/deep-learning-drizzleSoftware that can generate photos from paintings, turn horses into zebras, perform style transfer, and more.
View junyanz/CycleGANA MNIST-like fashion product database. Benchmark 👇
View zalandoresearch/fashion-mnistcvpr2024/cvpr2023/cvpr2022/cvpr2021/cvpr2020/cvpr2019/cvpr2018/cvpr2017 论文/代码/解读/直播合集,极市团队整理
View extreme-assistant/CVPR2024-Paper-Code-InterpretationA collaboration friendly studio for NeRFs
View nerfstudio-project/nerfstudioAI绘画资料合集(包含国内外可使用平台、使用教程、参数教程、部署教程、业界新闻等等) Stable diffusion、AnimateDiff、Stable Cascade 、Stable SDXL Turbo
View hua1995116/awesome-ai-paintingLow-code framework for building custom LLMs, neural networks, and other AI models
View ludwig-ai/ludwigSemantic segmentation models with 500+ pretrained convolutional and transformer-based backbones.
View qubvel-org/segmentation_models.pytorch中文小黑怪诞正文配图生成 Skill | 16:9 白底手绘 | 少量红橙蓝批注 | Codex Skill
View helloianneo/ian-xiaohei-illustrationsVisualize, query, and stream to train on multimodal robotics data.
View rerun-io/rerunOpenVINO™ is an open source toolkit for optimizing and deploying AI inference
View openvinotoolkit/openvinoConvert AI papers to GUI,Make it easy and convenient for everyone to use artificial intelligence technology。让每个人都简单方便的使用前沿人工智能技术
View Baiyuetribe/paper2guiImage-to-image translation with conditional adversarial nets
View phillipi/pix2pixPyTorch implementation of YOLOv3, YOLOv3-SPP, and YOLOv3-tiny for real-time object detection with training, validation, inference, and multi-format export.
View ultralytics/yolov3Streamlined interface for generating images with AI in Krita. Inpaint and outpaint with optional text prompt, no tweaking required.
View Acly/krita-ai-diffusionX-AnyLabeling: A lightweight, efficient, and unified cross-platform desktop application for annotating text, image, video, and multimodal data
View CVHub520/X-AnyLabelingopenFrameworks is a community-developed cross platform toolkit for creative coding in C++.
View openframeworks/openFrameworks🦙 LaMa Image Inpainting, Resolution-robust Large Mask Inpainting with Fourier Convolutions, WACV 2022
View advimman/lama🚀 World's largest GPT Image 2 prompt library, updated daily — 2000+ curated prompts with preview images, 16 languages.
View YouMind-OpenLab/awesome-gpt-image-2Best Practices, code samples, and documentation for Computer Vision.
View microsoft/computervision-recipesThe code for our newly accepted paper in Pattern Recognition 2020: "U^2-Net: Going Deeper with Nested U-Structure for Salient Object Detection."
View xuebinqin/U-2-NetTurn your two-bit doodles into fine artworks with deep neural networks, generate seamless textures from photos, transfer style from one image to another, perform example-based upscaling, but wait...
View alexjc/neural-doodleA collection of tutorials on state-of-the-art computer vision models and techniques. Explore everything from foundational architectures like ResNet to cutting-edge models like RF-DETR, YOLO11, SAM 3
View roboflow/notebooksRobust Video Matting in PyTorch, TensorFlow, TensorFlow.js, ONNX, CoreML!
View PeterL1n/RobustVideoMattingRF-DETR is a real-time object detection and segmentation model architecture developed by Roboflow, SOTA on COCO, designed for fine-tuning. [ICLR 2026]
View roboflow/rf-detrDeeplake is AI Data Runtime for Agents. It provides serverless postgres with a multimodal datalake, enabling scalable retrieval and training.
View activeloopai/deeplakeHello AI World guide to deploying deep-learning inference networks and deep vision primitives with TensorRT and NVIDIA Jetson.
View dusty-nv/jetson-inference深度学习面试宝典(含数学、机器学习、深度学习、计算机视觉、自然语言处理和SLAM等方向)
View amusi/Deep-Learning-Interview-BookText-to-3D & Image-to-3D & Mesh Exportation with NeRF + Diffusion.
View ashawkey/stable-dreamfusionMulti-Platform Package Manager for Stable Diffusion
View LykosAI/StabilityMatrixLab Materials for MIT 6.S191: Introduction to Deep Learning
View MITDeepLearning/introtodeeplearning[NeurIPS 2024 Best Paper Award][GPT beats diffusion🔥] [scaling laws in visual generation📈] Official impl. of "Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction".
View FoundationVision/VARLeading free and open-source face recognition system
View exadel-inc/CompreFaceAwesome curated collection of images and prompts generated by GPT-4o and gpt-image-1. Explore AI generated visuals created with ChatGPT and Sora
View jamez-bondos/awesome-gpt4o-images