Trending AI Image Projects
The image board collects the AI projects that work with pictures: diffusion and generation models, editing and upscaling tools, background removal, OCR, annotation and the computer-vision stack underneath them. Entries come from GitHub topic pages for AI image generation, text-to-image and computer vision, ranked by stars, so landmark model repositories sit beside newer tools that are still gaining ground. Every card shows language, stars and forks and opens a page with the project description, license, last push date, README and related image projects. If you are picking a library to build on, check the last push date on the detail page before committing — this board measures popularity, not maintenance.
Trending AI Image Projects
Developer-Y / cs-video-courses
List of Computer Science courses with video lectures.
View Developer-Y/cs-video-coursesultralytics / ultralytics
Ultralytics YOLO27, YOLO26, YOLO11, YOLOv8 — object detection, instance segmentation, semantic segmentation, image classification, pose estimation, object tracking
View ultralytics/ultralyticsultralytics / yolov5
Ultralytics YOLOv5 in PyTorch for object detection, instance segmentation, classification, training, and export.
View ultralytics/yolov5rohitg00 / ai-engineering-from-scratch
Learn it. Build it. Ship it for others.
View rohitg00/ai-engineering-from-scratchgoogle-ai-edge / mediapipe
Cross-platform, customizable ML solutions for live and streaming media.
View google-ai-edge/mediapipeashishpatel26 / 500-AI-Machine-learning-Deep-learning-Computer-vision-NLP-Projects-with-code
500 AI Machine learning Deep learning Computer vision NLP Projects with code
View ashishpatel26/500-AI-Machine-learning-Deep-learning-Computer-vision-NLP-Projects-with-codeCMU-Perceptual-Computing-Lab / openpose
OpenPose: Real-time multi-person keypoint detection library for body, face, hands, and foot estimation
View CMU-Perceptual-Computing-Lab/openposeeugeneyan / applied-ml
📚 Papers & tech blogs by companies sharing their work on data science & machine learning in production.
View eugeneyan/applied-mld2l-ai / d2l-en
Interactive deep learning book with multi-framework code, math, and discussions. Adopted at 500 universities from 70 countries including Stanford, MIT, Harvard, and Cambridge.
View d2l-ai/d2l-enAnil-matcha / Open-Generative-AI
Unrestricted Open-source alternative to AI video platforms — Free AI image & video generation studio with 600+ models (Flux, Midjourney, Kling, Sora, Veo). No content filters.
View Anil-matcha/Open-Generative-AIHumanSignal / label-studio
Label Studio is a multi-type data labeling and annotation tool with standardized output format
View HumanSignal/label-studiolucidrains / vit-pytorch
Implementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in Pytorch
View lucidrains/vit-pytorchjunyanz / pytorch-CycleGAN-and-pix2pix
Image-to-Image Translation in PyTorch
View junyanz/pytorch-CycleGAN-and-pix2pixgraphdeco-inria / gaussian-splatting
Original reference implementation of "3D Gaussian Splatting for Real-Time Radiance Field Rendering"
View graphdeco-inria/gaussian-splattingAccumulateMore / CV
✅(已完结)超级全面的 深度学习 笔记【土堆 Pytorch】【李沐 动手学深度学习】【吴恩达 深度学习】【大飞 大模型Agent】
View AccumulateMore/CVMaaAssistantArknights / MaaAssistantArknights
《明日方舟》小助手,全日常一键长草!| A one-click tool for the daily tasks of Arknights, supporting all clients.
View MaaAssistantArknights/MaaAssistantArknightsAlexeyAB / darknet
YOLOv4 / Scaled-YOLOv4 / YOLO - Neural Networks for Object Detection (Windows and Linux version of Darknet )
View AlexeyAB/darknethuggingface / datasets
🤗 The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools
View huggingface/datasetsscreenpipe / screenpipe
YC (S26) | Open Computer History | Record your screen continuously locally and provide context to your agents (Claude, Codex, Openclaw, Hermes, Runner...)
View screenpipe/screenpipeShusenTang / Dive-into-DL-PyTorch
本项目将《动手学深度学习》(Dive into Deep Learning)原书中的MXNet实现改为PyTorch实现。
View ShusenTang/Dive-into-DL-PyTorchmicrosoft / AirSim
Open source simulator for autonomous vehicles built on Unreal Engine / Unity, from Microsoft AI & Research
View microsoft/AirSimNVlabs / instant-ngp
Instant neural graphics primitives: lightning fast NeRF and more
View NVlabs/instant-ngpEvoLinkAI / awesome-gpt-image-2-API-and-Prompts
GPT-Image-2 API and Prompts
View EvoLinkAI/awesome-gpt-image-2-API-and-Promptscvat-ai / cvat
Computer Vision Annotation Tool (CVAT) is a leading platform for building high-quality visual datasets for vision AI.
View cvat-ai/cvatbharathgs / Awesome-pytorch-list
A comprehensive list of pytorch related content on github,such as different models,implementations,helper libraries,tutorials etc.
View bharathgs/Awesome-pytorch-listwkentaro / labelme
Image annotation with Python. Supports polygon, rectangle, circle, line, point, and AI-assisted annotation.
View wkentaro/labelmeNVIDIA / DeepLearningExamples
State-of-the-Art Deep Learning scripts organized by models - easy to train and deploy with reproducible accuracy and performance on enterprise-grade infrastructure.
View NVIDIA/DeepLearningExamplesdavisking / dlib
A toolkit for making real world machine learning and data analysis applications in C++
View davisking/dlibcarla-simulator / carla
Open-source simulator for autonomous driving research.
View carla-simulator/carlajacobgil / pytorch-grad-cam
Advanced AI Explainability for computer vision. Support for CNNs, Vision Transformers, Classification, Object detection, Segmentation, Image similarity and more.
View jacobgil/pytorch-grad-camkmario23 / deep-learning-drizzle
Drench yourself in Deep Learning, Reinforcement Learning, Machine Learning, Computer Vision, and NLP by learning from these exciting lectures!!
View kmario23/deep-learning-drizzlejunyanz / CycleGAN
Software that can generate photos from paintings, turn horses into zebras, perform style transfer, and more.
View junyanz/CycleGANzalandoresearch / fashion-mnist
A MNIST-like fashion product database. Benchmark 👇
View zalandoresearch/fashion-mnistextreme-assistant / CVPR2024-Paper-Code-Interpretation
cvpr2024/cvpr2023/cvpr2022/cvpr2021/cvpr2020/cvpr2019/cvpr2018/cvpr2017 论文/代码/解读/直播合集,极市团队整理
View extreme-assistant/CVPR2024-Paper-Code-Interpretationnerfstudio-project / nerfstudio
A collaboration friendly studio for NeRFs
View nerfstudio-project/nerfstudioludwig-ai / ludwig
Low-code framework for building custom LLMs, neural networks, and other AI models
View ludwig-ai/ludwigqubvel-org / segmentation_models.pytorch
Semantic segmentation models with 500+ pretrained convolutional and transformer-based backbones.
View qubvel-org/segmentation_models.pytorchrerun-io / rerun
Visualize, query, and stream to train on multimodal robotics data.
View rerun-io/rerunlucidrains / DALLE2-pytorch
Implementation of DALL-E 2, OpenAI's updated text-to-image synthesis neural network, in Pytorch
View lucidrains/DALLE2-pytorchopenvinotoolkit / openvino
OpenVINO™ is an open source toolkit for optimizing and deploying AI inference
View openvinotoolkit/openvinophillipi / pix2pix
Image-to-image translation with conditional adversarial nets
View phillipi/pix2pixultralytics / yolov3
PyTorch implementation of YOLOv3, YOLOv3-SPP, and YOLOv3-tiny for real-time object detection with training, validation, inference, and multi-format export.
View ultralytics/yolov3CVHub520 / X-AnyLabeling
X-AnyLabeling: A lightweight, efficient, and unified cross-platform desktop application for annotating text, image, video, and multimodal data
View CVHub520/X-AnyLabelingopenframeworks / openFrameworks
openFrameworks is a community-developed cross platform toolkit for creative coding in C++.
View openframeworks/openFrameworksadvimman / lama
🦙 LaMa Image Inpainting, Resolution-robust Large Mask Inpainting with Fourier Convolutions, WACV 2022
View advimman/lamamicrosoft / computervision-recipes
Best Practices, code samples, and documentation for Computer Vision.
View microsoft/computervision-recipesxuebinqin / U-2-Net
The code for our newly accepted paper in Pattern Recognition 2020: "U^2-Net: Going Deeper with Nested U-Structure for Salient Object Detection."
View xuebinqin/U-2-Netroboflow / notebooks
A collection of tutorials on state-of-the-art computer vision models and techniques. Explore everything from foundational architectures like ResNet to cutting-edge models like RF-DETR, YOLO11, SAM 3
View roboflow/notebooksPeterL1n / RobustVideoMatting
Robust Video Matting in PyTorch, TensorFlow, TensorFlow.js, ONNX, CoreML!
View PeterL1n/RobustVideoMattingroboflow / rf-detr
RF-DETR is a real-time object detection and segmentation model architecture developed by Roboflow, SOTA on COCO, designed for fine-tuning. [ICLR 2026]
View roboflow/rf-detractiveloopai / deeplake
Deeplake is AI Data Runtime for Agents. It provides serverless postgres with a multimodal datalake, enabling scalable retrieval and training.
View activeloopai/deeplakedusty-nv / jetson-inference
Hello AI World guide to deploying deep-learning inference networks and deep vision primitives with TensorRT and NVIDIA Jetson.
View dusty-nv/jetson-inferenceamusi / Deep-Learning-Interview-Book
深度学习面试宝典(含数学、机器学习、深度学习、计算机视觉、自然语言处理和SLAM等方向)
View amusi/Deep-Learning-Interview-BookMITDeepLearning / introtodeeplearning
Lab Materials for MIT 6.S191: Introduction to Deep Learning
View MITDeepLearning/introtodeeplearninglucidrains / imagen-pytorch
Implementation of Imagen, Google's Text-to-Image Neural Network, in Pytorch
View lucidrains/imagen-pytorchexadel-inc / CompreFace
Leading free and open-source face recognition system
View exadel-inc/CompreFacejamez-bondos / awesome-gpt4o-images
Awesome curated collection of images and prompts generated by GPT-4o and gpt-image-1. Explore AI generated visuals created with ChatGPT and Sora
View jamez-bondos/awesome-gpt4o-imagesmicrosoft / ailab
Experience, Learn and Code the latest breakthrough innovations with Microsoft AI
View microsoft/ailabXavierXiao / Dreambooth-Stable-Diffusion
Implementation of Dreambooth (https://arxiv.org/abs/2208.12242) with Stable Diffusion
View XavierXiao/Dreambooth-Stable-DiffusionHenryNdubuaku / maths-cs-ai-compendium
Become a cracked AI/ML researcher/engineer with this unconventional textbook covering maths, computing, and ML with intuition.
View HenryNdubuaku/maths-cs-ai-compendiumamusi / awesome-object-detection
Awesome Object Detection based on handong1587 github: https://handong1587.github.io/deep_learning/2015/10/09/object-detection.html
View amusi/awesome-object-detectionhybridgroup / gocv
Go package for computer vision using OpenCV 4 and beyond. Includes support for DNN, CUDA, OpenCV Contrib, and OpenVINO.
View hybridgroup/gocvopen-mmlab / mmagic
OpenMMLab Multimodal Advanced, Generative, and Intelligent Creation Toolbox. Unlock the magic 🪄: Generative-AI (AIGC), easy-to-use APIs, awsome model zoo, diffusion models
View open-mmlab/mmagicPeterL1n / BackgroundMattingV2
Real-Time High-Resolution Background Matting
View PeterL1n/BackgroundMattingV2lance-format / lance
Open Lakehouse Format for Multimodal AI. Convert from Parquet in 2 lines of code for 100x faster random access, vector index, and data versioning.
View lance-format/lanceSerpentAI / SerpentAI
Game Agent Framework. Helping you create AIs / Bots that learn to play any game you own!
View SerpentAI/SerpentAIpliang279 / awesome-multimodal-ml
Reading list for research topics in multimodal machine learning
View pliang279/awesome-multimodal-mlPaddlePaddle / models
Officially maintained, supported by PaddlePaddle, including CV, NLP, Speech, Rec, TS, big models and so on.
View PaddlePaddle/modelsNVIDIA / pix2pixHD
Synthesizing and manipulating 2048x1024 images with conditional GANs
View NVIDIA/pix2pixHDclovaai / donut
Official Implementation of OCR-free Document Understanding Transformer (Donut) and Synthetic Document Generator (SynthDoG), ECCV 2022
View clovaai/donutamusi / daily-paper-computer-vision
记录每天整理的计算机视觉/深度学习/机器学习相关方向的论文
View amusi/daily-paper-computer-visionopenMVG / openMVG
open Multiple View Geometry library. Basis for 3D computer vision and Structure from Motion.
View openMVG/openMVGSkalskiP / courses
This repository is a curated collection of links to various courses and resources about Artificial Intelligence (AI)
View SkalskiP/coursesliuruoze / EasyPR
(CGCSTCD'2017) An easy, flexible, and accurate plate recognition project for Chinese licenses in unconstrained situations. CGCSTCD = China Graduate Contest on Smart-city Technology and Creative Design
View liuruoze/EasyPRjason718 / awesome-self-supervised-learning
A curated list of awesome self-supervised methods
View jason718/awesome-self-supervised-learningKevinMusgrave / pytorch-metric-learning
The easiest way to use deep metric learning in your application. Modular, flexible, and extensible. Written in PyTorch.
View KevinMusgrave/pytorch-metric-learningdragen1860 / TensorFlow-2.x-Tutorials
TensorFlow 2.x version's Tutorials and Examples, including CNN, RNN, GAN, Auto-Encoders, FasterRCNN, GPT, BERT examples, etc. TF 2.0版入门实例代码,实战教程。
View dragen1860/TensorFlow-2.x-Tutorialspromptslab / Awesome-Prompt-Engineering
This repository contains a hand-curated resources for Prompt Engineering with a focus on Generative Pre-trained Transformer (GPT), ChatGPT, PaLM etc
View promptslab/Awesome-Prompt-EngineeringComputer Vision & Detection Projects
- opencv/opencv — ★ 90,853
- ultralytics/ultralytics — ★ 61,651
- ultralytics/yolov5 — ★ 58,015
- roboflow/supervision — ★ 50,364
- ashishpatel26/500-AI-Machine-learning-Deep-learning-Computer-vision-NLP-Projects-with-code — ★ 36,866
- CMU-Perceptual-Computing-Lab/openpose — ★ 34,447
- lucidrains/vit-pytorch — ★ 25,506
- spmallick/learnopencv — ★ 23,152
Image Generation & Diffusion Repositories
- Anil-matcha/Open-Generative-AI — ★ 28,583
- junyanz/CycleGAN — ★ 12,871
- lucidrains/DALLE2-pytorch — ★ 11,299
- lucidrains/imagen-pytorch — ★ 8,423
- jamez-bondos/awesome-gpt4o-images — ★ 8,147
- XavierXiao/Dreambooth-Stable-Diffusion — ★ 7,732
- open-mmlab/mmagic — ★ 7,466
Photo Editing, Upscaling & Restoration
- advimman/lama — ★ 10,260
- PeterL1n/BackgroundMattingV2 — ★ 7,190
- Developer-Y/cs-video-courses — ★ 83,509