About spmallick/learnopencv
spmallick/learnopencv is an open-source project on GitHub, mainly written in Jupyter Notebook. Learn OpenCV : C++ and Python Examples It currently holds 23,152 stars and 0 forks with 0 open issues, and was last pushed on an unknown date (repository created unknown).
Project Overview
AI Homed tracks it on the AI Image Projects board and on the AI AI Image Projects list.
GitHub Repository Details
README
LearnOpenCV
This repository contains code for Computer Vision, Deep learning, and AI research articles shared on our blog LearnOpenCV.com.
---
Build Production-Ready Computer Vision & AI Solutions
LearnOpenCV is maintained by BigVision.AI, a computer vision and AI consulting company. We help organizations design, build, optimize, and deploy production-ready AI solutions. Our team has deep expertise in computer vision, deep learning, multimodal AI, and edge deployment, with experience solving complex technical challenges across industries.
Have a project in mind? Talk with our expert AI solution builders.
List of Blog Posts
| Blog Post | Code| | ------------- |:-------------| | Train RF-DETR for Instance Segmentation with GPT-6 Astra | | | How 10x AI Engineers Train Models for Production: Lessons from PINTO's Face Alignment Project | | | SAM-3: What’s New, How It Works, and Why It Matters [Updated] | Code | | VGGT vs. VGGT-Ω (VGGT-Omega): A Complete Guide to Feed-Forward 3D Reconstruction | Code | | MiniCPM-o 4.5: A 9B Model That Can See, Hear, and Speak at the Same Time | Code | | Object Tracking using OpenCV (C++/Python) [Updated] | Code | | Object Detection with OpenCV 5 in C++: YOLO26 Pose and Segmentation [Updated] | Code | | Read, Write and Display a Video using OpenCV [Updated] | Code | | Histogram of Oriented Gradients Explained Using OpenCV [Updated] | Code | | Edge Detection Using OpenCV [Updated] | Code | | Read, Display and Write an Image Using OpenCV [Updated] | Code | | Cropping an Image Using OpenCV [Updated] | Code | | Image Resizing with OpenCV [Updated] | Code | | Filling Holes in an Image Using OpenCV (Python/C++) [Updated] | Code | | Barcode and QR Code Scanner Using OpenCV [Updated] | Code | | OCR Text Recognition Using Tesseract and OpenCV [Updated] | Code | | Color Spaces in OpenCV (C++ and Python) [Updated] | Code | | Human Pose Estimation with OpenCV (C++ and Python) [Updated] | Code | | Face Detection with OpenCV and Dlib (C++ and Python) [Updated] | Code | | Rotation Matrix to Euler Angles [Updated] | Code | | Hough Transform with OpenCV (C++ and Python) [Updated] | Code | | Otsu's Thresholding with OpenCV [Updated] | Code | | Stereo Camera Depth Estimation with OpenCV (Python and C++) [Updated] | Code | | Understanding Lens Distortion [Updated] | Code | | Blob Detection Using OpenCV (Python, C++) [Updated] | Code | | Homography Examples Using OpenCV (Python and C++) [Updated] | Code | | Convex Hull Using OpenCV in Python and C++ [Updated] | Code | | Hu Moments for Shape Matching with OpenCV (Python and C++) [Updated] | Code | | BRISQUE Image Quality Assessment with OpenCV [Updated] | Code | | Super Resolution in OpenCV [Updated] | Code | | Augmented Reality Using ArUco Markers in OpenCV [Updated] | Code | | Head Pose Estimation with OpenCV [Updated] | Code | | Average Face with OpenCV: C++ and Python Tutorial [Updated] | Code | | Image Alignment with ECC in OpenCV: C++ and Python [Updated] | Code | | Monocular SLAM in Python with OpenCV [Updated] | Code | | OpenCV QR Code Scanner in C++ and Python [Updated] | Code | | Optical Flow in OpenCV: Sparse and Dense Methods [Updated] | Code | | CNN-Based Image Colorization with OpenCV DNN [Updated] | Code | | Snake Game with OpenCV and Python [Updated] | Code | |Delaunay Triangulation and Voronoi Diagram using OpenCV ( C++ / Python) [Updated] | Code | | Contour Detection using OpenCV (Python/C++) [Updated] | Code | | Feature Based Image Alignment using OpenCV (C++/Python) [Updated] | Code | | Video Stabilization Using Point Feature Matching in OpenCV [Updated] | Code | |Camera Calibration using OpenCV [Updated] |Code| |Find the Center of a Blob (Centroid) using OpenCV (C++/Python) [Updated] | Code| | How to find frame rate or frames per second (fps) in OpenCV ( Python / C++ ) ? [Updated] | Code | | Install OpenCV 5 on Linux | Code | | Decoding Virat Kohli's Flick Shot: AI-Based 3D Motion Reconstruction | Code | | How to Run Object Detection with OpenCV 5 | | | World Cup 2026 Offside Technology: AI, Computer Vision, and the Connected Ball | | | AI Aced JEE Advanced 2026. Can You Trust It to Teach You? | | | How to Fine-Tune YOLO26 for Safety Gear and Sign Language Detection | Code | | How to Unlock 5 Vision Skills with the Moondream Cloud API | Code | | JEE Advanced 2026: We Tested AI on the Toughest Exam | | | How to Master Qwen3-VL Embedding and Reranker for Multimodal Search | Code | | How to Master YOLOE: Real-Time Open-Vocabulary Detection Made Easy | Code | | Vision Banana: How Image Generators Are Becoming Powerful Vision Models | | | YOLO26 Keypoint Estimation: Real-Time Pose Estimation with Ultralytics | Code | | RF-DETR Segmentation: Real-Time Detection & Instance Segmentation Guide | Code | | YOLO26 Instance Segmentation: Pixel-Perfect AI at Real-Time Speed | Code | | Multi-Object Tracking with Roboflow Trackers and OpenCV | Code | | Real-Time Face Blur and Pixelation with OpenCV YuNet | Code | | Breaking the Bottleneck: Achieving Native NMS-Free Inference with YOLO26 | Code | | YOLOv26: An Object Detector Built for Real-Time Deployment | Code | | Beyond Transformers: A Deep Dive into HOPE | | | Serving SGLang: Launch a Production-Style Server | | |Deployment on Edge: LLM Serving on Jetson using vLLM|Code| |Nested Learning: Is Deep Learning Architecture an Illusion?|| | How to Build a GitHub Code-Analyser Agent for Developer Productivity | Code | | The Existential Problems in LLM Serving | | | SAM 3D: Foundation Model for Single-Image 3D Reconstruction | | | Image-GS: Adaptive Image Reconstruction using 2D Gaussians | Code | | Ultimate Guide to Vector Databases and RAG Pipeline | Code | |What Makes DeepSeek OCR So Powerful|Code| | 2D Gaussian Splatting: Geometrically Accurate Radiance Field Reconstruction | Code | | TRM: Tiny Recursive Models | Code | |Deploying ML Models on Arduino: From Blink to Think|Code| | VideoRAG: Redefining Long-Context Video Comprehension | | | AI Agent in Action: Automating Desktop Tasks with VLMs | Code | | Top VLM Evaluation Metrics for Optimal Performance Analysis | Code | |Getting Started with VLM on Jetson Nano|Code| | VLM on Edge: Worth the Hype or Just a Novelty? | Code | | AnomalyCLIP : Harnessing CLIP for Weakly-Supervised Video Anomaly Recognition | Code | | AI_for_Video_Understanding_From_Content_Moderation_to_Summarization | Code | | Video-RAG: Training-Free Retrieval for Long-Video LVLMs | Code | | Object Detection and Spatial Understanding with VLMs ft. Qwen2.5-VL | Code | | LangGraph: Building Self-Correcting RAG Agent for Code Generation | Code | | Inside Sinusoidal Position Embeddings: A Sense of Order | Code | | Inside RoPE: Rotary Magic into Position Embeddings | Code | | SimLingo-Vision-Language-Action-Model-for-Autonomous-Driving | Code | | FineTuning Gemma 3n for Medical VQA on ROCOv2 | Code | | SmolLM3 Blueprint: SOTA 3B-Parameter LLM | | | LangGraph-A-Visual-Automation-and-Summarization-Pipeline | Code | | Fine-Tuning AnomalyCLIP: Class-Agnostic Zero-Shot Anomaly Detection | Code | | SigLIP 2: DeepMind’s Multilingual Vision-Language Model | | | MedGemma: Google’s Medico VLM for Clinical QA, Imaging, and More | Code | | Nanonets-OCR-s: Enabling Rich, Structured Markdown for Document Understanding | | | Optimizing VJEPA-2: Tackling Latency & Context in Real-Time Video Classification Scripts | Code | | V-JEPA 2: Meta’s Breakthrough in AI for the Physical World | Code | | NVIDIA Cosmos Reason1: Video Understanding | Code | | GR00T N1.5 Explained | | | LLaVA | Code | | SmolVLA: Affordable & Efficient VLA Robotics on Consumer GPUs | Code | | Fine-Tuning Grounding DINO: Open-Vocabulary Object Detection | Code | | Getting Started with Qwen3 – The Thinking Expert | Code | | Inside the GPU: A Comprehensive Guide to Modern Graphics Architecture | | | Distributed Parallel Training: PyTorch | Code | | MONAI: The Definitive Framework for Medical Imaging Powered by PyTorch | | | SANA-Sprint: The One-Step Revolution in High-Quality AI Image Synthesis | | | FramePack-Video-Diffusion-but-feels-like-Image-Diffusion | Code | | Model Weights File Formats in Machine Learning | | | Unsloth: A Guide from Basics to Fine-Tuning Vision Models | [Code](https: