Computervisie
342 projecten in AI & ML
Toont top 200 van 342 op Atlas Score; verfijn met de filters hierboven.
Paperless-ngx
@paperless-ngxScan, index, and archive all of your paper documents with an improved interface (fork of Paperless).
SnapOtter
@snapotter-hqSuite of 200+ web tools for converting and editing images, videos, audio and PDFs, including a layer-based image editor, OCR, transcription, background removal and batch processing pipelines (alternative to SmallPDF, iLovePDF, CloudConvert).
LocalAI
@mudlerRun your AI models locally and generate images and audio (alternative to OpenAI and Claude).
mirobody
@thetahealthThe AI-native health data engine — collect, translate, and reason with AI Agents over labs results, wearables & genomics.
Frigate
@blakeblackshearMonitor your security cameras with locally processed AI.
openvino
@openvinotoolkitOpenVINO™ is an open source toolkit for optimizing and deploying AI inference
unstructured
@Unstructured-IOConvert documents to structured data effortlessly. Unstructured is open-source ETL solution for transforming complex documents into clean, structured formats for language models. Visit our website to learn more about our enterprise grade Platform product for production grade workflows, partitioning, enrichments, chunking and embedding.
aiworkdeck
@zeweihanAI-native IDE workspace for legal and document-heavy workflows: files, agents, plugins, built-in document editing with tracked changes, OCR, evidence chains. VS Code for lawyers.
opencv
@opencvOpen Source Computer Vision Library
ultralytics
@ultralyticsUltralytics YOLO27, YOLO26, YOLO11, YOLOv8 — object detection, instance segmentation, semantic segmentation, image classification, pose estimation, object tracking
pdf-craft
@oomol-labPDF craft can convert PDF files into various other formats. This project will focus on processing PDF files of scanned books.
OCRmyPDF
@ocrmypdfOCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
unstract
@ZipstackLLM-Driven Extraction of Unstructured Data — Built for API Deployments & ETL Pipeline Workflows
opendataloader-pdf
@opendataloader-projectPDF Parser for AI-ready data. Automate PDF accessibility. Open-source.
MaaAssistantArknights
@MaaAssistantArknights《明日方舟》小助手,全日常一键长草!| A one-click tool for the daily tasks of Arknights, supporting all clients.
LunaTranslator
@HIllya51视觉小说翻译器 / Visual Novel Translator
Viseron
@roflcoopterSelf-hosted, local-only NVR and AI Computer Vision software. With features such as object detection, motion detection, face recognition and more, it gives you the power to keep an eye on your home, office or any other place you want to monitor.
ShareX
@ShareXShareX is a free and open-source application that enables users to capture or record any area of their screen with a single keystroke. It also supports uploading images, text, and various file types to a wide range of destinations.
mediapipe
@google-ai-edgeCross-platform, customizable ML solutions for live and streaming media.
Easydict
@tisfeng一个简洁优雅的词典翻译 macOS App。开箱即用,支持离线 OCR 识别,支持有道词典,🍎 苹果系统词典,🍎 苹果系统翻译,OpenAI,Gemini,DeepL,Google,Bing,腾讯,百度,阿里,小牛,彩云和火山翻译。A concise and elegant Dictionary and Translator macOS App for looking up words and translating text.
rerun
@rerun-ioVisualize, query, and stream to train on multimodal robotics data.
imgproxy
@imgproxyFast and secure standalone server for resizing and converting remote images.
X-AnyLabeling
@CVHub520X-AnyLabeling: A lightweight, efficient, and unified cross-platform desktop application for annotating text, image, video, and multimodal data, combining versatile built-in tools with state-of-the-art AI models and flexible multi-format export.
OmniTools
@iib0011Collection of powerful web-based tools for everyday tasks (coding, manipulating images/videos, PDFs or crunching numbers...).
Flyimg
@flyimgResize and crop images on the fly. Get optimised images with MozJPEG, WebP or PNG using ImageMagick, with an efficient caching system.
torchgeo
@torchgeoTorchGeo: datasets, samplers, transforms, and pre-trained models for geospatial data
supervision
@roboflowWe write your reusable computer vision tools. 💜
AI-For-Beginners
@microsoft12 Weeks, 24 Lessons, AI for All!
Graphite
@GraphiteEditorCommunity-built comprehensive 2D content creation appplication for graphic design, digital art, and interactive real-time motion graphics powered by a node-based procedural graphics engine
datasets
@huggingface🤗 The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools
Damselfly
@webreaperFast server-based photo management system for large collections of images. Includes face detection, face & object recognition, powerful search, and EXIF Keyword tagging. Runs on Linux, MacOS and Windows.
camera.ui
@camerauiThe modern, local-first platform for professional video surveillance.
Skill_Seekers
@yusufkaraaslanConvert documentation websites, GitHub repositories, and PDFs into Claude AI skills with automatic conflict detection
tesseract
@tesseract-ocrTesseract Open Source OCR Engine (main repository)
ai-engineering-from-scratch
@rohitg00Learn it. Build it. Ship it for others.
sharp
@lovellHigh performance Node.js image processing, the fastest module to resize JPEG, PNG, WebP, AVIF and TIFF images. Uses the libvips library.
cvat
@cvat-aiComputer Vision Annotation Tool (CVAT) is a leading platform for building high-quality visual datasets for vision AI. It offers open-source, cloud, and enterprise products, as well as labeling services, for image, video, and 3D annotation with AI-assisted labeling, quality assurance, team collaboration, analytics, and developer APIs.
kornia
@kornia🐍 Geometric Computer Vision Library for Spatial AI
eSearch
@xushengfeng截屏 离线OCR 搜索翻译 以图搜图 贴图 录屏 万向滚动截屏 屏幕翻译 Screenshot Offline OCR Search Translate Search for picture Paste the picture on the screen Screen recorder Omnidirectional scrolling screenshot Screen translator 支持Windows Linux macOS
CameraChessWeb
@PbatchRecord a chess game live and upload the PGN to Lichess
AgML
@Project-AgMLAgML is a centralized framework for agricultural machine learning. AgML provides access to public agricultural datasets for common agricultural deep learning tasks, with standard benchmarks and pretrained models, as well the ability to generate synthetic data and annotations.
eclaire
@eclaire-labsLocal-first, open-source AI assistant for your data. Unify tasks, notes, docs, photos, and bookmarks. Private, self-hosted, and extensible via APIs.
geemap
@gee-communityA Python package for interactive geospatial analysis and visualization with Google Earth Engine.
PaddleOCR
@PaddlePaddleTurn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
chineseocr_lite
@DayBreak-u超轻量级中文ocr,支持竖排文字识别, 支持ncnn、mnn、tnn推理 ( dbnet(1.8M) + crnn(2.5M) + anglenet(378KB)) 总模型仅4.7M
bisheng
@dataelementBISHENG is an open LLM devops platform for next generation Enterprise AI applications. Powerful and comprehensive features include: GenAI workflow, RAG, Agent, Unified model management, Evaluation, SFT, Dataset Management, Enterprise-level System Management, Observability and more.
RapidRAW
@CyberTimonA beautiful, non-destructive, and GPU-accelerated RAW image editor built with performance in mind.
Invoice-Downloader
@EthanYoQ电子发票整理与报销准备工具:从邮箱批量收集 PDF/OFD/XML 发票,OCR 识别、分类归档并生成 Excel 汇总;提供 Windows/macOS 桌面版与 DSH 插件。
HRConvert2
@zelon88A self-hosted, respource aware file conversion server supporting 488 formats in 26 languages.
scenic
@google-researchScenic: A Jax Library for Computer Vision Research and Beyond
fiftyone
@voxel51Refine high-quality datasets and visual AI models
PyMuPDF
@pymupdfPyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
autogluon
@autogluonFast and Accurate ML in 3 Lines of Code
Savant
@insight-platformPython Computer Vision & Video Analytics Framework With Batteries Included
maths-cs-ai-compendium
@HenryNdubuakuBecome a cracked AI/ML researcher/engineer with this unconventional textbook covering maths, computing, and ML with intuition.
paperlib
@Future-ScholarsAn open-source academic paper management tool.
labelme
@wkentaroImage annotation with Python. Supports polygon, rectangle, circle, line, point, and AI-assisted annotation.
liteparse
@run-llamaA fast, helpful, and open-source document parser
libvips
@libvipsA fast image processing library with low memory needs.
librealsense
@realsenseaiRealSense SDK
crosspaste-desktop
@CrossPasteCross-device clipboard sync for macOS, Windows & Linux — end-to-end encrypted, LAN-only, no cloud. OCR, CLI and MCP server built in.
TrafficLab-3D
@duy-phamduc68Create a digital-twin style traffic visualization using only mp4 CCTV footage and its Google Maps location.
digital-agriculture-datasets
@ricberOpen datasets for training and benchmarking AI and robotic systems in agriculture and off-road environments
vision
@pytorchDatasets, Transforms and Models specific to Computer Vision
rf-detr
@roboflowRF-DETR is a real-time object detection and segmentation model architecture developed by Roboflow, SOTA on COCO, designed for fine-tuning. [ICLR 2026]
RapidOCR
@RapidAI📄 Awesome OCR multiple programing languages toolkits based on ONNX Runtime, OpenVINO, MNN, PaddlePaddle, TensorRT and PyTorch.
MouseTooltipTranslator
@ttop32Mouseover Translate Any Language At Once - Chrome Extension: PDF Translator, EBOOK, EPUB, OCR, TTS, NETFLIX, YOUTUBE DUAL SUBTITLES, GOOGLE DOCS, AI, VIEWER, GMAIL, WRITING, IMAGE, DUAL SUBS, MANGA, HOVER, DICTIONARY, WEBTOON, EDGE, JAPANESE, ENGLISH
AutonomousVehicleControlBeginnersGuide
@ShisatoYanoPython sample codes and documents about Autonomous vehicle control algorithm. This project can be used as a technical guide book to study the algorithms and the software architectures for beginners.
pythoncode-tutorials
@x4nth055The Python Code Tutorials
deep-learning-for-image-processing
@WZMIAOMIAOdeep learning for image processing including classification and object-detection etc.
rembg
@danielgatisRembg is a tool to remove images background
chandra
@datalab-toOCR model that handles complex tables, forms, handwriting with full layout.
ludwig
@ludwig-aiLow-code framework for building custom LLMs, neural networks, and other AI models
Final2x
@EutropicAIa cross-platform image super-resolution tool
Docspell
@eikekAuto-tagging document organizer and archive.
AlbumentationsX
@albumentations-teamImage augmentation for computer vision. AGPL-3.0-only or commercial licensing.
HomeGenie
@genielabsHomeGenie: The Programmable Intelligence with 100% Local Agentic AI.
mrpt
@MRPT:zap: The Mobile Robot Programming Toolkit (MRPT)
STranslate
@STranslateA ready-to-go translation ocr tool developed with WPF/WPF 开发的一款即用即走的翻译、OCR工具
sports
@roboflowcomputer vision and sports
yolov5
@ultralyticsUltralytics YOLOv5 in PyTorch for object detection, instance segmentation, classification, training, and export.
tesseract.js
@napthaPure Javascript OCR for more than 100 Languages 📖🎉🖥
vit-pytorch
@lucidrainsImplementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in Pytorch
ImageMagick
@ImageMagickImageMagick is a free, open-source software suite for creating, editing, converting, and displaying images. It supports 200+ formats and offers powerful command-line tools and APIs for automation, scripting, and integration across platforms.
dlib
@daviskingA toolkit for making real world machine learning and data analysis applications in C++
cropperjs
@fengyuanchenJavaScript image cropper.
pytorch-grad-cam
@jacobgilAdvanced AI Explainability for computer vision. Support for CNNs, Vision Transformers, Classification, Object detection, Segmentation, Image similarity and more.
manga-image-translator
@zyddnysTranslate manga/image 一键翻译各类图片内文字 https://cotrans.touhou.ai/ (no longer working)
Yolov5-deepsort-inference
@SharpilessYolov5 deepsort inference,使用YOLOv5+Deepsort实现车辆行人追踪和计数,代码封装成一个Detector类,更容易嵌入到自己的项目中
ai-hardware-engineer-roadmap
@ai-hpcMaster AI inference, AI agent harness systems, and hardware engineering — then design a physical AI chip. That is the goal.
jt-doc-tools
@jasoncheng7115An all-in-one PDF and Office document platform: self-hosted, open source, under your control. 整合式 PDF / Office 文件處理平台,自架、開源、可控。
terratorch
@torchgeoA Python toolkit for fine-tuning Geospatial Foundation Models (GFMs).
habitat-lab
@facebookresearchA modular high-level library to train embodied AI agents across a variety of tasks and environments.
GLM-OCR
@zai-orgGLM-OCR: Accurate × Fast × Comprehensive
doctr
@mindeedocTR (Document Text Recognition) - a seamless, high-performing & accessible library for OCR-related tasks powered by Deep Learning. Ongoing development and maintenance by t2k.
sumatrapdf-plus
@dengxiboSumatraPDF fork: Chinese EPUB/MOBI, smart PDF dark mode, OCR, TTS, offline dictionary.
FAST
@FAST-ImagingA framework for high-performance medical image processing, neural network inference and visualization
cucim
@rapidsaicuCIM - RAPIDS GPU-accelerated image processing library
finsight-pro
@Ali-MarandiAI-Powered Financial Analysis Desktop App — 7 Engines, 17+ Ratios, Bankruptcy Prediction, TSETMC Live, 100% Offline
carla
@carla-simulatorOpen-source simulator for autonomous driving research.
ddddocr
@sml2h3带带弟弟 通用验证码识别OCR pypi版
PaddleDetection
@PaddlePaddleObject Detection toolkit based on PaddlePaddle. It supports object detection, instance segmentation, multiple object tracking and real-time multi-person keypoint detection.
Pillow
@python-pillowPython Imaging Library (fork)
video-subtitle-extractor
@YaoFANGUK视频硬字幕提取,生成srt文件。无需申请第三方API,本地实现文本识别。基于深度学习的视频字幕提取框架,包含字幕区域检测、字幕内容提取。A GUI tool for extracting hard-coded subtitle (hardsub) from videos and generating srt files.
OMRChecker
@Udayraj123Evaluate OMR sheets fast and accurately using a scanner 🖨 or your phone 🤳.
korean-jangbu-for
@kimlawtech한국 스타트업, 1인 법인, 프리랜서, 개인 사업자를 위한 장부 자동 생성 Claude Code 스킬. 카드명세서 PDF·은행 CSV → 재무제표·세무사 전달 CSV 자동 생성. Level 2 민감정보 마스킹 적용.
MeterEye
@epynicA camera reads my electricity meter's LCD so I don't have to. ESP32-CAM + custom 7-segment decoder, no ML, no cloud
MATLAB-Simulink-Challenge-Project-Hub
@mathworksThis MATLAB and Simulink Challenge Project Hub contains a list of research and design project ideas. These projects will help you gain practical experience and insight into technology trends and industry directions.
datascience
@sreeharierkThis repository is a compilation of free resources for learning Data Science.
OC_SORT
@noahcao[CVPR2023] The official repo for OC-SORT: Observation-Centric SORT on video Multi-Object Tracking. OC-SORT is simple, online and robust to occlusion/non-linear motion.
Papermerge
@papermergeDocument management system focused on scanned documents (electronic archives). Features file browsing in similar way to dropbox/google drive. OCR, full text search, text overlay/selection.
filepond
@pqina🌊 A flexible and fun JavaScript file upload library
segmentation_models.pytorch
@qubvel-orgSemantic segmentation models with 500+ pretrained convolutional and transformer-based backbones.
yolov3
@ultralyticsPyTorch implementation of YOLOv3, YOLOv3-SPP, and YOLOv3-tiny for real-time object detection with training, validation, inference, and multi-format export.
Biodiversity
@microsoftMicrosoft AI for Good Lab — Biodiversity research hub. Open-source AI models, edge devices, and tools for biodiversity monitoring and conservation. Your source for MegaDetector, SPARROW, PytorchWildlife, Bioacoustics, and more.
ConstructionActionRecognition
@S1mpleyangThis github project aims to recongnize unsafe actions of workers on the construction site.
Object-Detection-Metrics
@rafaelpadillaMost popular metrics used to evaluate object detection algorithms.
MinerU
@opendatalabTransforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
500-AI-Machine-learning-Deep-learning-Computer-vision-NLP-Projects-with-code
@ashishpatel26500 AI Machine learning Deep learning Computer vision NLP Projects with code
CV
@AccumulateMore✅(已完结)超级全面的 深度学习 笔记【土堆 Pytorch】【李沐 动手学深度学习】【吴恩达 深度学习】【大飞 大模型Agent】
screenpipe
@screenpipeYC (S26) | Open Computer History | Record your screen continuously locally and provide context to your agents (Claude, Codex, Openclaw, Hermes, Runner...)
SimpleITK
@SimpleITKSimpleITK: a layer built on top of the Insight Toolkit (ITK), intended to simplify and facilitate ITK's use in rapid prototyping, education and interpreted languages.
nQuantCpp
@mcychannQuantCpp includes top 6 color quantization algorithms for visual c++ producing high quality optimized images.
awesome-industrial-anomaly-detection
@M-3LABPaper list and datasets for industrial image anomaly/defect detection (updating). 工业异常/瑕疵检测论文及数据集检索库(持续更新)。
Slicer
@SlicerMulti-platform, free open source software for visualization and image computing.
Deep-Learning-for-Solar-Panel-Recognition
@saizkCNN models for Solar Panel Detection and Segmentation in Aerial Images.
agribound
@montimajAn AI-powered field boundary delineation toolkit combining satellite foundation models, embeddings, and global training data for accurate agricultural parcel/field boundary mapping.
grass
@OSGeoGRASS - free and open-source geospatial processing engine
geospatial
@opengeosA Python package for installing commonly used packages for geospatial analysis and data visualization with only one command.
cs-video-courses
@Developer-YList of Computer Science courses with video lectures.
scikit-image
@scikit-imageImage processing in Python
Chinese-CLIP
@OFA-SysChinese version of CLIP which achieves Chinese cross-modal retrieval and representation generation.
BundleTrack
@wenbowen123[IROS 2021] BundleTrack: 6D Pose Tracking for Novel Objects without Instance or Category-Level 3D Models
lex
@i-dot-aiUK legal API for AI agents and researchers.
open_clip
@mlfoundationsAn open source implementation of CLIP.
colmap
@colmapCOLMAP - Structure-from-Motion and Multi-View Stereo
techniques
@satellite-image-deep-learningTechniques for deep learning with satellite & aerial imagery
omniparse
@adithya-s-kIngest, parse, and optimize any data format ➡️ from documents to multimedia ➡️ for enhanced compatibility with GenAI frameworks
OTB
@orfeotoolboxGithub mirror of https://gitlab.orfeo-toolbox.org/orfeotoolbox/otb
google-images-download
@hardikvasaPython Script to download hundreds of images from 'Google Images'. It is a ready-to-run code!
openMVG
@openMVGopen Multiple View Geometry library. Basis for 3D computer vision and Structure from Motion.
lucida
@egeorcunBackground removal that keeps what matters: glass, camouflage, text, glow and line art. BiRefNet fine-tune, MIT.
smile
@haifenglStatistical Machine Intelligence & Learning Engine
catalyst
@catalyst-teamAccelerated deep learning R&D
ImageSharp
@SixLaborsA modern, cross-platform, 2D Graphics library for .NET
Learning-Deep-Learning
@patrick-llgcPaper reading notes on Deep Learning and Machine Learning
opendatacam
@opendatacamAn open source tool to quantify the world
edgeyolo
@LSH9832an edge-real-time anchor-free object detector with decent performance
Umi-OCR
@hiroi-soraOCR software, free and offline. 开源、免费的离线OCR软件。支持截屏/批量导入图片,PDF文档识别,排除水印/页眉页脚,扫描/生成二维码。内置多国语言库。
pysot
@STVIRSenseTime Research platform for single object tracking, implementing algorithms like SiamRPN and SiamMask.
efficientsam3
@SimonZeng7108EfficientSAM3 compresses SAM3 into lightweight, edge-friendly models via progressive knowledge distillation for fast promptable concept segmentation and tracking.
wild_visual_navigation
@leggedroboticsWild Visual Navigation: A system for fast traversability learning via pre-trained models and online self-supervision
raster-vision
@azaveaAn open source library and framework for deep learning on satellite and aerial imagery.
introtodeeplearning
@MITDeepLearningLab Materials for MIT 6.S191: Introduction to Deep Learning
Parsr
@axa-groupTransforms PDF, Documents and Images into Enriched Structured Data
ballerine
@ballerine-ioOpen-source infrastructure and data orchestration platform for risk decisioning
Awesome-pytorch-list
@bharathgsA comprehensive list of pytorch related content on github,such as different models,implementations,helper libraries,tutorials etc.
blind_watermark
@guofei9987Blind&Invisible Watermark ,图片盲水印,提取水印无须原图!
Meshroom
@alicevisionNode-based Visual Programming Toolbox
openFrameworks
@openframeworksopenFrameworks is a community-developed cross platform toolkit for creative coding in C++.
javacv
@bytedecoJava interface to OpenCV, FFmpeg, and more
deep-learning-drizzle
@kmario23Drench yourself in Deep Learning, Reinforcement Learning, Machine Learning, Computer Vision, and NLP by learning from these exciting lectures!!
CVPR2026-Papers-with-Code
@amusiCVPR 2026 论文和开源项目合集
HyperGAN
@HyperGANComposable GAN framework with api and user interface
3DObjectTracking
@DLR-RMAlgorithms and Publications on 3D Object Tracking
openeyes
@mandarwagh9OpenEyes is an open-source robot vision framework for edge devices
Virgilio
@virgili0Your new Mentor for Data Science E-Learning.
EasyOCR
@JaidedAIReady-to-use OCR with 80+ supported languages and all popular writing scripts including Latin, Chinese, Arabic, Devanagari, Cyrillic and etc.
learnopencv
@spmallickLearn OpenCV : C++ and Python Examples
pcl
@PointCloudLibraryPoint Cloud Library (PCL)
Deep-Learning-for-Tracking-and-Detection
@abhineet123Collection of papers, datasets, code and other resources for object tracking and detection using deep learning
mv-extractor
@LukasBommesExtract frames and motion vectors from H.264 and MPEG-4 encoded video.
elpv-dataset
@zae-bayernA dataset of functional and defective solar cells extracted from EL images of solar modules
AirSim
@microsoftOpen source simulator for autonomous vehicles built on Unreal Engine / Unity, from Microsoft AI & Research
gocv
@hybridgroupGo package for computer vision using OpenCV 4 and beyond. Includes support for DNN, CUDA, OpenCV Contrib, and OpenVINO.
awesome-image-registration
@Awesome-Image-Registration-Organizationimage registration related books, papers, videos, and toolboxes
open-semantic-search
@opensemanticsearchOpen Source research tool to search, browse, analyze and explore large document collections by Semantic Search Engine and Open Source Text Mining & Text Analytics platform (Integrates ETL for document processing, OCR for images & PDF, named entity recognition for persons, organizations & locations, metadata management by thesaurus & ontologies, search user interface & search apps for fulltext search, faceted search & knowledge graph)
topconf-paper-figure-gallery
@qwdwqfwq3,516 curated Figure 1 / teasers from ICLR, ICML, NeurIPS, CVPR, ACL, AAAI (2023-2026), with tier badges and FigureForge retrieval-augmented figure drafting
LaTeX-OCR
@lukas-blecherpix2tex: Using a ViT to convert images of equations into LaTeX code.
Dolphin
@bytedanceThe official repo for “Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting”, ACL, 2025.
Compiler-Construction
@Hazrat-Ali9🐳 Compiler Construction CS 🚃 student or a curious 🚒 developer this repo ✈ will guide you through 🚁 the process of building 🚂 your own compiler 🚅 step-by-step projects 🚢 in C/C++, Python 🪀 and Java along with 🎳 mini-languages Includes 🥎 tools like Lex/Yacc 🏉 ANTLR LLVM 🎲 and Bison 🍋
zerox
@getomni-aiOCR & Document Extraction using vision models
nerfstudio
@nerfstudio-projectA collaboration friendly studio for NeRFs
YOLOX
@Megvii-BaseDetectionYOLOX is a high-performance anchor-free YOLO, exceeding yolov3~v5 with MegEngine, ONNX, TensorRT, ncnn, and OpenVINO supported. Documentation: https://yolox.readthedocs.io/
Halide
@halidea language for fast, portable data-parallel computation
VIAME
@VIAMEVideo and Image Analytics for Multiple Environments
psmoveapi
@thpCross-platform library for 6DoF tracking of the PS Move Motion Controller. Sensor fusion, computer vision, ambient display (LED orb).
docwire
@docwireDocWire SDK: local-first C++20 data-processing infrastructure for deterministic, auditable, on-premise workflows. Supports 100+ formats, OCR, local AI, and opt-in cloud AI pipelines
gaussian-splatting
@graphdeco-inriaOriginal reference implementation of "3D Gaussian Splatting for Real-Time Radiance Field Rendering"
notebooks
@roboflowA collection of tutorials on state-of-the-art computer vision models and techniques. Explore everything from foundational architectures like ResNet to cutting-edge models like RF-DETR, YOLO11, SAM 3, and Qwen3-VL.
norfair
@tryolabsLightweight Python library for adding real-time multi-object tracking to any detector.
vehicle_counting_tensorflow
@ahmetozlu:oncoming_automobile: "MORE THAN VEHICLE COUNTING!" This project provides prediction for speed, color and size of the vehicles with TensorFlow Object Counting API.
deepstream-services-library
@prominenceaiA shared library of on-demand DeepStream Pipeline Services for Python and C/C++
Thermogram
@s-duA powerful app for processing DJI infrared (IR) images, featuring advanced thermal analysis, conversion, and visualization capabilities. Supports DJI Matrice 4 (M4T), Matrice 30 (M30T), Mavic 3 (M3T), and Mavic 2 Enterprise Advanced drones
fixxer
@oaklensartAI-powered photography workflow automation
lama
@advimman🦙 LaMa Image Inpainting, Resolution-robust Large Mask Inpainting with Fourier Convolutions, WACV 2022
pytorch-metric-learning
@KevinMusgraveThe easiest way to use deep metric learning in your application. Modular, flexible, and extensible. Written in PyTorch.
UCF-SST-CitySim1-Dataset
@UCF-SST-LabOfficial github page of UCF SST CitySim Dataset
noaa-apt
@martinberNOAA APT weather satellite image decoder, for Linux, Windows, RPi 2+, OSX and Android+Termux
note4yaoo
@uptonkingdaily notes