Computervisie
652 projecten in AI & ML
Toont top 200 van 652 op Atlas Score; verfijn met de filters hierboven.
mirobody
@thetahealthThe AI-native health data engine — collect, translate, and reason with AI Agents over labs results, wearables & genomics.
SnapOtter
@snapotter-hqSuite of 200+ web tools for converting and editing images, videos, audio and PDFs, including a layer-based image editor, OCR, transcription, background removal and batch processing pipelines (alternative to SmallPDF, iLovePDF, CloudConvert).
Frigate
@blakeblackshearMonitor your security cameras with locally processed AI.
LunaTranslator
@HIllya51视觉小说翻译器 / Visual Novel Translator
Viseron
@roflcoopterSelf-hosted, local-only NVR and AI Computer Vision software. With features such as object detection, motion detection, face recognition and more, it gives you the power to keep an eye on your home, office or any other place you want to monitor.
OpenCV
@opencvOpen Source Computer Vision Library
X-AnyLabeling
@CVHub520X-AnyLabeling: A lightweight, efficient, and unified cross-platform desktop application for annotating text, image, video, and multimodal data, combining versatile built-in tools with state-of-the-art AI models and flexible multi-format export.
Easydict
@tisfeng一个简洁优雅的词典翻译 macOS App。开箱即用,支持离线 OCR 识别,支持有道词典,🍎 苹果系统词典,🍎 苹果系统翻译,OpenAI,Gemini,DeepL,Google,Bing,腾讯,百度,阿里,小牛,彩云和火山翻译。A concise and elegant Dictionary and Translator macOS App for looking up words and translating text.
BallonsTranslator
@dmMaze深度学习辅助漫画翻译工具, 支持一键机翻和简单的图像/文本编辑 | Yet another computer-aided comic/manga translation tool powered by deeplearning
OCRmyPDF
@ocrmypdfOCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
opendataloader-pdf
@opendataloader-projectPDF Parser for AI-ready data. Automate PDF accessibility. Open-source.
MaaAssistantArknights
@MaaAssistantArknights《明日方舟》小助手,全日常一键长草!| A one-click tool for the daily tasks of Arknights, supporting all clients.
imgproxy
@imgproxyFast and secure standalone server for resizing and converting remote images.
eSearch
@xushengfeng截屏 离线OCR 搜索翻译 以图搜图 贴图 录屏 万向滚动截屏 屏幕翻译 Screenshot Offline OCR Search Translate Search for picture Paste the picture on the screen Screen recorder Omnidirectional scrolling screenshot Screen translator 支持Windows Linux macOS
camera.ui
@camerauiThe modern, local-first platform for professional video surveillance.
ShareX
@ShareXShareX is a free and open-source application that enables users to capture or record any area of their screen with a single keystroke. It also supports uploading images, text, and various file types to a wide range of destinations.
rerun
@rerun-ioVisualize, query, and stream to train on multimodal robotics data.
RapidRAW
@CyberTimonA beautiful, non-destructive, and GPU-accelerated RAW image editor built with performance in mind.
compressO
@codeforreal1Convert any video/image into a tiny size. 100% free & open-source. Available for Mac, Windows & Linux.
LLPlayer
@umlx5hThe media player for language learning, with dual subtitles, AI-generated subtitles, real-time translation, and more!
clearcam
@roryclearAdd object detection, tracking, mobile notifications, and search to any security camera.
Skill_Seekers
@yusufkaraaslanConvert documentation websites, GitHub repositories, and PDFs into Claude AI skills with automatic conflict detection
Graphite
@GraphiteEditorCommunity-built comprehensive 2D content creation appplication for graphic design, digital art, and interactive real-time motion graphics powered by a node-based procedural graphics engine
Invoice-Downloader
@EthanYoQ电子发票整理与报销准备工具:从邮箱批量收集 PDF/OFD/XML 发票,OCR 识别、分类归档并生成 Excel 汇总;提供 Windows/macOS 桌面版与 DSH 插件。
CameraChessWeb
@PbatchRecord a chess game live and upload the PGN to Lichess
eclaire
@eclaire-labsLocal-first, open-source AI assistant for your data. Unify tasks, notes, docs, photos, and bookmarks. Private, self-hosted, and extensible via APIs.
chineseocr_lite
@DayBreak-u超轻量级中文ocr,支持竖排文字识别, 支持ncnn、mnn、tnn推理 ( dbnet(1.8M) + crnn(2.5M) + anglenet(378KB)) 总模型仅4.7M
sherloq
@GuidoBartoliAn open-source digital image forensic toolset
manga-translator-ui
@hgmzhn基于manga-image-translator 实现的开源漫画AI翻译桌面工具。支持日、韩、英文漫画自动处理,集成OpenAl、Gemini等多翻译引擎;实现OCR文字检测、原文擦除、AI翻译、图像修复、译文排版完整链路,自带可视化编辑器,支持自定义文本样式,一键部署开箱即用。
sharp
@lovellHigh performance Node.js image processing, the fastest module to resize JPEG, PNG, WebP, AVIF and TIFF images. Uses the libvips library.
bisheng
@dataelementBISHENG is an open LLM devops platform for next generation Enterprise AI applications. Powerful and comprehensive features include: GenAI workflow, RAG, Agent, Unified model management, Evaluation, SFT, Dataset Management, Enterprise-level System Management, Observability and more.
PyMuPDF
@pymupdfPyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
librealsense
@realsenseaiRealSense SDK
sparrow
@katanamlStructured data extraction, instruction calling and agentic workflows with ML, LLM and Vision LLM
MaaFramework
@MaaXYZ基于图像识别的自动化黑盒测试框架 | An automation black-box testing framework based on image recognition
Damselfly
@webreaperFast server-based photo management system for large collections of images. Includes face detection, face & object recognition, powerful search, and EXIF Keyword tagging. Runs on Linux, MacOS and Windows.
ImageToolbox
@T8RIN🖼️ Image Toolbox is a powerful app for advanced image manipulation. It offers dozens of features, from basic tools like crop and draw to filters, OCR, and a wide range of image processing options
obsidian-omnisearch
@scambierA search engine that "just works" for Obsidian. Supports OCR and PDF indexing.
tesseract
@tesseract-ocrTesseract Open Source OCR Engine (main repository)
RapidOCR
@RapidAI📄 Awesome OCR multiple programing languages toolkits based on ONNX Runtime, OpenVINO, MNN, PaddlePaddle, TensorRT and PyTorch.
obs-backgroundremoval
@royshilAn OBS plugin for removing background in portrait images (video), making it easy to replace the background when recording or streaming.
KillerPDF
@SteveTheKillerFree and open-source PDF editor for Windows with a built-in PDF 2.0 engine. View, annotate, OCR, merge, split, crop, rotate, compare, edit text, draw, sign, fill forms, print, flatten, and open password-protected PDFs without a subscription.
LichtFeld-Studio
@MrNeRFTrain, inspect, edit, automate, and export 3D Gaussian Splatting scenes from a single native application.
MaaNTE
@1bananachickenMaaNTE. Nevertheless to Everless automatic assistant 异环小助手
crosspaste-desktop
@CrossPasteCross-device clipboard sync for macOS, Windows & Linux — end-to-end encrypted, LAN-only, no cloud. OCR, CLI and MCP server built in.
openexr
@AcademySoftwareFoundationThe OpenEXR project provides the specification and reference implementation of the EXR file format, the professional-grade image storage format of the motion picture industry.
Flyimg
@flyimgResize and crop images on the fly. Get optimised images with MozJPEG, WebP or PNG using ImageMagick, with an efficient caching system.
scenic
@google-researchScenic: A Jax Library for Computer Vision Research and Beyond
paperlib
@Future-ScholarsAn open-source academic paper management tool.
Final2x
@EutropicAIa cross-platform image super-resolution tool
Translumo
@ramjkeAdvanced real-time screen translator for games, hardcoded subtitles in videos, static text and etc.
PaddleOCR
@PaddlePaddleTurn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
labelme
@wkentaroImage annotation with Python. Supports polygon, rectangle, circle, line, point, and AI-assisted annotation.
liteparse
@run-llamaA fast, helpful, and open-source document parser
libvips
@libvipsA fast image processing library with low memory needs.
fiftyone
@voxel51Refine high-quality datasets and visual AI models
rf-detr
@roboflowRF-DETR is a real-time object detection and segmentation model architecture developed by Roboflow, SOTA on COCO, designed for fine-tuning. [ICLR 2026]
STranslate
@STranslateA ready-to-go translation ocr tool developed with WPF/WPF 开发的一款即用即走的翻译、OCR工具
anylabeling
@vietanhdevEffortless AI-assisted data labeling with AI support from YOLO, Segment Anything (SAM+SAM2/2.1+SAM3), MobileSAM!!
FAST
@FAST-ImagingA framework for high-performance medical image processing, neural network inference and visualization
cucim
@rapidsaicuCIM - RAPIDS GPU-accelerated image processing library
markdownify-mcp
@zcaceresA Model Context Protocol server for converting almost anything to Markdown
paperless-gpt
@icereedUse LLMs and LLM Vision (OCR) to handle paperless-ngx - Document Digitalization powered by AI
chandra
@datalab-toOCR model that handles complex tables, forms, handwriting with full layout.
koharu
@koharu-rsAI-powered manga translator, written in Rust.
images
@weservSource code of wsrv.nl (formerly images.weserv.nl), to be used on your own server(s).
OSS-DocumentScanner
@ossappscollectiveDocument scanning app
Savant
@insight-platformPython Computer Vision & Video Analytics Framework With Batteries Included
HomeGenie
@genielabsHomeGenie: The Programmable Intelligence with 100% Local Agentic AI.
digital-agriculture-datasets
@ricberOpen datasets for training and benchmarking AI and robotic systems in agriculture and off-road environments
mrpt
@MRPT:zap: The Mobile Robot Programming Toolkit (MRPT)
terratorch
@torchgeoA Python toolkit for fine-tuning Geospatial Foundation Models (GFMs).
D-FINE
@PeterandeD-FINE: Redefine Regression Task of DETRs as Fine-grained Distribution Refinement [ICLR 2025 Spotlight]
vision
@pytorchDatasets, Transforms and Models specific to Computer Vision
remove-ai-watermarks
@wiltodeltaRemove visible and invisible AI watermarks and provenance metadata from images and video. Python library and CLI for SynthID, C2PA, EXIF, IPTC, XMP, and common generative-AI marks.
Text-Grab
@TheJoeFinUse OCR in Windows quickly and easily with Text Grab. With optional background process and notifications.
RuVector
@ruvnetRuVector provides High Performance, Real-Time decisions and agent memory , Self-Learning Ai, Vector GNN DB built in Rust.
semantic-router
@aurelio-labsSuperfast AI decision making and intelligent processing of multi-modal data.
spritefusion-pixel-snapper
@Hugo-DzA tool to snap pixels to a perfect grid. Designed to fix messy and inconsistent pixel art generated by AI.
comic-translate
@ogkalu2AI comic and manga translator app/browser extension for automatically translating comics, manga, manhwa, BDs, fumetti, and more in multiple languages and formats (Images, PDF, EPUB, CBR, CBZ etc).
Docspell
@eikekAuto-tagging document organizer and archive.
webp_server_go
@webp-shGo version of WebP Server. A tool that will serve your JPG/PNG/BMP/SVGs as WebP/AVIF format with compression, on-the-fly.
TrafficLab-3D
@duy-phamduc68Create a digital-twin style traffic visualization using only mp4 CCTV footage and its Google Maps location.
deep-learning-for-image-processing
@WZMIAOMIAOdeep learning for image processing including classification and object-detection etc.
rembg
@danielgatisRembg is a tool to remove images background
GLM-OCR
@zai-orgGLM-OCR: Accurate × Fast × Comprehensive
stitching
@OpenStitchingA Python package for fast and robust Image Stitching
3dgrut
@nv-tlabsRay tracing and hybrid rasterization of Gaussian particles
mokuro
@kha-whiteRead Japanese manga inside browser with selectable text.
detectron2
@facebookresearchDetectron2 is a platform for object detection, segmentation and other visual recognition tasks.
manga-image-translator
@zyddnysTranslate manga/image 一键翻译各类图片内文字 https://cotrans.touhou.ai/ (no longer working)
chafa
@hpjansson📺🗿 Terminal graphics for the 21st century.
TensorRT-YOLO
@laugh12321🚀 Easier & Faster YOLO Deployment Toolkit for NVIDIA 🛠️
rse-grand-challenge
@DIAGNijmegenA platform for end-to-end development of machine learning solutions in biomedical imaging
ai-hardware-engineer-roadmap
@ai-hpcMaster AI inference, AI agent harness systems, and hardware engineering — then design a physical AI chip. That is the goal.
jt-doc-tools
@jasoncheng7115An all-in-one PDF and Office document platform: self-hosted, open source, under your control. 整合式 PDF / Office 文件處理平台,自架、開源、可控。
video-subtitle-extractor
@YaoFANGUK视频硬字幕提取,生成srt文件。无需申请第三方API,本地实现文本识别。基于深度学习的视频字幕提取框架,包含字幕区域检测、字幕内容提取。A GUI tool for extracting hard-coded subtitle (hardsub) from videos and generating srt files.
postprocessing
@pmndrsA post processing library for three.js.
cropperjs
@fengyuanchenJavaScript image cropper.
PaddleX
@PaddlePaddleAll-in-One Development Tool based on PaddlePaddle
PySceneDetect
@Breakthrough:movie_camera: Python and OpenCV-based scene cut/transition detection program & library.
photon
@silvia-odwyer⚡ Rust/WebAssembly image processing library
Pix2Text
@breezedeusAn Open-Source Python3 tool with SMALL models for recognizing layouts, tables, math formulas (LaTeX), and text in images, converting them into Markdown format. A free alternative to Mathpix, empowering seamless conversion of visual content into text-based representations. 80+ languages are supported.
OpenImageIO
@AcademySoftwareFoundationReading, writing, and processing images in a wide variety of file formats, using a format-agnostic API, aimed at VFX applications.
edit-mind
@IliasHadLocal-first Video Knowledge Base. Index your video library with multi-modal analysis (YOLO, DeepFace, Whisper), search semantically via natural language, Docker-ready.
lightly-train
@lightly-aiAll-in-one training for vision models (YOLO, ViTs, RT-DETR, DINOv3): pretraining, fine-tuning, distillation.
BiaPy
@BiaPyXOpen source Python library for building bioimage analysis pipelines
Papermerge
@papermergeDocument management system focused on scanned documents (electronic archives). Features file browsing in similar way to dropbox/google drive. OCR, full text search, text overlay/selection.
vit-pytorch
@lucidrainsImplementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in Pytorch
alpamayo
@NVlabsNVIDIA Alpamayo 1 Nano is an open 10B reasoning VLA model for autonomous vehicles that pairs driving trajectories with Chain-of-Causation reasoning.
ImageMagick
@ImageMagickImageMagick is a free, open-source software suite for creating, editing, converting, and displaying images. It supports 200+ formats and offers powerful command-line tools and APIs for automation, scripting, and integration across platforms.
opencvsharp
@shimatOpenCV wrapper for .NET
datascience
@sreeharierkThis repository is a compilation of free resources for learning Data Science.
gowall
@AchnoA tool to convert a Wallpaper's color scheme / palette, OCR with VLM's Traditional & Hybrid, Image Compression ,color palette extraction, image upsacling with Adversarial Networks and more image processing features.
TRex
@amebalabsCopy any text on your screen, stop retyping.
pixeltable
@pixeltableThe backend agents build with - Multimodal database, orchestration, and serving in one file
yolo_ros
@mgonzs13Ultralytics YOLOv8, YOLOv9, YOLOv10, YOLOv11, YOLOv12 for ROS 2
habitat-lab
@facebookresearchA modular high-level library to train embodied AI agents across a variety of tasks and environments.
PaddleDetection
@PaddlePaddleObject Detection toolkit based on PaddlePaddle. It supports object detection, instance segmentation, multiple object tracking and real-time multi-person keypoint detection.
PhotoEditor
@burhanrashid52A Photo Editor library with simple, easy support for image editing using paints,text,filters,emoji and Sticker like stories.
bild
@anthonynsimonImage processing algorithms in pure Go
openrecall
@openrecallOpenRecall is a fully open-source, privacy-first alternative to proprietary solutions like Microsoft's Windows Recall. With OpenRecall, you can easily access your digital history, enhancing your memory and productivity without compromising your privacy.
retain-pdf
@wxyhgk在保留版面、公式与结构的前提下进行 PDF 翻译,适用于科研与技术文档
spirula-studio
@harry7557558Cross-vendor 3D Gaussian Splatting trainer - video to splat to mesh, Vulkan or CUDA.
yolov5
@ultralyticsUltralytics YOLOv5 in PyTorch for object detection, instance segmentation, classification, training, and export.
tesseract.js
@napthaPure Javascript OCR for more than 100 Languages 📖🎉🖥
Pillow
@python-pillowPython Imaging Library (fork)
ade-cli
@landing-aiThe official CLI for Agentic Document Extraction (ADE) by LandingAI — parse documents and extract schema-shaped data from your terminal
image-actions
@calibreappA Github Action that automatically compresses JPEGs, PNGs, WebPs & AVIFs in Pull Requests.
DepthFlow
@BrokenSource🌊 Images to 3D parallax videos
RMT
@zclucasRMT (RuoMengTu) is a free, open-source macro tool built on AHKv2. Let the code handle the tedious work—you have more meaningful things to do.
finsight-pro
@Ali-MarandiAI-Powered Financial Analysis Desktop App — 7 Engines, 17+ Ratios, Bankruptcy Prediction, TSETMC Live, 100% Offline
GeoDeep
@uav4geoFree and open source library for AI object detection and semantic segmentation in geospatial rasters. 🚀
carla
@carla-simulatorOpen-source simulator for autonomous driving research.
OTB
@orfeotoolboxGithub mirror of https://gitlab.orfeo-toolbox.org/orfeotoolbox/otb
basis_universal
@BinomialLLCBasis Universal GPU Texture Codec
ocrs
@robertknightRust library and CLI tool for OCR (extracting text from images)
was-node-suite-comfyui
@WASasquatchWAS-NS Reborn; Tools for image processing, filters, masking, text, logic, numbers, latents, files, 3D scenes, and animation.
ddddocr
@sml2h3带带弟弟 通用验证码识别OCR pypi版
segmentation_models.pytorch
@qubvel-orgSemantic segmentation models with 500+ pretrained convolutional and transformer-based backbones.
yolov3
@ultralyticsPyTorch implementation of YOLOv3, YOLOv3-SPP, and YOLOv3-tiny for real-time object detection with training, validation, inference, and multi-format export.
imagor
@cshumFast, secure image processing server and Go library, using libvips
docs
@sismicsLightweight document management system packed with all the features you can expect from big expensive solutions
pywt
@PyWaveletsPyWavelets - Wavelet Transforms in Python
uniface
@yakhyoUniFace: A Unified Face Analysis Library for Python | Detection, alignment, landmarks, face-mesh, recognition, parsing, gaze, attributes and anti-spoofing under one API.
symforce
@symforce-orgFast symbolic computation, code generation, and nonlinear optimization for robotics
amica
@semperaiAmica is an open source interface for interactive communication with 3D characters with voice synthesis and speech recognition.
SnapX
@SnapXLSnapX is a free, open-source, cross-platform tool that lets you capture or record any area of your screen and instantly share it with a single keypress. Upload images, videos, text, and more to multiple supported destinations—all with ease. ShareX fork
Object-Detection-Metrics
@rafaelpadillaMost popular metrics used to evaluate object detection algorithms.
TOCropViewController
@TimOliverA view controller for iOS that allows users to crop portions of UIImage objects
manga-ocr
@kha-whiteOptical character recognition for Japanese text, with the main focus being Japanese manga
PhotonCamera
@eszdmanAndroid Camera that uses Enhanced image processing
omniparse
@adithya-s-kIngest, parse, and optimize any data format ➡️ from documents to multimedia ➡️ for enhanced compatibility with GenAI frameworks
habitat-sim
@facebookresearchA flexible, high-performance 3D simulator for Embodied AI research.
Simd
@ermig1979C++ image processing and machine learning library with using of SIMD: SSE, AVX, AVX-512, AMX for x86/x64, NEON, SVE for ARM, HVX for Hexagon
MORT
@kmonkeyheadMORT 번역기 프로젝트 - Real-time game translator with OCR
streamlit-webrtc
@whitphxReal-time video and audio processing on Streamlit
Converseen
@Faster3ckConverseen is a batch image converter and resizer
MeterEye
@epynicA camera reads my electricity meter's LCD so I don't have to. ESP32-CAM + custom 7-segment decoder, no ML, no cloud
alpasim
@NVlabsAlpaSim is an open-source autonomous vehicle simulation platform designed for development and testing of end-to-end AV policies
filepond
@pqina🌊 A flexible and fun JavaScript file upload library
scikit-image
@scikit-imageImage processing in Python
jabref
@JabRefDesktop app for managing BibTeX and BibLaTeX (.bib) libraries
unrealcv
@unrealcvUnrealCV: Connecting Computer Vision to Unreal Engine
iOS-OCR-Server
@riddlelingAn iOS OCR Server Using Apple’s Vision Framework
tapnet
@google-deepmindTracking Any Point (TAP)
rpaframework
@robocorpCollection of open-source libraries and tools for Robotic Process Automation (RPA), designed to be used with both Robot Framework and Python
qupath
@qupathQuPath - Open-source bioimage analysis for research
imagej2
@imagejOpen scientific N-dimensional image processing :microscope: :sparkler:
react-native-fast-tflite
@margelo🧬 High-performance TensorFlow Lite library for React Native with GPU acceleration
cosmo-edge
@cosmo-wander-aiProduction-grade C++ edge AI engine for video analytics and on-device VLM across Sophon, Rockchip RKNN, and x86, with visual orchestration, real-time OSD, events, and reproducible benchmarks.
SVGcode
@tomayacConvert color bitmap images to color SVG vector images.
DICOMautomaton
@hdclarkA multipurpose tool for medical physics.
MuscleMap
@MuscleMapMuscleMap: An Open-Source, Community-Supported Consortium for Whole-Body Quantitative MRI of Muscle
comfyui_LLM_party
@heshengtaoLLM Agent Framework in ComfyUI includes MCP sever, Omost,GPT-sovits, ChatTTS,GOT-OCR2.0, and FLUX prompt nodes,access to Feishu,discord,and adapts to all llms with similar openai / aisuite interfaces, such as o1,ollama, gemini, grok, qwen, GLM, deepseek, kimi,doubao. Adapted to local llms, vlm, gguf such as llama-3.3 Janus-Pro, Linkage graphRAG
OnnxOCR
@jingsongliujing基于PaddleOCR重构,并且脱离PaddlePaddle深度学习训练框架的轻量级OCR,推理速度超快 —— A lightweight OCR system based on PaddleOCR, decoupled from the PaddlePaddle deep learning training framework, with ultra-fast inference speed.
MinerU
@opendatalabTransforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
WebPlotDigitizer
@automeris-ioComputer vision assisted tool to extract numerical data from plot images.
awesome-industrial-anomaly-detection
@M-3LABPaper list and datasets for industrial image anomaly/defect detection (updating). 工业异常/瑕疵检测论文及数据集检索库(持续更新)。
Slicer
@SlicerMulti-platform, free open source software for visualization and image computing.
EmbeddedSystem
@SummerGift:books: 计算机体系架构、嵌入式系统基础与主流编程语言相关内容总结
Deep-Learning-for-Solar-Panel-Recognition
@saizkCNN models for Solar Panel Detection and Segmentation in Aerial Images.
ConstructionActionRecognition
@S1mpleyangThis github project aims to recongnize unsafe actions of workers on the construction site.
unblink
@zapdos-labsCamera monitoring with VLM
google-images-download
@hardikvasaPython Script to download hundreds of images from 'Google Images'. It is a ready-to-run code!
tesserocr
@sirfzA Python wrapper for the tesseract-ocr API
govips
@davidbyttowA lightning fast image processing and resizing library for Go
frigate-hass-integration
@blakeblackshearFrigate integration for Home Assistant
exifcleaner
@szTheoryCross-platform desktop GUI app to clean image metadata
photonix
@photonixappA modern, web-based photo management server. Run it on your home server and it will let you find the right photo from your collection on any device. Smart filtering is made possible by object recognition, face recognition, location awareness, color analysis and other ML algorithms.
pymatting
@pymattingA Python library for alpha matting
OpenCVTutorials
@fendouaiOpenCV-Python4.1 中文文档
xtreme1
@xtreme1-ioXtreme1 is an all-in-one data labeling and annotation platform for multimodal data training and supports 3D LiDAR point cloud, image, and LLM.
TurboOCR
@aiptimizerTurboOCR, >200 img/s OmnidocBench. TensorRT FP16, PP-OCRv6, HTTP + gRPC
agribound
@montimajAn AI-powered field boundary delineation toolkit combining satellite foundation models, embeddings, and global training data for accurate agricultural parcel/field boundary mapping.
replicate
@ultralyticsReady-to-use Cog deployments and CI/CD for running Ultralytics YOLO11, YOLO World, YOLOE, and YOLO26 models on Replicate.
openMVG
@openMVGopen Multiple View Geometry library. Basis for 3D computer vision and Structure from Motion.
flyingmouse-format
@LaoFeng-mouse飞鼠格式 FlyingMouse Format - Windows 免费文件格式转换工具(离线可用,内置 FFmpeg/LibreOffice/Poppler/Tesseract)。图片/文档/表格/PPT/PDF/音视频/WPS 格式互转 + OCR + 批量转换;音频仅支持普通格式。作者:牢蜂(LaoFeng)|仅供个人免费使用,禁止商业售卖/转卖/套壳
deepseek-ocr.rs
@TimmyOVORust multi‑backend OCR/VLM engine (DeepSeek‑OCR-1/2, PaddleOCR‑VL, DotsOCR) with DSQ quantization and an OpenAI‑compatible server & CLI – run locally without Python.
android-gpuimage-plus
@wysaidAndroid Image & Camera Filters Based on OpenGL.