ToolAI.io · GitHub Channel

Computer Vision · Open-source AI Projects on GitHub

A fact-based index of LLM, agent, MCP, RAG and AI development projects, with license, setup, download, screenshots and related resources for each entry.

Public repository data

Project index

20 projects
Computer Vision

supervision

roboflow/supervision

Reusable utilities for computer vision data handling, detection annotation and tracking workflows.

★ 49.9K⑂ 4.8KPython
MITQ85
Computer Vision

Segment Anything: Generate Image Segmentation Masks with Points and Boxes

facebookresearch/segment-anything

Meta’s SAM image segmentation project supports prompt inputs such as points and boxes, and can also automatically generate candidate masks for an entire image. It is suitable for interactive annotation and image processing workflows.

★ 54.8K⑂ 6.4KPython
Apache-2.0Q84
Computer Vision

DINOv2: visual features for classification and retrieval

facebookresearch/dinov2

Meta’s self-supervised vision models provide pretrained backbones and task heads for feature extraction, visual retrieval and downstream adaptation.

★ 13.3K⑂ 1.3KPython
Apache-2.0Q84
Computer Vision

PaddleDetection

PaddlePaddle/PaddleDetection

Object Detection toolkit based on PaddlePaddle. It supports object detection, instance segmentation, multiple object tracking and real-time multi-person keypoint detection.

★ 14.3K⑂ 3KPython
Apache-2.0Q80
Computer Vision

PaddleSeg

PaddlePaddle/PaddleSeg

Easy-to-use image segmentation library with awesome pre-trained model zoo, supporting wide-range of practical tasks in Semantic Segmentation, Interactive Segmentation, Panoptic Segmentation, Image Matting, 3D Segmentation, etc.

★ 9.4K⑂ 1.7KPython
Apache-2.0Q80
Computer Vision

sam2

facebookresearch/sam2

The repository provides code for running inference with the Meta Segment Anything Model 2 (SAM 2), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.

★ 19.6K⑂ 2.5KJupyter Notebook
Apache-2.0Q79
Computer Vision

VGGT: Estimating Cameras, Depth, and 3D Structure from Multi-View Images

facebookresearch/vggt

Visual Geometry Grounded Transformer predicts camera parameters, depth maps, point maps, and point tracks from scene images, making it suitable for preliminary geometric estimation in 3D vision research and reconstruction workflows.

★ 14.4K⑂ 1.5KPython
License not detectedQ77

Page 2 / 2 · 20 projects

Recently updated

TypeSafe Python SDKtypesafe-ai/typesafe-sdk-python★ 218 TypeSafe JavaScript SDKtypesafe-ai/typesafe-sdk-js★ 231 TypeSafe Agent Skillstypesafe-ai/skills★ 2K cherry-studioCherryHQ/cherry-studio★ 51.5K onyxonyx-dot-app/onyx★ 32K siyuansiyuan-note/siyuan★ 46.2K

Most starred

ECCaffaan-m/ECC★ 248.8K Hermes Agentnousresearch/hermes-agent★ 227.1K tensorflowtensorflow/tensorflow★ 198.8K AutoGPTSignificant-Gravitas/AutoGPT★ 187.1K ollamaollama/ollama★ 180.2K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174.9K