ToolAI.io · Canal do GitHub

Visão computacional · Projetos de AI de código aberto no GitHub

Um índice baseado em fatos de projetos de desenvolvimento de LLM, agentes, MCP, RAG e AI, com licença, configuração, download, capturas de tela e recursos relacionados para cada item.

Dados do repositório público

Índice do projeto

20 Projetos
Visão computacional

supervision

roboflow/supervision

Reusable utilities for computer vision data handling, detection annotation and tracking workflows.

★ 49,9K⑂ 4,8KPython
MITQ85
Visão computacional

Segment Anything: Generate Image Segmentation Masks with Points and Boxes

facebookresearch/segment-anything

Meta’s SAM image segmentation project supports prompt inputs such as points and boxes, and can also automatically generate candidate masks for an entire image. It is suitable for interactive annotation and image processing workflows.

★ 54,8K⑂ 6,4KPython
Apache-2.0Q84
Visão computacional

DINOv2: visual features for classification and retrieval

facebookresearch/dinov2

Meta’s self-supervised vision models provide pretrained backbones and task heads for feature extraction, visual retrieval and downstream adaptation.

★ 13,3K⑂ 1,3KPython
Apache-2.0Q84
Visão computacional

PaddleDetection

PaddlePaddle/PaddleDetection

Object Detection toolkit based on PaddlePaddle. It supports object detection, instance segmentation, multiple object tracking and real-time multi-person keypoint detection.

★ 14,3K⑂ 3KPython
Apache-2.0Q80
Visão computacional

PaddleSeg

PaddlePaddle/PaddleSeg

Easy-to-use image segmentation library with awesome pre-trained model zoo, supporting wide-range of practical tasks in Semantic Segmentation, Interactive Segmentation, Panoptic Segmentation, Image Matting, 3D Segmentation, etc.

★ 9,4K⑂ 1,7KPython
Apache-2.0Q80
Visão computacional

sam2

facebookresearch/sam2

The repository provides code for running inference with the Meta Segment Anything Model 2 (SAM 2), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.

★ 19,6K⑂ 2,5KJupyter Notebook
Apache-2.0Q79
Visão computacional

VGGT: Estimating Cameras, Depth, and 3D Structure from Multi-View Images

facebookresearch/vggt

Visual Geometry Grounded Transformer predicts camera parameters, depth maps, point maps, and point tracks from scene images, making it suitable for preliminary geometric estimation in 3D vision research and reconstruction workflows.

★ 14,4K⑂ 1,5KPython
Licença não detectadaQ77

Página 2 / 2 · 20 projetos

Atualizado recentemente

TypeSafe Python SDKtypesafe-ai/typesafe-sdk-python★ 218 TypeSafe JavaScript SDKtypesafe-ai/typesafe-sdk-js★ 231 TypeSafe Agent Skillstypesafe-ai/skills★ 2K cherry-studioCherryHQ/cherry-studio★ 51,5K onyxonyx-dot-app/onyx★ 32K siyuansiyuan-note/siyuan★ 46,2K

Mais estrelados

ECCaffaan-m/ECC★ 248,8K Hermes Agentnousresearch/hermes-agent★ 227,1K tensorflowtensorflow/tensorflow★ 198,8K AutoGPTSignificant-Gravitas/AutoGPT★ 187,1K ollamaollama/ollama★ 180,2K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174,9K