ToolAI.io · Kênh GitHub

Thị giác máy tính · Các dự án AI mã nguồn mở trên GitHub

Chỉ mục dựa trên dữ kiện về các dự án phát triển LLM, tác nhân, MCP, RAG và AI, kèm giấy phép, thiết lập, tải xuống, ảnh chụp màn hình và tài nguyên liên quan cho từng mục.

Dữ liệu kho lưu trữ công khai

Mục lục dự án

20 Dự án
Thị giác máy tính

supervision

roboflow/supervision

Reusable utilities for computer vision data handling, detection annotation and tracking workflows.

★ 49,9K⑂ 4,8KPython
MITQ85
Thị giác máy tính

Segment Anything: Generate Image Segmentation Masks with Points and Boxes

facebookresearch/segment-anything

Meta’s SAM image segmentation project supports prompt inputs such as points and boxes, and can also automatically generate candidate masks for an entire image. It is suitable for interactive annotation and image processing workflows.

★ 54,8K⑂ 6,4KPython
Apache-2.0Q84
Thị giác máy tính

DINOv2: visual features for classification and retrieval

facebookresearch/dinov2

Meta’s self-supervised vision models provide pretrained backbones and task heads for feature extraction, visual retrieval and downstream adaptation.

★ 13,3K⑂ 1,3KPython
Apache-2.0Q84
Thị giác máy tính

PaddleDetection

PaddlePaddle/PaddleDetection

Object Detection toolkit based on PaddlePaddle. It supports object detection, instance segmentation, multiple object tracking and real-time multi-person keypoint detection.

★ 14,3K⑂ 3KPython
Apache-2.0Q80
Thị giác máy tính

PaddleSeg

PaddlePaddle/PaddleSeg

Easy-to-use image segmentation library with awesome pre-trained model zoo, supporting wide-range of practical tasks in Semantic Segmentation, Interactive Segmentation, Panoptic Segmentation, Image Matting, 3D Segmentation, etc.

★ 9,4K⑂ 1,7KPython
Apache-2.0Q80
Thị giác máy tính

sam2

facebookresearch/sam2

The repository provides code for running inference with the Meta Segment Anything Model 2 (SAM 2), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.

★ 19,6K⑂ 2,5KJupyter Notebook
Apache-2.0Q79
Thị giác máy tính

VGGT: Estimating Cameras, Depth, and 3D Structure from Multi-View Images

facebookresearch/vggt

Visual Geometry Grounded Transformer predicts camera parameters, depth maps, point maps, and point tracks from scene images, making it suitable for preliminary geometric estimation in 3D vision research and reconstruction workflows.

★ 14,4K⑂ 1,5KPython
Không phát hiện giấy phépQ77

Trang 2 / 2 · 20 dự án

Đã cập nhật gần đây

TypeSafe Python SDKtypesafe-ai/typesafe-sdk-python★ 218 TypeSafe JavaScript SDKtypesafe-ai/typesafe-sdk-js★ 231 TypeSafe Agent Skillstypesafe-ai/skills★ 2K cherry-studioCherryHQ/cherry-studio★ 51,5K onyxonyx-dot-app/onyx★ 32K siyuansiyuan-note/siyuan★ 46,2K

Được gắn sao nhiều nhất

ECCaffaan-m/ECC★ 248,8K Hermes Agentnousresearch/hermes-agent★ 227,1K tensorflowtensorflow/tensorflow★ 198,8K AutoGPTSignificant-Gravitas/AutoGPT★ 187,1K ollamaollama/ollama★ 180,2K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174,9K