추론, 배포 및 런타임
NVIDIA/TensorRT-LLM
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way
★ 14.4K⑂ 2.7KC++
NOASSERTIONQ82
AI 코딩 및 개발자 도구
paulduvall/ai-development-patterns
A comprehensive collection of AI development patterns for building software with AI assistance, organized by implementation maturity and development lifecycle phases. Includes Foundation, Development, and Operations patterns with practical examples and anti-patterns.
★ 638⑂ 53Python
MITQ82
AI 코딩 및 개발자 도구
xai-org/grok-build
SpaceXAI's coding agent harness and TUI. Fullscreen, mouse interactive, extensible.
★ 22.6K⑂ 4.3KRust
Apache-2.0Q79
컴퓨터 비전
facebookresearch/sam2
The repository provides code for running inference with the Meta Segment Anything Model 2 (SAM 2), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
★ 19.6K⑂ 2.5KJupyter Notebook
Apache-2.0Q79
에이전트 및 멀티 에이전트
wukongim/wukongim
More than just IM 不只是即时通讯(IM)
★ 4.9K⑂ 702Go
라이선스가 감지되지 않음Q79