Inferenza, distribuzione e runtime
NVIDIA/TensorRT-LLM
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way
★ 14,4K⑂ 2,7KC++
NOASSERTIONQ82
Programmazione AI e strumenti per sviluppatori
paulduvall/ai-development-patterns
A comprehensive collection of AI development patterns for building software with AI assistance, organized by implementation maturity and development lifecycle phases. Includes Foundation, Development, and Operations patterns with practical examples and anti-patterns.
★ 638⑂ 53Python
MITQ82
Programmazione AI e strumenti per sviluppatori
xai-org/grok-build
SpaceXAI's coding agent harness and TUI. Fullscreen, mouse interactive, extensible.
★ 22,6K⑂ 4,3KRust
Apache-2.0Q79
Visione artificiale
facebookresearch/sam2
The repository provides code for running inference with the Meta Segment Anything Model 2 (SAM 2), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
★ 19,6K⑂ 2,5KJupyter Notebook
Apache-2.0Q79
Agenti e multi-agente
wukongim/wukongim
More than just IM 不只是即时通讯(IM)
★ 4,9K⑂ 702Go
Licenza non rilevataQ79