Inferência, implantação e execução
NVIDIA/TensorRT-LLM
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way
★ 14,4K⑂ 2,7KC++
NOASSERTIONQ82
Programação com AI e ferramentas para desenvolvedores
paulduvall/ai-development-patterns
A comprehensive collection of AI development patterns for building software with AI assistance, organized by implementation maturity and development lifecycle phases. Includes Foundation, Development, and Operations patterns with practical examples and anti-patterns.
★ 638⑂ 53Python
MITQ82
Programação com AI e ferramentas para desenvolvedores
xai-org/grok-build
SpaceXAI's coding agent harness and TUI. Fullscreen, mouse interactive, extensible.
★ 22,6K⑂ 4,3KRust
Apache-2.0Q79
Visão computacional
facebookresearch/sam2
The repository provides code for running inference with the Meta Segment Anything Model 2 (SAM 2), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
★ 19,6K⑂ 2,5KJupyter Notebook
Apache-2.0Q79
Agentes e multiagente
wukongim/wukongim
More than just IM 不只是即时通讯(IM)
★ 4,9K⑂ 702Go
Licença não detectadaQ79