Inférence, déploiement et exécution
NVIDIA/TensorRT-LLM
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way
★ 14,4K⑂ 2,7KC++
NOASSERTIONQ82
Évaluation, observabilité et sécurité
NVIDIA/garak
the LLM vulnerability scanner
★ 8,9K⑂ 1,2KPython
Apache-2.0Q84
Évaluation, observabilité et sécurité
openai/evals
Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks
★ 19,2K⑂ 3,1KPython
NOASSERTIONQ82
LLM et modèles de fondation
openai/gpt-oss
gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI
★ 20,3K⑂ 2,1KPython
Apache-2.0Q86
Agents et multi-agents
microsoft/SkillOpt
SkillOpt is a text-space optimizer that trains reusable natural-language skills for frozen LLM agents through trajectory-driven edits, validation-gated updates, and deployable best_skill.md artifacts
★ 16,2K⑂ 1,5KPython
MITQ86
AI multimodale
PaddlePaddle/PaddleOCR
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages
★ 88K⑂ 11,2KPython
Apache-2.0Q86