Run agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference
NVIDIA · Dòng mô hình
Nemotron
Nemotron model family from NVIDIA.
Dòng thời gian phát hành
Dự án
Ongoing research training transformer models at scale
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.
Security scanner for AI agent skills. Detect vulnerabilities, malicious patterns, and security risks.
NVIDIA® TensorRT™ is an SDK for high-performance deep learning inference on NVIDIA GPUs. This repository contains the open source components of TensorRT.
the LLM vulnerability scanner
NVIDIA’s robot models and reference code offer multimodal action prediction, demonstration-data adaptation, fine-tuning and deployment tooling.
Compiles typed Python functions into compute kernels with geometry, physics and differentiable programming capabilities for simulation and machine learning.
Configures NVIDIA GPU access for container runtimes such as Docker, supporting containerized inference, training and CUDA workloads.
A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.
Open-source deep-learning framework for building, training, and fine-tuning deep learning models using state-of-the-art Physics-ML methods
Agent Skills for NVIDIA products — install into Claude Code, Codex, and other coding agents to run Physical AI, robotics, simulation, CUDA, and RAG workflows end to end.
AIStore: scalable storage for AI applications