A GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learning training and inference applications.
Repositories
NVIDIA repositories
Deep Learning GPU Training System
State-of-the-Art Deep Learning scripts organized by models - easy to train and deploy with reproducible accuracy and performance on enterprise-grade infrastructure.
Style transfer, deep learning, feature transform
A Fusion Code Generator for NVIDIA GPUs (commonly known as "nvFuser")
NVIDIA Isaac GR00T N1.7 - A Foundation Model for Generalist Robots.
Ongoing research training transformer models at scale
The NVIDIA NeMo Agent toolkit is an open-source library for efficiently connecting and optimizing teams of AI agents.
NeMo Retriever Library is a scalable, performance-oriented document content and metadata extraction microservice. NeMo Retriever Library uses specialized NVIDIA NIM microservices to find, contextualize, and extract text, tables, charts and images that you can use in downstream generative applications.
A toolkit for processing speech data and creating speech datasets
NeMo text processing for ASR and TTS
Toolkit for efficient experimentation with Speech Recognition, Text2Speech and NLP
OpenShell is the safe, private runtime for autonomous AI agents.
NVIDIA® TensorRT™ is an SDK for high-performance deep learning inference on NVIDIA GPUs. This repository contains the open source components of TensorRT.
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.
A PyTorch Extension: Tools for easy mixed precision and distributed training in Pytorch
A tool for testing and validating container requirements against versioned manifests
NVIDIA Fleet Intelligence Agent - Host agent for GPU telemetry collection and attestation
[ARCHIVED] The C++ Standard Library for your entire system. See https://github.com/NVIDIA/cccl
Numbast is a tool to build an automated pipeline that converts CUDA APIs into Numba bindings.