Skip to content
@NVIDIA

NVIDIA Corporation

Pinned Loading

  1. cosmos cosmos Public

    NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.

    Jupyter Notebook 12k 899

  2. NemoClaw NemoClaw Public

    Run agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference

    TypeScript 22.6k 3.1k

  3. TensorRT-LLM TensorRT-LLM Public

    TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. Tensor…

    Python 14.8k 2.8k

  4. cutlass cutlass Public

    CUDA Templates and Python DSLs for High-Performance Linear Algebra

    C++ 10.5k 2.1k

  5. warp warp Public

    A Python framework for GPU-accelerated simulation, robotics, and machine learning.

    Python 7.2k 641

  6. open-gpu-kernel-modules open-gpu-kernel-modules Public

    NVIDIA Linux open GPU kernel module source

    C 17.4k 1.9k

Repositories

Showing 10 of 812 repositories
  • aicr Public

    Tooling for optimized, validated, and reproducible GPU-accelerated AI runtime in Kubernetes

    NVIDIA/aicr's past year of commit activity
    Go 432 Apache-2.0 110 160 23 Updated Oct 2, 2026
  • Model-Optimizer Public

    A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.

    NVIDIA/Model-Optimizer's past year of commit activity
    Python 5,165 Apache-2.0 720 99 372 Updated Oct 2, 2026
  • TensorRT-LLM Public

    TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.

    NVIDIA/TensorRT-LLM's past year of commit activity
    Python 14,757 2,792 609 929 Updated Oct 2, 2026
  • garak Public

    the LLM vulnerability scanner

    NVIDIA/garak's past year of commit activity
    Python 9,408 Apache-2.0 1,330 273 (15 issues need help) 211 Updated Oct 2, 2026
  • cuda-python Public

    CUDA Python: Performance meets Productivity

    NVIDIA/cuda-python's past year of commit activity
    Cython 3,390 Apache-2.0 336 213 91 Updated Oct 2, 2026
  • physicsnemo Public

    Open-source deep-learning framework for building, training, and fine-tuning deep learning models using state-of-the-art Physics-ML methods

    NVIDIA/physicsnemo's past year of commit activity
    Python 3,310 Apache-2.0 803 21 76 Updated Oct 2, 2026
  • flashdreams Public

    high-performance inference and serving library for interactive autoregressive video and world models

    NVIDIA/flashdreams's past year of commit activity
    Python 510 61 48 78 Updated Oct 2, 2026
  • nvalchemi-toolkit Public

    ALCHEMI Toolkit is a developer toolkit for accelerating training and inference for AI in chemistry and material science.

    NVIDIA/nvalchemi-toolkit's past year of commit activity
    Python 168 Apache-2.0 39 6 16 Updated Oct 2, 2026
  • cudf Public

    cuDF - GPU DataFrame Library

    NVIDIA/cudf's past year of commit activity
    C++ 9,769 Apache-2.0 1,125 1,176 (31 issues need help) 208 Updated Oct 2, 2026
  • cuda-quantum Public

    C++ and Python support for the CUDA Quantum programming model for heterogeneous quantum-classical workflows

    NVIDIA/cuda-quantum's past year of commit activity
    C++ 1,151 Apache-2.0 472 379 (13 issues need help) 177 Updated Oct 2, 2026