Skip to content
@NVIDIA

NVIDIA Corporation

Pinned Loading

  1. cosmos cosmos Public

    NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.

    Jupyter Notebook 12k 908

  2. NemoClaw NemoClaw Public

    Run agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference

    TypeScript 22.7k 3.1k

  3. TensorRT-LLM TensorRT-LLM Public

    TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. Tensor…

    Python 14.8k 2.8k

  4. cutlass cutlass Public

    CUDA Templates and Python DSLs for High-Performance Linear Algebra

    C++ 10.5k 2.1k

  5. warp warp Public

    A Python framework for GPU-accelerated simulation, robotics, and machine learning.

    Python 7.2k 650

  6. open-gpu-kernel-modules open-gpu-kernel-modules Public

    NVIDIA Linux open GPU kernel module source

    C 17.5k 1.9k

Repositories

Showing 10 of 815 repositories
  • cluster-readiness-engine Public

    NVIDIA Cluster Readiness Engine

    NVIDIA/cluster-readiness-engine's past year of commit activity
    Go 73 Apache-2.0 31 26 11 Updated Oct 9, 2026
  • cccl Public

    CUDA Core Compute Libraries

    NVIDIA/cccl's past year of commit activity
  • nvcf Public

    Platform for deploying and routing GPU-accelerated inference, streaming, and batch workloads at scale.

    NVIDIA/nvcf's past year of commit activity
    Go 231 Apache-2.0 90 365 116 Updated Oct 9, 2026
  • garak Public

    the LLM vulnerability scanner

    Python 9,512 Apache-2.0 1,354 269 (15 issues need help) 207 Updated Oct 9, 2026
  • multi-storage-client Public

    Unified high-performance Python client for object and file stores.

    NVIDIA/multi-storage-client's past year of commit activity
    Python 85 Apache-2.0 27 1 0 Updated Oct 9, 2026
  • NemoClaw Public

    Run agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference

    NVIDIA/NemoClaw's past year of commit activity
    TypeScript 22,688 Apache-2.0 3,133 624 (2 issues need help) 138 Updated Oct 9, 2026
  • TensorRT-LLM Public

    TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.

    NVIDIA/TensorRT-LLM's past year of commit activity
    Python 14,781 2,804 628 973 Updated Oct 9, 2026
  • Model-Optimizer Public

    A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.

    NVIDIA/Model-Optimizer's past year of commit activity
    Python 5,255 Apache-2.0 736 101 365 Updated Oct 9, 2026
  • TransformerEngine Public

    A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hopper, Ada and Blackwell GPUs, to provide better performance with lower memory utilization in both training and inference.

    NVIDIA/TransformerEngine's past year of commit activity
    Python 3,571 Apache-2.0 857 161 201 Updated Oct 9, 2026
  • cuda-quantum Public

    C++ and Python support for the CUDA Quantum programming model for heterogeneous quantum-classical workflows

    NVIDIA/cuda-quantum's past year of commit activity
    C++ 1,159 Apache-2.0 480 380 (12 issues need help) 174 Updated Oct 9, 2026