Skip to content
@NVIDIA

NVIDIA Corporation

Pinned Loading

  1. cosmos cosmos Public

    NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.

    Jupyter Notebook 11.8k 881

  2. NemoClaw NemoClaw Public

    Run agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference

    TypeScript 22.5k 3.1k

  3. TensorRT-LLM TensorRT-LLM Public

    TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. Tensor…

    Python 14.6k 2.8k

  4. cutlass cutlass Public

    CUDA Templates and Python DSLs for High-Performance Linear Algebra

    C++ 10.5k 2.1k

  5. warp warp Public

    A Python framework for GPU-accelerated simulation, robotics, and machine learning.

    Python 7.1k 621

  6. open-gpu-kernel-modules open-gpu-kernel-modules Public

    NVIDIA Linux open GPU kernel module source

    C 17.4k 1.9k

Repositories

Showing 10 of 804 repositories
  • xla_staging Public Forked from openxla/xla

    A machine learning compiler for GPUs, CPUs, and ML accelerators

    NVIDIA/xla_staging's past year of commit activity
    C++ 3 Apache-2.0 949 0 0 Updated Sep 18, 2026
  • cccl Public

    CUDA Core Compute Libraries

    NVIDIA/cccl's past year of commit activity
    C++ 2,514 487 1,641 (9 issues need help) 274 Updated Sep 18, 2026
  • NemoClaw Public

    Run agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference

    NVIDIA/NemoClaw's past year of commit activity
    TypeScript 22,489 Apache-2.0 3,092 542 (1 issue needs help) 123 Updated Sep 18, 2026
  • TensorRT-LLM Public

    TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.

    NVIDIA/TensorRT-LLM's past year of commit activity
    Python 14,643 2,756 596 913 Updated Sep 18, 2026
  • earth2studio Public

    Open-source deep-learning framework for exploring, building and deploying AI weather/climate workflows.

    NVIDIA/earth2studio's past year of commit activity
    Python 1,136 Apache-2.0 258 11 27 Updated Sep 18, 2026
  • nvcf Public

    Platform for deploying and routing GPU-accelerated inference, streaming, and batch workloads at scale.

    NVIDIA/nvcf's past year of commit activity
    Go 215 Apache-2.0 72 319 91 Updated Sep 18, 2026
  • OpenShell Public

    OpenShell is the safe, private runtime for autonomous AI agents.

    NVIDIA/OpenShell's past year of commit activity
    Rust 8,683 Apache-2.0 1,261 380 (2 issues need help) 146 Updated Sep 18, 2026
  • Model-Optimizer Public

    A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.

    NVIDIA/Model-Optimizer's past year of commit activity
    Python 3,823 Apache-2.0 602 95 300 Updated Sep 18, 2026
  • NVSentinel Public

    NVSentinel detects and remediates GPU faults on Kubernetes nodes

    NVIDIA/NVSentinel's past year of commit activity
    Go 384 Apache-2.0 131 61 (3 issues need help) 30 Updated Sep 18, 2026
  • Megatron-LM Public

    Ongoing research training transformer models at scale

    NVIDIA/Megatron-LM's past year of commit activity
    Python 17,938 4,529 421 (3 issues need help) 945 Updated Sep 18, 2026