Skip to content
@NVIDIA

NVIDIA Corporation

Pinned Loading

  1. cosmos cosmos Public

    NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.

    Jupyter Notebook 11.3k 798

  2. NemoClaw NemoClaw Public

    Run agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference

    TypeScript 22k 3k

  3. TensorRT-LLM TensorRT-LLM Public

    TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. Tensor…

    Python 14.3k 2.6k

  4. cutlass cutlass Public

    CUDA Templates and Python DSLs for High-Performance Linear Algebra

    C++ 10.2k 2k

  5. warp warp Public

    A Python framework for GPU-accelerated simulation, robotics, and machine learning.

    Python 6.9k 575

  6. open-gpu-kernel-modules open-gpu-kernel-modules Public

    NVIDIA Linux open GPU kernel module source

    C 17.2k 1.8k

Repositories

Showing 10 of 778 repositories
  • TensorRT-LLM Public

    TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.

    NVIDIA/TensorRT-LLM's past year of commit activity
    Python 14,270 2,626 620 962 Updated Jul 31, 2026
  • Model-Optimizer Public

    A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.

    NVIDIA/Model-Optimizer's past year of commit activity
    Python 3,357 Apache-2.0 517 84 235 Updated Jul 31, 2026
  • NemoClaw Public

    Run agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference

    NVIDIA/NemoClaw's past year of commit activity
    TypeScript 22,001 Apache-2.0 2,989 223 (2 issues need help) 83 Updated Jul 31, 2026
  • NeMo-Relay Public

    Multi-language agent runtime and library for execution scope management, lifecycle events, and middleware on tool and LLM calls.

    NVIDIA/NeMo-Relay's past year of commit activity
    Rust 87 Apache-2.0 50 7 13 Updated Jul 31, 2026
  • OSMO Public

    The developer-first platform for scaling complex Physical AI workloads across heterogeneous compute—unifying training GPUs, simulation clusters, and edge devices in a simple YAML

    Python 199 Apache-2.0 45 75 36 Updated Jul 31, 2026
  • infra-controller Public

    NVIDIA Infra Controller - Hardware Lifecycle Management and multitenant networking

    NVIDIA/infra-controller's past year of commit activity
    Rust 242 Apache-2.0 164 677 (1 issue needs help) 77 Updated Jul 31, 2026
  • NeMo-Fabric Public

    NVIDIA NeMo Fabric

    NVIDIA/NeMo-Fabric's past year of commit activity
    Python 17 Apache-2.0 8 0 7 Updated Jul 31, 2026
  • NeMo-Retriever Public

    NeMo Retriever Library is a scalable, performance-oriented document content and metadata extraction microservice. NeMo Retriever Library uses specialized NVIDIA NIM microservices to find, contextualize, and extract text, tables, charts and images that you can use in downstream generative applications.

    NVIDIA/NeMo-Retriever's past year of commit activity
    Python 2,956 Apache-2.0 341 128 (1 issue needs help) 102 Updated Jul 31, 2026
  • nvcf Public

    Platform for deploying and routing GPU-accelerated inference, streaming, and batch workloads at scale.

    Go 188 Apache-2.0 40 159 39 Updated Jul 31, 2026
  • skills Public

    Agent Skills for NVIDIA products — install into Claude Code, Codex, and other coding agents to run Physical AI, robotics, simulation, CUDA, and RAG workflows end to end.

    Python 2,747 Apache-2.0 319 6 7 Updated Jul 31, 2026

Top languages

Loading…

Most used topics

Loading…