NVIDIA catalog
Models, skills and blueprints for GPU jobs.
Browse NVIDIA workloads inside ICPX before creating a compute job.
nvidia
Build and Deploy a Multi-Agent Chatbot
Deploy a multi-agent chatbot system and chat with agents on your Spark
nvidia
Build Knowledge Graphs with txt2kg
Extract triples with Ollama or vLLM, store them in a graph database, and explore them in a GPU-accelerated web UI
nvidia
CLI Coding Agent
Build local CLI coding agents with Ollama
nvidia
Connect Multiple DGX Spark through a Switch
Set up a cluster of DGX Spark devices that are connected through Switch
nvidia
Connect Three DGX Spark in a Ring Topology
Connect and set up three DGX Spark devices in a ring topology
nvidia
Connect Two DGX Stations for Distributed Workloads
Combined memory and compute over a direct high-speed link
nvidia
Connect Two Sparks
Connect two Spark devices and setup them up for inference and fine-tuning
nvidia
CUDA-X Data Science
Install and use NVIDIA cuML and NVIDIA cuDF to accelerate UMAP, HDBSCAN, pandas and more with zero code changes
nvidia
cuTile Kernels
Run cuTile kernel benchmarks, FMHA implementation, and LLM inference on DGX Spark and B300
nvidia
DGX Dashboard
Monitor your DGX system and launch JupyterLab
nvidia
DGX Station AI Skills and dgx-assist
Inspect DGX Station software and route version-aware, CLI-backed workflows
nvidia
Fine-tune with NeMo
Use NVIDIA NeMo to fine-tune models locally
nvidia
FLUX.1 Dreambooth LoRA Fine-tuning
Fine-tune FLUX.1-dev 12B model using Dreambooth LoRA for custom image generation
nvidia
Generate Images and Videos with ComfyUI
Node-based diffusion workflows for images and videos with FLUX, Wan, HunyuanVideo, and Stable Diffusion
nvidia
How to Build a Multi-GPU AI PC - A Practical Guide
Many people explore local generative AI for privacy and to avoid token limits, but newer models require significant memory and compute—leading some to adopt multi-GPU setups.
nvidia
How to Fine-Tune an LLM on NVIDIA GPUs With Unsloth
Fine-tune popular AI models faster in Unsloth with NVIDIA RTX AI PCs, RTX PRO workstations, and DGX Spark—plus explore the new Nemotron Nano 3 family of open models.
nvidia
How to Get Started With Large Language Models on NVIDIA RTX PCs
Learn about using LLMs locally on PCs and workstations with Ollama, AnythingLLM, and LM Studio.
nvidia
Install and Use Isaac Sim and Isaac Lab
Build Isaac Sim and Isaac Lab from source for Spark
nvidia
Install and Use NVIDIA PAIR
Run local AI requests through PAIR and route independent Ollama or LM Studio requests across compatible systems.
nvidia
Isaac GR00T N1.6 Fine-Tuning
Fine-tune and benchmark NVIDIA's GR00T N1.6 robotics foundation model on DGX Station
nvidia
Live VLM WebUI
Real-time Vision Language Model interaction with webcam streaming
nvidia
LLaMA Factory
Install and fine-tune models with LLaMA Factory
nvidia
LM Studio on DGX Spark
Deploy LM Studio and serve LLMs on a Spark device; use LM Link to access models remotely.
nvidia
Local Coding Agent
Run local CLI coding agents with Claude Code and Ollama on DGX Station (NVIDIA GB300) using qwen3.6:27b
nvidia
Local Healthcare Agent on DGX Station
Run healthcare AI agents that analyze patient data and predict protein structures in an OpenShell sandbox on DGX Station
nvidia
Multi-modal Inference
Setup multi-modal inference with TensorRT
nvidia
NCCL for Multiple Sparks
Install and test NCCL on two, three, or four Sparks
nvidia
Nemotron Model Family on DGX Spark
Deploy Nemotron 3 model family (Nemotron-3-Nano or Nemotron-3-Super) on DGX Spark
nvidia
NIM on Spark
Deploy a NIM on Spark
nvidia
NVFP4 Pretraining with Megatron Bridge
Pretrain Llama 3.1 8B with NVFP4 mixed precision on DGX Station using Megatron Bridge
nvidia
Open WebUI with Ollama
Install Open WebUI and use Ollama to chat with models on your Spark
nvidia
Portfolio Optimization
GPU-Accelerated portfolio optimization using cuOpt and cuML
nvidia
Quantize Models to NVFP4 with NVIDIA Model Optimizer
Cut memory ~3.5× vs FP16 while keeping accuracy close to FP8, then validate with an OpenAI-compatible endpoint
nvidia
Register DGX Station to Brev
Link your DGX Station to Brev for remote access and sharing
nvidia
Run Hermes Agent with a Local LLM
Chat from the terminal against local vLLM with the self-improving Nous Research agent (Telegram optional)
nvidia
Run models with llama.cpp on DGX Spark
Build llama.cpp with CUDA and serve models via an OpenAI-compatible API
nvidia
Run NemoClaw with a Local LLM
Build a local AI assistant in an OpenShell sandbox with vLLM inference and optional Telegram
nvidia
Run OpenClaw with a Local LLM
Install a local-first AI agent and connect it to a private OpenAI-compatible model endpoint
nvidia
Secure AI Agents with OpenShell
Isolate OpenClaw with kernel-level policies and route inference to a local model
nvidia
Serve LLMs with SGLang
High-throughput serving with RadixAttention, structured output, and an OpenAI-compatible API
nvidia
Serve LLMs with vLLM
High-throughput serving for 30+ models, with continuous batching and an OpenAI-compatible API
nvidia
Set Up Local Network Access
NVIDIA Sync helps set up and configure SSH access
nvidia
Single-cell RNA Sequencing
An end-to-end GPU-powered workflow for scRNA-seq using RAPIDS
nvidia
Spark & Reachy Photo Booth
AI augmented photo booth using the DGX Spark and Reachy Mini.
nvidia
Topic Modeling
Extract insights from massive text datasets using cuML's GPU-accelerated BERTopic
nvidia
TRT LLM for Inference
Install and use TensorRT-LLM on DGX Spark
nvidia
Unsloth on DGX Spark
Optimized fine-tuning with Unsloth
nvidia
Vibe Coding in VS Code
Use DGX Spark as a local or remote Vibe Coding assistant with Ollama and Continue
nvidia
Vision-Language Model Fine-tuning
Fine-tune Vision-Language Models for image and video understanding tasks using Qwen2.5-VL and InternVL3
nvidia
🦞 Set Up Example NemoClaw Agents 🦞
Ready-to-run application examples for your NemoClaw sandbox — policy, prompt, and personalization for each workflow
nvidia
Build and Deploy a Multi-Agent Chatbot
Deploy a multi-agent chatbot system and chat with agents on your Spark
nvidia
Build Knowledge Graphs with txt2kg
Extract triples with Ollama or vLLM, store them in a graph database, and explore them in a GPU-accelerated web UI
nvidia
CLI Coding Agent
Build local CLI coding agents with Ollama
nvidia
Connect Multiple DGX Spark through a Switch
Set up a cluster of DGX Spark devices that are connected through Switch
nvidia
Connect Three DGX Spark in a Ring Topology
Connect and set up three DGX Spark devices in a ring topology
nvidia
Connect Two DGX Stations for Distributed Workloads
Combined memory and compute over a direct high-speed link
nvidia
Connect Two Sparks
Connect two Spark devices and setup them up for inference and fine-tuning
nvidia
CUDA-X Data Science
Install and use NVIDIA cuML and NVIDIA cuDF to accelerate UMAP, HDBSCAN, pandas and more with zero code changes
nvidia
cuTile Kernels
Run cuTile kernel benchmarks, FMHA implementation, and LLM inference on DGX Spark and B300
nvidia
DGX Dashboard
Monitor your DGX system and launch JupyterLab
nvidia
DGX Station AI Skills and dgx-assist
Inspect DGX Station software and route version-aware, CLI-backed workflows
nvidia
Fine-tune with NeMo
Use NVIDIA NeMo to fine-tune models locally
nvidia
FLUX.1 Dreambooth LoRA Fine-tuning
Fine-tune FLUX.1-dev 12B model using Dreambooth LoRA for custom image generation
nvidia
Generate Images and Videos with ComfyUI
Node-based diffusion workflows for images and videos with FLUX, Wan, HunyuanVideo, and Stable Diffusion
nvidia
How to Build a Multi-GPU AI PC - A Practical Guide
Many people explore local generative AI for privacy and to avoid token limits, but newer models require significant memory and compute—leading some to adopt multi-GPU setups.
nvidia
How to Fine-Tune an LLM on NVIDIA GPUs With Unsloth
Fine-tune popular AI models faster in Unsloth with NVIDIA RTX AI PCs, RTX PRO workstations, and DGX Spark—plus explore the new Nemotron Nano 3 family of open models.
nvidia
How to Get Started With Large Language Models on NVIDIA RTX PCs
Learn about using LLMs locally on PCs and workstations with Ollama, AnythingLLM, and LM Studio.
nvidia
Install and Use Isaac Sim and Isaac Lab
Build Isaac Sim and Isaac Lab from source for Spark
nvidia
Install and Use NVIDIA PAIR
Run local AI requests through PAIR and route independent Ollama or LM Studio requests across compatible systems.
nvidia
Isaac GR00T N1.6 Fine-Tuning
Fine-tune and benchmark NVIDIA's GR00T N1.6 robotics foundation model on DGX Station
nvidia
Live VLM WebUI
Real-time Vision Language Model interaction with webcam streaming
nvidia
LLaMA Factory
Install and fine-tune models with LLaMA Factory
nvidia
LM Studio on DGX Spark
Deploy LM Studio and serve LLMs on a Spark device; use LM Link to access models remotely.
nvidia
Local Coding Agent
Run local CLI coding agents with Claude Code and Ollama on DGX Station (NVIDIA GB300) using qwen3.6:27b
nvidia
Local Healthcare Agent on DGX Station
Run healthcare AI agents that analyze patient data and predict protein structures in an OpenShell sandbox on DGX Station
nvidia
Multi-modal Inference
Setup multi-modal inference with TensorRT
nvidia
NCCL for Multiple Sparks
Install and test NCCL on two, three, or four Sparks
nvidia
Nemotron Model Family on DGX Spark
Deploy Nemotron 3 model family (Nemotron-3-Nano or Nemotron-3-Super) on DGX Spark
nvidia
NIM on Spark
Deploy a NIM on Spark
nvidia
NVFP4 Pretraining with Megatron Bridge
Pretrain Llama 3.1 8B with NVFP4 mixed precision on DGX Station using Megatron Bridge
nvidia
Open WebUI with Ollama
Install Open WebUI and use Ollama to chat with models on your Spark
nvidia
Portfolio Optimization
GPU-Accelerated portfolio optimization using cuOpt and cuML
nvidia
Quantize Models to NVFP4 with NVIDIA Model Optimizer
Cut memory ~3.5× vs FP16 while keeping accuracy close to FP8, then validate with an OpenAI-compatible endpoint
nvidia
Register DGX Station to Brev
Link your DGX Station to Brev for remote access and sharing
nvidia
Run Hermes Agent with a Local LLM
Chat from the terminal against local vLLM with the self-improving Nous Research agent (Telegram optional)
nvidia
Run models with llama.cpp on DGX Spark
Build llama.cpp with CUDA and serve models via an OpenAI-compatible API
nvidia
Run NemoClaw with a Local LLM
Build a local AI assistant in an OpenShell sandbox with vLLM inference and optional Telegram
nvidia
Run OpenClaw with a Local LLM
Install a local-first AI agent and connect it to a private OpenAI-compatible model endpoint
nvidia
Secure AI Agents with OpenShell
Isolate OpenClaw with kernel-level policies and route inference to a local model
nvidia
Serve LLMs with SGLang
High-throughput serving with RadixAttention, structured output, and an OpenAI-compatible API
nvidia
Serve LLMs with vLLM
High-throughput serving for 30+ models, with continuous batching and an OpenAI-compatible API
nvidia
Set Up Local Network Access
NVIDIA Sync helps set up and configure SSH access
nvidia
Single-cell RNA Sequencing
An end-to-end GPU-powered workflow for scRNA-seq using RAPIDS
nvidia
Spark & Reachy Photo Booth
AI augmented photo booth using the DGX Spark and Reachy Mini.
nvidia
Topic Modeling
Extract insights from massive text datasets using cuML's GPU-accelerated BERTopic
nvidia
TRT LLM for Inference
Install and use TensorRT-LLM on DGX Spark
nvidia
Unsloth on DGX Spark
Optimized fine-tuning with Unsloth
nvidia
Vibe Coding in VS Code
Use DGX Spark as a local or remote Vibe Coding assistant with Ollama and Continue
nvidia
Vision-Language Model Fine-tuning
Fine-tune Vision-Language Models for image and video understanding tasks using Qwen2.5-VL and InternVL3
nvidia
🦞 Set Up Example NemoClaw Agents 🦞
Ready-to-run application examples for your NemoClaw sandbox — policy, prompt, and personalization for each workflow