All Skills — Page 14
DOCA Ethernet Queue Development
Develop and debug DOCA Ethernet RX/TX queues on BlueField DPUs and ConnectX NICs.
DOCA GPUNetIO Development Skill
Helps developers connect CUDA kernels on NVIDIA GPUs to DOCA network queues for GPU-side packet I/O and debugging.
DOCA STA Storage Target Acceleration
Build and debug RDMA NVMe-oF storage targets accelerated by DOCA STA on BlueField.
DOCA Structured Tools Contract
A governed fallback contract for consolidating DOCA environment, device, capability, validation, and host-DPU state data.
DOCA Telemetry Exporter Development
Guides DOCA applications in defining, emitting, and debugging structured telemetry for external consumers.
DOCA Upgrade Control
Safely gate DOCA upgrades and rollbacks with explicit confirmation.
Holoscan SDK Setup Guide
Inspects a Linux host and selects the most suitable Holoscan SDK installation path.
Jetson USB Port Customization
Safely enable, disable, or change Jetson USB port roles through a kernel device-tree overlay.
Jetson Headless Memory Optimizer
Reclaim memory on GUI-free Jetson devices through safe, reversible system-service changes.
NeMo AutoModel Recipe Development
Build, modify, and validate NeMo AutoModel training and evaluation recipes.
NeMo MBridge CPU Offload
Configure, validate, and troubleshoot CPU offloading in Megatron Bridge to relieve GPU memory pressure.
NeMo MBridge CUDA Graphs
Configure, validate, and benchmark CUDA Graph capture in Megatron Bridge to reduce host-driver overhead.
NeMo Relay Call Instrumentation
Wrap tool and LLM calls with Relay lifecycle events, middleware, and guardrails.
NeMo Relay Context Isolation
Keeps NeMo Relay scope stacks independent across concurrent requests and async workflows while preserving ancestry propagation.
NeMo Relay Typed Wrappers & Codecs
Add typed boundaries to NeMo Relay integrations while preserving predictable JSON middleware semantics and caller-visible behavior.
Omniverse Realtime Viewer
Routes USD viewer requests to the right architecture and guides rendering, interaction, UI, and validation.
NVIDIA Physical AI Defect Image Generation
Orchestrate defect-image generation, augmentation, inference, and labeling for AOI datasets on OSMO.
NVIDIA NuRec Neural Reconstruction Router
Routes NuRec requests to the right upstream workflow skill.
RAG Performance Benchmark
Benchmark a deployed NVIDIA RAG Blueprint server and expose latency, throughput, and bottleneck behavior from one YAML config.
TAO CLIP Fine-Tuning and Deployment
Train and deploy NVIDIA TAO CLIP for image-text retrieval, zero-shot classification, and embedding extraction.
TAO on SLURM
Run TAO training, evaluation, and inference on a remote GPU cluster through SSH, SLURM, and Pyxis/Enroot.
TAO Platform Execution SDK
Submit, monitor, and manage NVIDIA TAO GPU training jobs across supported platforms.
TAO NVIDIA GPU Host Setup
Checks and standardizes NVIDIA drivers, CUDA, and container runtime prerequisites for TAO GPU hosts.
TAO FastFoundationStereo Real-Time Stereo Depth
Run TAO FastFoundationStereo workflows that produce low-latency disparity maps from stereo images.