NVIDIA Claude Skills Reviews — Page 5
NeMo Relay Call Instrumentation
Wrap tool and LLM calls with Relay lifecycle events, middleware, and guardrails.
NeMo Relay Context Isolation
Keeps NeMo Relay scope stacks independent across concurrent requests and async workflows while preserving ancestry propagation.
NeMo Relay Typed Wrappers & Codecs
Add typed boundaries to NeMo Relay integrations while preserving predictable JSON middleware semantics and caller-visible behavior.
Omniverse Realtime Viewer
Routes USD viewer requests to the right architecture and guides rendering, interaction, UI, and validation.
NVIDIA Physical AI Defect Image Generation
Orchestrate defect-image generation, augmentation, inference, and labeling for AOI datasets on OSMO.
NVIDIA NuRec Neural Reconstruction Router
Routes NuRec requests to the right upstream workflow skill.
RAG Performance Benchmark
Benchmark a deployed NVIDIA RAG Blueprint server and expose latency, throughput, and bottleneck behavior from one YAML config.
TAO CLIP Fine-Tuning and Deployment
Train and deploy NVIDIA TAO CLIP for image-text retrieval, zero-shot classification, and embedding extraction.
TAO on SLURM
Run TAO training, evaluation, and inference on a remote GPU cluster through SSH, SLURM, and Pyxis/Enroot.
TAO Platform Execution SDK
Submit, monitor, and manage NVIDIA TAO GPU training jobs across supported platforms.
TAO NVIDIA GPU Host Setup
Checks and standardizes NVIDIA drivers, CUDA, and container runtime prerequisites for TAO GPU hosts.
TAO FastFoundationStereo Real-Time Stereo Depth
Run TAO FastFoundationStereo workflows that produce low-latency disparity maps from stereo images.
VSS Video Summarizer
Creates timestamped narrative summaries of recorded videos through LVS, with a VLM fallback.
AutoMagicCalib Calibration Stack Launcher
Deploy the AutoMagicCalib microservice and web UI with Docker Compose for a ready-to-use camera calibration stack.
CUDA-Q Quantum Onboarding
Guides developers from CUDA-Q installation to quantum kernels, GPU simulation, and real QPU execution.
cuOpt Numerical Optimization API
Guide agents through GPU-accelerated LP, MILP, and QP modeling with cuOpt
cuOpt Numerical Optimization Formulation
Turn natural-language problems into clear LP, MILP, or QP mathematical formulations.
cuPyNumeric Parallel Shard Loader
Builds processor-sized parallel loading paths from sharded on-disk data into distributed cuPyNumeric arrays.
Clinical ASR Flywheel: Environment Setup
Validate that a clinical ASR evaluation environment can complete a TTS-to-ASR round trip through NVIDIA-hosted speech services.
DOCA Bench Benchmarking Skill
Measure DOCA library throughput, latency, and bandwidth reproducibly on real NVIDIA networking hardware.
DOCA CollectX Telemetry Deployment
Deploy, operate, and debug CollectX telemetry collectors so counters reach downstream exporters.
DOCA DMA Development Guide
Guides hands-on DOCA DMA memory-copy development on BlueField and ConnectX systems.
DOCA Flow Performance Measurement
Guides reproducible, defensible measurements of host-side and DPU-CPU DOCA Flow control-plane rule rates.
DOCA Flow Tune
Guides engineers through snapshotting, analyzing, and optimizing live or captured DOCA Flow pipelines.