NVIDIA Dev & Engineering Skills Reviews — Page 5
Jetson Memory Optimizer
Reclaim unused DRAM on headless or no-camera Jetson deployments.
Megatron-Core Testing Guide
Design, run, and reproduce distributed Megatron-Core tests with CI-aligned workflows.
MoE Dispatcher Selection Guide
Choose alltoall, DeepEP, or HybridEP from the hardware, expert-parallel degree, and optimization stage.
MoE Hardware Configuration Reference
Plan MoE training layouts and throughput expectations across NVIDIA GPU platforms.
NV-Reason-CXR Chest X-ray Reasoning Test
Runs reproducible command-shape and live inference smoke tests for chest X-ray reasoning.
Omniverse USD Performance Tuning
Diagnose USD scene performance problems and coordinate evidence-based optimization from runtime setup through reporting.
CuTile Autotuning Guide
Design, implement, and validate low-overhead autotuning for CuTile GPU kernels.
Catheter Navigation DRR Renderer
Generate a single digitally reconstructed radiograph from CT data or a synthetic phantom for fluoroscopy previews and renderer smoke tests.
i4h Catheter Navigation Workflow Guide
Orients agents to the right stage of NVIDIA’s endovascular catheter navigation workflow.
I4H Scene Editing Workflow
Edit existing Isaac for Healthcare scenes live through a controlled bridge session.
Isaac for Healthcare Workflow Validator
Run policy or state-machine rollouts and record verification episodes to HDF5.
DeepStream SOP Compliance Inference
Build and operate a GPU-accelerated service that checks industrial work-step order and SOP compliance in video.
DOCA PCC Custom Congestion Control
Guides host-side loading and troubleshooting of custom PCC algorithms on BlueField
DOCA Rivermax Receive Development
Guides developers through building, validating, and debugging DOCA Rivermax receive applications for real-time network streams.
Jetson Pinmux Customizer
Generate Jetson carrier pinmux BCT configuration from an XLSM.
Megatron-Core Linting & Formatting
Standardize Megatron-LM Python checks and formatting before code review.
NeMo MBridge Hierarchical Context Parallelism
Enable, troubleshoot, and verify hierarchical context parallelism in Megatron-Bridge.
NeMo-RL Auto Research
Turns NeMo-RL research goals into reproducible, iterative experiments.
NeMo-RL Session Memory
Preserve coding-agent context across interruptions, restarts, and handoffs.
TAO Hugging Face Model Integrator
Connect Hugging Face computer-vision models to NVIDIA TAO training, ONNX export, and TensorRT deployment workflows.
TileGym Transformers Kernel Integration
Integrate TileGym kernels into Hugging Face Transformers models without editing Transformers source.
Catheter Navigation Setup Verification
Checks host, GPU, and Python path readiness for catheter navigation.
DOCA DPA High-Level Tracer
Capture and decode BlueField DPA programming events to diagnose kernel behavior, communication latency, and trace failures.
DOCA GPUNetIO WRITE Bandwidth Benchmark
Build, run, and interpret sustained RDMA WRITE bandwidth measurements driven by CUDA kernels through the DOCA GPUNetIO surface.