Data & Analysis ✓ NVIDIA · Official anomaly-generationsynthetic-dataimage-generationfine-tuninggpu-computingtao

PAIDF AnomalyGen

Fine-tune, generate, evaluate, and refine synthetic anomaly images.

FollowSkills review · FSRS-2.0
Not recommended
47/ 100 5-point scale 2.4 / 5
Trust13 / 25 · 2.6/5

The document limits the declared tools to Read and Bash and provides a non-root container, mount-permission preflight, path validation, and confirmation before writing the training config. However, it performs Docker and GPU operations, downloads about 140 GB of checkpoints, and passes HF_TOKEN without sufficient disclosure of token exposure, network data flow, cost confirmation, rollback, or cleanup, so points are deducted.

Reliability8 / 20 · 2.0/5

The phases, parameter validation, output verification, and several failure modes are documented in useful detail. Static evidence provides no tests or reproducible CI covering this skill's key paths, while background training, long-running execution, and recovery from some failures still shift work to the user, so the score remains below the static ceiling.

Adaptability8 / 15 · 2.7/5

The audience, three modes, inputs, outputs, and several non-fit conditions are reasonably clear. The skill nevertheless depends tightly on a CUDA GPU, Docker/NVIDIA Container Toolkit, a specific container, and external data services, with no Chinese-language or mainland-China reachability guidance; core downloads may be inaccessible, so points are deducted.

Convention9 / 15 · 3.0/5

The documentation includes a phase table, quick start, parameter tables, reference-document routing, error handling, output layout, Apache-2.0 licensing, and version 0.1.0. It lacks standard Instructions/Examples sections, an author email, a changelog, and a clearly stated maintenance owner/update path; the supplied benchmark also reports hierarchy and schema findings, so points are deducted.

Effectiveness6 / 15 · 2.0/5

The skill describes an end-to-end workflow covering fine-tuning, generation, evaluation, search, filtering, regeneration, and declared CSV/log outputs. However, there are no statically verifiable representative outputs for the core pipeline, the workflow is costly, and results still require interpretation; the benchmark covers only a planning task rather than actual pipeline completion, so the score is limited.

Verifiability3 / 10 · 1.5/5

A committed benchmark report and evaluation specification provide some evidence, with limited reported security, correctness, and discoverability results. There is only one non-executing planning task and no tests, CI, or independent corroboration covering PAIDF's key phases, preventing a higher static verifiability score.

Evidence confidence:Low Reviewed Jul 29, 2026 Reviewed revision ce70ca7f1966
The upstream repository has new commits since this review. The score still applies to the reviewed revision shown and may not cover the latest changes.
Before you use it
  • Before running, explicitly confirm GPU, disk, time, and roughly 140 GB of download costs, and avoid exposing HF_TOKEN through shell history or container inspection in untrusted environments.
  • Core preparation and checkpoint retrieval depend on Hugging Face, Roboflow, GitHub, and a container image; mainland-China network restrictions may prevent completion.
  • The document launches training in the background but does not provide a complete process-monitoring, cancellation, checkpoint-recovery, or resource-cleanup procedure.
  • The benchmark contains only one planning task and does not establish that the actual generation, evaluation, search, and regeneration pipeline runs successfully.
See the full review method →

What does this skill do, and when should you use it?

This skill provides the full PAIDF AnomalyGen multi-phase pipeline for fine-tuning on a new anomaly dataset, generating synthetic anomaly images, and evaluating output quality. It supports full, finetune_only, and inference_only modes. The pipeline can search per-sample guidance and crop_ratio parameters, then filter and regenerate samples by nn_score. It requires Docker, NVIDIA Container Toolkit, a CUDA GPU, and may download about 140 GB of pretrained checkpoints.

The skill reads a dataset root, defect-specification JSONL, and optionally an existing checkpoint; verifies or downloads checkpoints; prepares training and inference JSONL; and runs fine-tuning, SDG generation, evaluation, per-sample parameter search, assembly, filtering, and regeneration. It executes shell scripts and Python utilities in Docker or the supported host environment. Outputs include original, rounds, searched, and regens directories, with files such as SDG_result.csv, per_sample.csv, eval.log, search_summary.csv, and regen_summary.csv.

  1. A computer-vision team fine-tunes AnomalyGen on a new dataset containing texture, text, or CAD defects.
  2. A data engineer reuses an existing fine-tuned checkpoint to generate anomaly images without running Phase 1.
  3. A quality team evaluates generated images using nn_score, mnn_score, and FID outputs.
  4. A researcher searches guidance and crop_ratio values separately for each sample.
  5. An operator filters low-scoring samples by nn_threshold and regenerates them automatically.

What are this skill's strengths and limitations?

Pros
  • Covers fine-tuning, generation, evaluation, search, assembly, filtering, and regeneration in one workflow.
  • Supports checkpoint reuse through inference_only as well as fine-tuning-only execution.
  • Produces per-sample and aggregate evaluation artifacts for verification.
  • Supports 2b and 14b models, plus multi-GPU fine-tuning and SDG.
Limitations
  • Requires a CUDA GPU, Docker, and NVIDIA Container Toolkit; host execution additionally requires the cosmos-predict2 environment.
  • Checkpoint downloads can total about 140 GB and require HF_TOKEN.
  • Execution depends on dataset structure, defect specifications, mask files, and mount permissions.
  • The source provides no test-suite or broader platform-coverage evidence for this individual skill.

How do you install this skill?

Install the skill through the repository README's skills CLI flow:
npx skills add nvidia/skills --skill paidf-anomalygen --yes
The skill becomes available when the agent next loads skills and encounters a relevant task. The source does not require separately cloning the skill directory.

How do you use this skill?

Ask the agent to “fine-tune AnomalyGen,” “generate anomaly images,” “run PAIDF SDG,” “evaluate SDG output quality,” or “run per-sample search.” Before execution, provide mode, name, dataset_dir, defect_spec, and num_SDG; inference_only additionally requires both checkpoint_dir and step. Run commands from the repository root with ANOMALYGEN_SCRIPTS exported. Full mode requires reading both the fine-tuning and inference references, and the pipeline is intended to run without mid-run pauses.

FAQ

Does this skill require network access?
It may need network access to verify or download pretrained checkpoints; the quick start describes an approximately 140 GB download requiring HF_TOKEN.
Can I generate images without fine-tuning again?
Yes. Use inference_only with an existing checkpoint_dir and step, supplied together.
Can I run only fine-tuning?
Yes. finetune_only runs Phases 0–1 and ignores num_SDG.
What failure cases are documented?
The source documents missing mask directories, empty or short AMP output, mid-round SDG failures, and off-boundary step values, with troubleshooting references for these cases.

More skills from this repository

All from NVIDIA/skills

Data & Analysis ✓ NVIDIA · Official

NVIDIA Physical AI Defect Image Generation

Orchestrate defect-image generation, augmentation, inference, and labeling for AOI datasets on OSMO.

Data & Analysis ✓ NVIDIA · Official

NV-Generate-MR

Generate synthetic body MRI volumes through NVIDIA’s rflow-mr workflow.

Data & Analysis ✓ NVIDIA · Official

cuPyNumeric Parallel Shard Loader

Builds processor-sized parallel loading paths from sharded on-disk data into distributed cuPyNumeric arrays.

Data & Analysis ✓ NVIDIA · Official

NeMo Data Designer Synthetic Data Skill

Build synthetic datasets and declarative data-generation pipelines from a natural-language description.

Dev & Engineering ✓ NVIDIA · Official

NeMo MBridge Recipe Recommender

Recommends adjustable Megatron Bridge recipes for your model, GPU budget, and training goal.

Data & Analysis ✓ NVIDIA · Official

Data Designer Synthetic Data Skill

Build synthetic datasets and data-generation pipelines from a natural-language specification.

Data & Analysis ✓ NVIDIA · Official

TAO Standard Training Workflow

Run a controlled TAO train, evaluate, and export workflow on labeled data.

Data & Analysis ✓ NVIDIA · Official

Synthetic Brain MRI Generator

Generate synthetic brain MRI volumes through NVIDIA’s documented workflow.

Data & Analysis ✓ NVIDIA · Official

Nemotron Retrieval Recipes

Plan, debug, tune, evaluate, export, and deploy Nemotron embedding and reranking recipes.

Data & Analysis ✓ NVIDIA · Official

Nemotron Customization Pipelines

Plan, configure, and chain existing Nemotron steps for data preparation, training, alignment, conversion, optimization, and evaluation.

Dev & Engineering ✓ NVIDIA · Official

TAO Platform Execution SDK

Submit, monitor, and manage NVIDIA TAO GPU training jobs across supported platforms.

Dev & Engineering ✓ NVIDIA · Official

Clinical ASR Flywheel: Environment Setup

Validate that a clinical ASR evaluation environment can complete a TTS-to-ASR round trip through NVIDIA-hosted speech services.

Dev & Engineering ✓ NVIDIA · Official

NVIDIA Skill Finder

Finds the right NVIDIA agent skill for product, hardware, and workflow requests.

Data & Analysis ✓ NVIDIA · Official

Cosmos-Embed Video Retrieval & Fine-Tuning

Use Cosmos-Embed1 to embed, retrieve, deduplicate, and fine-tune video-text data.

Data & Analysis ✓ NVIDIA · Official

NV Brain MRI Diffusion Fine-Tuner

Fine-tune the NV-Generate-CTMR diffusion UNet from NIfTI brain MRI training volumes.

Data & Analysis ✓ NVIDIA · Official

Clinical ASR Fine-Tuning

Fine-tune Parakeet TDT v2 with NeMo for clinical vocabulary and close the loop with offline KER evaluation.

Automation & Ops ✓ NVIDIA · Official

Isaac for Healthcare E2E Workflow

Run the Isaac for Healthcare pipeline from recording through validation.

Dev & Engineering ✓ NVIDIA · Official

DALI Dynamic Mode Assistant

Helps agents write, review, and migrate NVIDIA DALI imperative dynamic-mode code.

Dev & Engineering ✓ NVIDIA · Official

cuPyNumeric Migration Readiness

Assess whether NumPy code is ready to scale on GPUs before committing to a substantial cuPyNumeric port.

Data & Analysis ✓ NVIDIA · Official

Nemotron Speech ASR Customization Orchestrator

Chooses and coordinates the lowest-cost path for adapting speech recognition to a domain or language.

Related skills