Choose the right Cluster Validation Suite (CVS) config template for your workload#
2026-09-25
5 min read time
For the complete field-by-field schema for each test suite’s config file, see Cluster Validation Suite (CVS) test configuration files reference.
Use cvs config copy <path> --output <dest> to copy a template, or cvs config list <path> to browse templates in a directory.
Platform, health, RCCL, and other diagnostic/network configs use fixed filenames (see Burn-in / Diag and Network below). Training and inference workloads use the naming patterns in their respective sections.
Threshold pairs#
Some training and inference suites ship a matching threshold file for each config: the same basename with _threshold inserted before .json (for example …_single_threshold.json). Copy both files and keep them in the same directory. The config references its threshold file via a threshold_json field.
Burn-in / Diag#
The following suites are available for burn-in and diagnostic workloads:
Suite |
Config location |
List / README |
|---|---|---|
Platform |
|
|
Health |
|
|
Preflight |
|
|
Network#
The following suites are available for network testing:
Suite |
Config location |
List / README |
|---|---|---|
IB Perf |
|
|
RCCL |
|
|
MORI |
|
|
Training#
Training workload templates use:
{gpu}_{framework}_{model}_{mode}.json
Segment |
Meaning |
|---|---|
|
Target GPU architecture — for example |
|
Training stack — for example |
|
Model identifier — for example |
|
Topology — |
Example filenames#
input/config_file/training/jaxmaxtext/
├── mi3xx_jaxmaxtext_llama-3.3-70b_single.json
└── mi3xx_jaxmaxtext_llama-3.3-70b_distributed.json
mi3xx_jaxmaxtext_llama-3.3-70b_distributed.json → mi3xx · jaxmaxtext · llama-3.3-70b · distributed.
Available suites#
Suite |
Config location |
List / README |
|---|---|---|
JAX MaxText |
|
|
Megatron |
|
|
TorchTitan |
|
|
Aorta |
|
|
Inference#
Inference templates use one of two filename patterns:
{gpu}_{framework}_{model}_{mode}.json
{gpu}_{framework}_{model}_{precision}_{mode}.json
Use the second form when precision (fp8, mxfp4, bf16, and similar) is a separate token after the model name.
Segment |
Meaning |
|---|---|
|
Target GPU architecture — for example |
|
Inference stack — for example |
|
Model identifier — for example |
|
Optional quantization or dtype token — for example |
|
Topology or workload shape — |
Example filenames#
Without {precision}:
input/config_file/inference/xdit/
├── mi3xx_pytorch_xdit_flux1_dev_single.json
└── mi3xx_pytorch_xdit_wan22_14b_single.json
mi3xx_pytorch_xdit_flux1_dev_single.json → mi3xx · pytorch_xdit · flux1_dev · single.
With {precision}:
input/config_file/inference/vllm/
├── mi3xx_vllm_llama33-70b_fp8_single.json
└── mi3xx_vllm_llama33-70b_fp8_distributed.json
mi3xx_vllm_llama33-70b_fp8_single.json → mi3xx · vllm · llama33-70b · fp8 · single.
Available suites#
Suite |
Config location |
List / README |
|---|---|---|
ATOM |
|
|
vLLM |
|
|
SGLang |
|
|
xDiT |
|
|