Cluster Status
Check running workloads
Section titled “Check running workloads”sparkrun cluster statussparkrun status # aliasThis shows all sparkrun containers running on your cluster hosts, grouped by job.
Output
Section titled “Output”The status command displays:
- Running jobs — containers grouped by cluster ID with role (head/worker/solo), host, status, and container image
- IP mapping — when InfiniBand was detected during launch, shows the complementary IP alongside each host (e.g. management IP if you defined the cluster by IB IPs, or vice versa)
- Idle hosts — hosts that have no sparkrun containers running
- Pending operations — downloads or image distributions currently in progress that haven’t launched containers yet, with elapsed time and a note that they will consume VRAM once launched
- Copy-paste commands — ready-to-use
sparkrun logsandsparkrun stopcommands for each running job
Example output
Section titled “Example output”$ sparkrun cluster statusJob: qwen3.5-35b-a3b-fp8-sglang (tp=2) (2 container(s)) node_0 10.24.11.13 (ib: 192.168.11.13) Up 35 seconds scitrera/dgx-spark-sglang:0.5.9-dev1-329817e2-t5 node_1 10.24.11.14 (ib: 192.168.11.14) Up 9 seconds scitrera/dgx-spark-sglang:0.5.9-dev1-329817e2-t5 logs: sparkrun logs qwen3.5-35b-a3b-fp8-sglang --hosts 10.24.11.13,10.24.11.14 --tp 2 stop: sparkrun stop qwen3.5-35b-a3b-fp8-sglang --hosts 10.24.11.13,10.24.11.14 --tp 2
Idle hosts (no sparkrun containers): 10.24.11.16 10.24.11.17
Total: 2 container(s) across 4 host(s)Options
Section titled “Options”# Check a specific clustersparkrun cluster status --cluster mylab
# Check specific hostssparkrun cluster status --hosts 192.168.11.13,192.168.11.14If no options are given, uses the default cluster.