Skip to content

Command reference

This page is generated from the same file the build validates commands against. Every shell block and every lab script on the site is parsed, and every option passed to a tool listed here is checked against this list. An option that does not exist fails the build.

The list is deliberately not a manual. It records which options exist, so that the course cannot invent one; what each option does is taught on the page that uses it, linked from each entry.

33 tools, 0 with option lists captured from the pinned version's help output, generated 2026-09-08.

Tools marked seed have option lists written from documentation rather than captured from --help; the build treats an unknown option on such a tool as a warning, not an error, until the capture is made on the lab machines.

aiderAider 0.86.0 · seed

-h --help --model --edit-format --weak-model --api-key --set-env --message -m --test-cmd --auto-test --map-tokens --read --git --no-git --auto-commits --no-auto-commits --yes-always --cache-prompts --verify-ssl --model-settings-file --openai-api-base --openai-api-key --no-show-model-warnings

codexOpenAI Codex CLI 0.153.4 · seed

-h --help --oss --local-provider --sandbox --ask-for-approval -a --dangerously-bypass-approvals-and-sandbox --yolo --full-auto --image --search --config -c --model -m --cd -C

Subcommands: exec, resume, cloud, mcp, login, logout, completion, app

exoexo main · seed

-h --help -q --quiet -v --verbose -m --force-master --no-api --api-port --no-worker --no-downloads --offline --no-batch --legacy-daemon --bootstrap-peers --namespace --zenoh-port --discovery-port --fast-synch --no-fast-synch

ggml-rpc-serverllama.cpp v0.4.0 · seed

--help -h -H --host -p --port -c --cache -t --threads -d --device

iperf3iperf3 3.21 · seed

-h --help -s --server -c --client -p --port -t --time -P --parallel -b --bandwidth -u --udp -R --reverse -J --json -i --interval -B --bind -A --affinity -V --verbose -4 -6 -v --version --bitrate --bidir -l --length -w --window -M --set-mss -Z --zerocopy --logfile -D --daemon -1 --one-off

llama-benchllama.cpp v0.4.0 · seed

--help -h --version -m --model -p --n-prompt -n --n-gen -pg -b --batch-size -ub --ubatch-size -ngl --n-gpu-layers -t --threads -r --repetitions -o --output --numa -sm --split-mode -ts --tensor-split -fa --flash-attn --mmap --mlock -v --verbose -ctk --cache-type-k -ctv --cache-type-v --rpc -dev --device -rpc --list-devices --no-warmup

llama-clillama.cpp v0.4.0 · seed

--help -h --version -m --model -p --prompt -f --file -n --predict -c --ctx-size -b --batch-size -ngl --n-gpu-layers -t --threads --temp --top-k --top-p --min-p --repeat-penalty -i --interactive --interactive-first -cnv --conversation --color -s --seed --mlock --no-mmap --mirostat --grammar --grammar-file --chat-template -ins --instruct --multiline-input --in-prefix --in-suffix -r --reverse-prompt --log-disable --verbose-prompt --simple-io --lora --control-vector -e --escape --numa -fa --flash-attn --rpc -ts --tensor-split -dev --device

llama-imatrixllama.cpp v0.4.0 · seed

--help -h -lv --verbosity -m --model -f --file -o --output-file -ofreq --output-frequency --output-format --save-frequency --process-output --in-file --parse-special --chunk --from-chunk --chunks --no-ppl --show-statistics -ngl --n-gpu-layers

llama-perplexityllama.cpp v0.4.0 · seed

--help -h -m --model -f --file -ngl --n-gpu-layers -c --ctx-size --kl-divergence-base --kl-divergence

llama-serverllama.cpp v0.4.0 · seed

--help -h --version -m --model -mu --model-url -hf --hf-repo --hf-file -a --alias -t --threads -tb --threads-batch -c --ctx-size -n --predict -b --batch-size -ub --ubatch-size -np --parallel --cont-batching -ngl --n-gpu-layers -sm --split-mode -ts --tensor-split -mg --main-gpu --mlock --no-mmap --numa --host --port --path --api-key --api-key-file --embedding --embeddings --reranking --metrics --slots --props -to --timeout --chat-template --chat-template-file --jinja -fa --flash-attn --rope-scaling --rope-freq-base --rope-freq-scale --yarn-orig-ctx --yarn-ext-factor --yarn-attn-factor --cache-type-k --cache-type-v --no-context-shift --grp-attn-n --grp-attn-w --log-disable -v --verbose -s --seed --temp --top-k --top-p --min-p --repeat-penalty --repeat-last-n --presence-penalty --frequency-penalty --mirostat --grammar --grammar-file --json-schema --lora --lora-scaled --control-vector -sp --special --verbose-prompt --spec-draft-model -md --spec-draft-n-max --spec-draft-n-min --spec-draft-p-min -ngld --spec-type --cache-prompt --no-cache-prompt --slot-save-path -ctk -ctv --no-slots --no-metrics --rpc -dev --device --gpu-layers --cache-reuse --reasoning-format --reasoning-budget

mlx_lm.cache_promptunpinned · seed

--model --prompt --prompt-cache-file --max-kv-size --kv-bits --quantized-kv-start

mlx_lm.fuseunpinned · seed

--model --save-path --adapter-path --hf-path --upload-repo --dequantize --export-gguf --gguf-path

mlx_lm.generatemlx-lm 0.31.3 · seed

--model --prompt --max-tokens --temp --top-p --seed --adapter-path --trust-remote-code --colorize --eos-token --draft-model --num-draft-tokens --prompt-cache-file --max-kv-size --kv-bits --kv-group-size --system-prompt --chat-template-config --ignore-chat-template --verbose --quantized-kv-start

mlx.distributed_configunpinned · seed

--verbose --hosts --ignore-unreachable --hostfile --over --output-hostfile --auto-setup --no-auto-setup --dot --backend --env

mlx.launchmlx-lm 0.31.3 · seed

--hosts --backend --hostfile -n --verbose --label --env --print-python --repeat-hosts --mpi-arg --connections-per-ip --starting-port -p --cwd --nccl-port --python

opencodeOpenCode 1.18.29 · seed

-h --help -v --version --print-logs --log-level --pure

Subcommands: run, auth, models, serve, mcp, agent, session, upgrade

python -m sglang.launch_serverSGLang 0.5.19 · seed

--model-path --host --port --tp-size --mem-fraction-static --context-length --quantization --trust-remote-code --chat-template --api-key --served-model-name --enable-torch-compile --grammar-backend --tool-call-parser --reasoning-parser --attention-backend --speculative-algorithm --speculative-draft-model-path --speculative-num-steps --speculative-eagle-topk --speculative-num-draft-tokens --disable-radix-cache --max-running-requests --chunked-prefill-size --enable-metrics --log-level --log-requests --kv-cache-dtype --dtype --tensor-parallel-size --model --disaggregation-mode --disaggregation-transfer-backend --disaggregation-bootstrap-port --disaggregation-ib-device --base-gpu-id

rayunpinned · seed

-h --help --version

Subcommands: start, stop, status, list, up, down, attach

rpc-serverllama.cpp v0.4.0 · seed

--help -h -H --host -p --port -c --cache -t --threads -d --device

trtllm-serveTensorRT-LLM 1.2.1 · seed

-h --help --host --port --tp_size --pp_size --max_batch_size --kv_cache_free_gpu_memory_fraction --backend --max_seq_len --max_num_tokens --kv_cache_dtype --tool_parser --reasoning_parser --served_model_name --log_level --trust_remote_code --extra_llm_api_options

vllmvLLM 0.28.0 · seed

-h --help --version

Subcommands: serve, bench, chat, complete