| title | DGDR Templates |
|---|---|
| subtitle | Ready-to-apply DynamoGraphDeploymentRequest manifests for profiling and generating DynamoGraphDeployments. |
A DynamoGraphDeploymentRequest (DGDR) describes a model, workload, hardware, and optional latency
targets. Dynamo profiles that intent and generates the DynamoGraphDeployment (DGD) that serves
traffic. Use these templates when you want Dynamo to choose a deployment configuration instead of
writing the complete DGD yourself.
Each manifest is embedded from
examples/deployments/dgdr/.
Open an example, use the copy button, and adjust its model, backend, hardware, and namespace before
applying it.
- Install the Dynamo Kubernetes platform and operator.
- Create
hf-token-secretin the target namespace when the model requires authentication. - For namespace-restricted operator installations, set
hardware.gpuSku,hardware.vramMb, andhardware.numGpusPerNodeexplicitly.
Release installations select the matching profiler image when spec.image is omitted. For a local
operator build without a known release version, set spec.image to a compatible dynamo-planner
image.
Apply a template with:
export NAMESPACE=dynamo-cloud
kubectl apply -f rapid.yaml -n ${NAMESPACE}<Code src="../../../../../examples/deployments/dgdr/rapid.yaml" title="rapid.yaml" language="yaml" maxLines={0} />
<Code src="../../../../../examples/deployments/dgdr/thorough.yaml" title="thorough.yaml" language="yaml" maxLines={0} />
<Code src="../../../../../examples/deployments/dgdr/planner.yaml" title="planner.yaml" language="yaml" maxLines={0} />
<Code src="../../../../../examples/deployments/dgdr/moe-sglang.yaml" title="moe-sglang.yaml" language="yaml" maxLines={0} />
<Code src="../../../../../examples/deployments/dgdr/review-before-deploy.yaml" title="review-before-deploy.yaml" language="yaml" maxLines={0} />
<Code src="../../../../../examples/deployments/dgdr/generated-dgd-override.yaml" title="generated-dgd-override.yaml" language="yaml" maxLines={0} />
<Code src="../../../../../examples/deployments/dgdr/model-cache.yaml" title="model-cache.yaml" language="yaml" maxLines={0} />
<Code src="../../../../../examples/deployments/dgdr/mocker.yaml" title="mocker.yaml" language="yaml" maxLines={0} />
<Code src="../../../../../examples/deployments/dgdr/profiling-artifacts.yaml" title="profiling-artifacts.yaml" language="yaml" maxLines={0} />
Watch the request until it reaches Deployed or Ready:
kubectl get dgdr <request-name> -n ${NAMESPACE} -wFor a request with autoApply: false, extract the generated DGD:
kubectl get dgdr <request-name> -n ${NAMESPACE} \
-o jsonpath='{.status.profilingResults.selectedConfig}' > generated-dgd.yamlSee Auto Deploy with DGDR for the task-oriented workflow and the DGDR Reference for field definitions, lifecycle phases, and validation behavior.