Skip to content

Commit 37fd769

Browse files
authored
Prevent Incorrect Load-Aware Routing for AMD SGLang Backends (llm-d#2025)
Signed-off-by: weizhoublue <weizhou.lan@daocloud.io>
1 parent 3af10f0 commit 37fd769

1 file changed

Lines changed: 1 addition & 0 deletions

File tree

guides/optimized-baseline/modelserver/amd/sglang/kustomization.yaml

Lines changed: 1 addition & 0 deletions
Original file line numberDiff line numberDiff line change
@@ -14,6 +14,7 @@ labels:
1414
llm-d.ai/model: Qwen3-32B
1515
llm-d.ai/accelerator-variant: gpu
1616
llm-d.ai/accelerator-vendor: amd
17+
llm-d.ai/engine-type: sglang
1718
includeSelectors: true
1819
includeTemplates: true
1920
patches:

0 commit comments

Comments
 (0)