🚀 The feature, motivation and pitch
hi @hongxiayang
+viz @powderluv @chunfangamd @andyluo7
PD disagg is already the current state of optimization that ppl run in prod. currently vLLM router only has CUDA GPU tests and doesn't have AMD GPU test. can u look into fixing this? AMD supports RIXL in vLLM router but there is no tests for that. it should also have tests in vLLM router for MoRI kvcache transfer once it implements that
https://github.com/vllm-project/router/blob/4df9bbc6562bd0ccaddb991efad581bfd14843a5/.buildkite/pipeline.yml#L193-L276
Alternatives
ROCm doesn't improve their user experience & code quality
Additional context
No response
Before submitting a new issue...
🚀 The feature, motivation and pitch
hi @hongxiayang
+viz @powderluv @chunfangamd @andyluo7
PD disagg is already the current state of optimization that ppl run in prod. currently vLLM router only has CUDA GPU tests and doesn't have AMD GPU test. can u look into fixing this? AMD supports RIXL in vLLM router but there is no tests for that. it should also have tests in vLLM router for MoRI kvcache transfer once it implements that
https://github.com/vllm-project/router/blob/4df9bbc6562bd0ccaddb991efad581bfd14843a5/.buildkite/pipeline.yml#L193-L276
Alternatives
ROCm doesn't improve their user experience & code quality
Additional context
No response
Before submitting a new issue...