We read every piece of feedback, and take your input very seriously.
To see all available qualifiers, see our documentation.
There was an error while loading. Please reload this page.
1 parent 4fb529c commit b0c8987Copy full SHA for b0c8987
1 file changed
perf-changelog.yaml
@@ -4950,4 +4950,4 @@
4950
- "Add GB300 Dynamo-vLLM AgentX aggregate TP4/TP8 low-latency sweeps at conc [1,4,8,16]."
4951
- "Add NVIDIA/srt-slurm#229-derived GB300 P/D topologies: 1P/1D DEP8/DEP8 at c512, 2P/1D DEP8/DEP8 at c1280, and 3P/1D DEP8/DEP16 at c1536."
4952
- "Use vllm/vllm-openai:nightly-dev-arm64-cu13.0.1-c188b96 for all new recipes."
4953
- pr-link: https://github.com/SemiAnalysisAI/InferenceX/pull/2260
+ pr-link: https://github.com/SemiAnalysisAI/InferenceX/pull/2269
0 commit comments