We read every piece of feedback, and take your input very seriously.
To see all available qualifiers, see our documentation.
There was an error while loading. Please reload this page.
1 parent 60f0789 commit 8bb369fCopy full SHA for 8bb369f
1 file changed
perf-changelog.yaml
@@ -4960,4 +4960,4 @@
4960
- "Add GB300 Dynamo-vLLM AgentX aggregate TP4/TP8 low-latency sweeps at conc [1,4,8,16]."
4961
- "Add NVIDIA/srt-slurm#229-derived GB300 P/D topologies: 1P/1D DEP8/DEP8 at c512, 2P/1D DEP8/DEP8 at c1280, and 3P/1D DEP8/DEP16 at c1536."
4962
- "Use vllm/vllm-openai:nightly-dev-arm64-cu13.0.1-c188b96 for all new recipes."
4963
- pr-link: https://github.com/SemiAnalysisAI/InferenceX/pull/2260
+ pr-link: https://github.com/SemiAnalysisAI/InferenceX/pull/2269
0 commit comments