[AMD][MI355X] update model for gpt-oss by chunfangamd · Pull Request #1670 · SemiAnalysisAI/InferenceX

chunfangamd · 2026-06-05T03:16:29Z

Switch the MoE backend to CK to leverage the benefits from this PR: vllm-project/vllm#42098

Note

Low Risk
Benchmark-only model ID and changelog update; no serving code or sweep matrix changes in the diff.

Overview
Updates the GPT-OSS FP4 MI355X vLLM benchmark entry to load openai/gpt-oss-120b instead of amd/gpt-oss-120b-w-mxfp4-a-fp8, aligning this MI355X sweep with the other AMD GPT-OSS vLLM configs that already point at the OpenAI checkpoint. Image (vllm/vllm-openai-rocm:v0.22.0), runner, and fixed-seq-len search space are unchanged.

Adds a perf-changelog.yaml record for gptoss-fp4-mi355x-vllm documenting the model swap (PR #1670).

^{Reviewed by Cursor Bugbot for commit 770268c. Bugbot is set up for automated code reviews on this repo. Configure here.}

Record the GPT-OSS MI355X vLLM model update (amd/gpt-oss-120b-w-mxfp4-a-fp8 -> openai/gpt-oss-120b).

github-actions · 2026-06-05T03:16:38Z

Thanks for the contribution! For vLLM & SGLang, please ensure that your recipes is similar to the official vLLM recipes and/or the SGLang cookbook

If it is not, please create a PR first before we can merge your single node PR into the master branch. Let's ensure that the documentation is first class such that the entire ML community can benefit from your hard work! Thank you

PR authors are responsible for ensuring that after merging, all GitHub Action jobs fully pass. A lot of the time, failures are just flakes and simply re-running the failed jobs will fix it. If re-running failed jobs is attempted, PR authors are responsible for ensuring it passes. See GitHub's docs on re-running failed jobs: https://docs.github.com/en/actions/how-tos/manage-workflow-runs/re-run-workflows-and-jobs#re-running-failed-jobs-in-a-workflow

As a rule of thumb, generally, PR authors should request a review & get a PR approval from the respective companies' CODEOWNERS before requesting a review from core maintainers.

If additional help is needed, PR authors can reach out to core maintainers over Slack.

github-actions · 2026-06-05T03:16:38Z

Thanks for the contribution! For vLLM & SGLang, please ensure that your recipes is similar to the official vLLM recipes and/or the SGLang cookbook

If it is not, please create a PR first before we can merge your single node PR into the master branch. Let's ensure that the documentation is first class such that the entire ML community can benefit from your hard work! Thank you

PR authors are responsible for ensuring that after merging, all GitHub Action jobs fully pass. A lot of the time, failures are just flakes and simply re-running the failed jobs will fix it. If re-running failed jobs is attempted, PR authors are responsible for ensuring it passes. See GitHub's docs on re-running failed jobs: https://docs.github.com/en/actions/how-tos/manage-workflow-runs/re-run-workflows-and-jobs#re-running-failed-jobs-in-a-workflow

As a rule of thumb, generally, PR authors should request a review & get a PR approval from the respective companies' CODEOWNERS before requesting a review from core maintainers.

If additional help is needed, PR authors can reach out to core maintainers over Slack.

github-actions · 2026-06-05T03:16:38Z

Thanks for the contribution! For vLLM & SGLang, please ensure that your recipes is similar to the official vLLM recipes and/or the SGLang cookbook

If it is not, please create a PR first before we can merge your single node PR into the master branch. Let's ensure that the documentation is first class such that the entire ML community can benefit from your hard work! Thank you

PR authors are responsible for ensuring that after merging, all GitHub Action jobs fully pass. A lot of the time, failures are just flakes and simply re-running the failed jobs will fix it. If re-running failed jobs is attempted, PR authors are responsible for ensuring it passes. See GitHub's docs on re-running failed jobs: https://docs.github.com/en/actions/how-tos/manage-workflow-runs/re-run-workflows-and-jobs#re-running-failed-jobs-in-a-workflow

As a rule of thumb, generally, PR authors should request a review & get a PR approval from the respective companies' CODEOWNERS before requesting a review from core maintainers.

If additional help is needed, PR authors can reach out to core maintainers over Slack.

billishyahao

LGTM

github-actions · 2026-06-05T04:56:27Z

see unofficial run visualizer at https://inferencex.semianalysis.com/inference?unofficialRun=26993379366
see unofficial run visualizer at https://inferencex.semianalysis.com/evaluation?unofficialRun=26993379366

github-actions · 2026-06-05T06:37:45Z

see unofficial run visualizer at https://inferencex.semianalysis.com/inference?unofficialRun=26998598877
see unofficial run visualizer at https://inferencex.semianalysis.com/evaluation?unofficialRun=26998598877

ukannika and others added 2 commits June 5, 2026 03:11

Update amd-master.yaml

0ab344e

Add perf-changelog entry placeholder

3fa6c61

Record the GPT-OSS MI355X vLLM model update (amd/gpt-oss-120b-w-mxfp4-a-fp8 -> openai/gpt-oss-120b).

chunfangamd requested a review from a team June 5, 2026 03:16

github-project-automation Bot added this to InferenceMAX Board Jun 5, 2026

chunfangamd requested review from 1am9trash, billishyahao, seungrokj and yctseng0211 as code owners June 5, 2026 03:16

Update the changelog

48f5b3c

claude Bot reviewed Jun 5, 2026

View reviewed changes

Comment thread perf-changelog.yaml Outdated

chunfangamd added the full-sweep-enabled label Jun 5, 2026

billishyahao approved these changes Jun 5, 2026

View reviewed changes

Merge branch 'main' into chun-uma/gptoss-regression-fix-1

770268c

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

[AMD][MI355X] update model for gpt-oss#1670

[AMD][MI355X] update model for gpt-oss#1670
chunfangamd wants to merge 4 commits into
mainfrom
chun-uma/gptoss-regression-fix-1

chunfangamd commented Jun 5, 2026 •

edited by cursor Bot

Loading

Uh oh!

github-actions Bot commented Jun 5, 2026

Uh oh!

github-actions Bot commented Jun 5, 2026

Uh oh!

github-actions Bot commented Jun 5, 2026

Uh oh!

Uh oh!

billishyahao left a comment

Uh oh!

github-actions Bot commented Jun 5, 2026

Uh oh!

github-actions Bot commented Jun 5, 2026

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

3 participants

Conversation

chunfangamd commented Jun 5, 2026 • edited by cursor Bot Loading Uh oh! There was an error while loading. Please reload this page.

Uh oh!

Uh oh!

github-actions Bot commented Jun 5, 2026

Uh oh!

github-actions Bot commented Jun 5, 2026

Uh oh!

github-actions Bot commented Jun 5, 2026

Uh oh!

Uh oh!

billishyahao left a comment

Choose a reason for hiding this comment

Uh oh!

github-actions Bot commented Jun 5, 2026

Uh oh!

github-actions Bot commented Jun 5, 2026

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

3 participants

chunfangamd commented Jun 5, 2026 •

edited by cursor Bot

Loading