fix: cap SWA/chunked-local runtime admission to startup pool-sizing bound by AlpinDale · Pull Request #1659 · dphnAI/aphrodite-engine

AlpinDale · 2026-04-28T08:27:13Z

No description provided.

…ound Signed-off-by: AlpinDale <alpindale@gmail.com>

chatgpt-codex-connector

💡 Codex Review

Here are some automated review suggestions for this pull request.

Reviewed commit: 87d8d64015

ℹ️ About Codex in GitHub

Your team has set up Codex to review pull requests in this repo. Reviews are triggered when you

Open a pull request for review
Mark a draft as ready
Comment "@codex review".

If Codex has suggestions, it will comment; otherwise it will react with 👍.

Codex can also answer questions or update the PR. Try commenting "@codex address that feedback".

chatgpt-codex-connector · 2026-04-28T08:32:25Z

+            # `sum(reservations) <= pool` <=> `sum(peak_real_held) <= pool`.
+            # Drift between the two would re-introduce the deadlock from
+            # issue #39734 or, worse, mid-prefill OOM.
+            num_required_blocks = min(num_required_blocks, self._max_admission_blocks_per_request)


Compute capped admission against live blocks, not slot length

For running chunked-local requests, this cap is applied before the fast-path subtracts len(req_to_blocks), but req_to_blocks includes historical null placeholders from skipped chunks. Once the slot list grows past the cap, get_num_blocks_to_allocate can return 0 even when a new real KV block is still needed in the current chunk (no new skipped block freed in that step). In Scheduler._schedule_running this underestimation bypasses the free-block guard in KVCacheManager.allocate_slots, and allocate_new_blocks can then raise ValueError when BlockPool.get_new_blocks is called with insufficient free blocks.

Useful? React with 👍 / 👎.

fix: cap SWA/chunked-local runtime admission to startup pool-sizing b…

87d8d64

…ound Signed-off-by: AlpinDale <alpindale@gmail.com>

chatgpt-codex-connector Bot reviewed Apr 28, 2026

View reviewed changes

AlpinDale merged commit 42769da into main Apr 28, 2026
1 check failed

AlpinDale deleted the fix/kv-admission branch April 28, 2026 08:42

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Uh oh!

Uh oh!

fix: cap SWA/chunked-local runtime admission to startup pool-sizing bound#1659

fix: cap SWA/chunked-local runtime admission to startup pool-sizing bound#1659
AlpinDale merged 1 commit into
mainfrom
fix/kv-admission

AlpinDale commented Apr 28, 2026

Uh oh!

chatgpt-codex-connector Bot left a comment

Uh oh!

chatgpt-codex-connector Bot Apr 28, 2026

Uh oh!

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

1 participant

Uh oh!

Uh oh!

Conversation

AlpinDale commented Apr 28, 2026

Uh oh!

chatgpt-codex-connector Bot left a comment

Choose a reason for hiding this comment

💡 Codex Review

Uh oh!

chatgpt-codex-connector Bot Apr 28, 2026

Choose a reason for hiding this comment

Uh oh!

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

1 participant