Specdec Bench: vLLM reqid, SGL path, conc > 1 metric fix #541

IzzyPutterman · 2025-11-12T06:21:00Z

What does this PR do?

SGLang Fix for actually passing the draft model path to the engine

vLLM Fix for multiturn to not overlap request_id strings

Acceptance Rate Fix for potential race condition on multiturn datasets in writing back AR

Overview: ?

Usage

# Add a code snippet demonstrating how to use this

Testing

Before your PR is "Ready for review"

Make sure you read and follow Contributor guidelines and your commits are signed.
Is this change backward compatible?: Yes/No
Did you write any new necessary tests?: Yes/No
Did you add or update any necessary documentation?: Yes/No
Did you update Changelog?: Yes/No

Additional Information

Signed-off-by: Izzy Putterman <[email protected]>

codecov · 2025-11-12T06:33:35Z

Codecov Report

✅ All modified and coverable lines are covered by tests.
✅ Project coverage is 74.37%. Comparing base (7d0f7a9) to head (1f7e8cb).
⚠️ Report is 38 commits behind head on main.

Additional details and impacted files

@@           Coverage Diff           @@
##             main     #541   +/-   ##
=======================================
  Coverage   74.37%   74.37%           
=======================================
  Files         182      182           
  Lines       18219    18219           
=======================================
  Hits        13550    13550           
  Misses       4669     4669

☔ View full report in Codecov by Sentry.
📢 Have feedback on the report? Share it here.

🚀 New features to boost your workflow:

❄️ Test Analytics: Detect flaky tests, report on failures, and find test suite problems.

yeyu-nvidia · 2025-11-13T22:44:16Z

Can you add some description of this PR?

IzzyPutterman · 2025-11-14T06:26:04Z

Can you add some description of this PR?

Updated the description

h-guo18

LGTM

## What does this PR do? **SGLang** Fix for actually passing the draft model path to the engine **vLLM** Fix for multiturn to not overlap request_id strings **Acceptance Rate** Fix for potential race condition on multiturn datasets in writing back AR **Overview:** ? ## Usage  ```python # Add a code snippet demonstrating how to use this ``` ## Testing  ## Before your PR is "*Ready for review*"  - **Make sure you read and follow [Contributor guidelines](https://github.com/NVIDIA/TensorRT-Model-Optimizer/blob/main/CONTRIBUTING.md)** and your commits are signed. - **Is this change backward compatible?**: Yes/No  - **Did you write any new necessary tests?**: Yes/No - **Did you add or update any necessary documentation?**: Yes/No - **Did you update [Changelog](https://github.com/NVIDIA/TensorRT-Model-Optimizer/blob/main/CHANGELOG.rst)?**: Yes/No  ## Additional Information  Signed-off-by: Izzy Putterman <[email protected]>

Specdec Bench: vLLM reqid, SGL path, conc > 1 metric fix

1f7e8cb

Signed-off-by: Izzy Putterman <[email protected]>

IzzyPutterman requested a review from a team as a code owner November 12, 2025 06:21

IzzyPutterman requested a review from ChenhanYu November 12, 2025 06:21

IzzyPutterman requested review from kevalmorabia97 and yeyu-nvidia November 13, 2025 22:32

IzzyPutterman requested a review from h-guo18 November 19, 2025 22:53

h-guo18 approved these changes Nov 21, 2025

View reviewed changes

yeyu-nvidia approved these changes Nov 26, 2025

View reviewed changes

IzzyPutterman merged commit 80c5491 into main Nov 26, 2025
26 checks passed

IzzyPutterman deleted the iputterman/specdec-bench-11-11 branch November 26, 2025 05:13

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Uh oh!

Specdec Bench: vLLM reqid, SGL path, conc > 1 metric fix #541

Specdec Bench: vLLM reqid, SGL path, conc > 1 metric fix #541

IzzyPutterman commented Nov 12, 2025 •

edited

Loading

Uh oh!

codecov bot commented Nov 12, 2025 •

edited

Loading

Uh oh!

yeyu-nvidia commented Nov 13, 2025

Uh oh!

IzzyPutterman commented Nov 14, 2025

Uh oh!

h-guo18 left a comment

Uh oh!

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

4 participants

Specdec Bench: vLLM reqid, SGL path, conc > 1 metric fix #541

Specdec Bench: vLLM reqid, SGL path, conc > 1 metric fix #541

Conversation

IzzyPutterman commented Nov 12, 2025 • edited Loading Uh oh! There was an error while loading. Please reload this page.

Uh oh!

What does this PR do?

Usage

Testing

Before your PR is "Ready for review"

Additional Information

Uh oh!

codecov bot commented Nov 12, 2025 • edited Loading Uh oh! There was an error while loading. Please reload this page.

Uh oh!

Codecov Report

Uh oh!

yeyu-nvidia commented Nov 13, 2025

Uh oh!

IzzyPutterman commented Nov 14, 2025

Uh oh!

h-guo18 left a comment

Choose a reason for hiding this comment

Uh oh!

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

4 participants

IzzyPutterman commented Nov 12, 2025 •

edited

Loading

codecov bot commented Nov 12, 2025 •

edited

Loading