vllm
[Perf] Optimize compute maxsim using batched version, 3.2% E2E throughput improvement
#36710
Merged

[Perf] Optimize compute maxsim using batched version, 3.2% E2E throughput improvement #36710

noooop merged 2 commits into main from wentao-optimize-compute-maxsim
yewentao256
yewentao256 optimize compute maxsim using batched version
14cfdb51
yewentao256 yewentao256 requested a review from noooop noooop 146 days ago
yewentao256 yewentao256 requested a review from WoosukKwon WoosukKwon 146 days ago
yewentao256 yewentao256 requested a review from njhill njhill 146 days ago
mergify mergify added frontend
mergify mergify added v1
yewentao256 yewentao256 added ready
gemini-code-assist
gemini-code-assist commented on 2026-03-10
DarkLight1337
DarkLight1337 commented on 2026-03-11
DarkLight1337
DarkLight1337 commented on 2026-03-11
yewentao256
DarkLight1337
DarkLight1337 approved these changes on 2026-03-11
yewentao256 Merge branch 'main' into wentao-optimize-compute-maxsim
0eb1b437
noooop noooop merged c34ba6b9 into main 145 days ago
noooop noooop deleted the wentao-optimize-compute-maxsim branch 145 days ago

Login to write a write a comment.

Login via GitHub

Assignees
No one assigned
Labels
Milestone