vllm
[Bugfix][Quantization] Fix FP8 + EP
#13784
Merged

[Bugfix][Quantization] Fix FP8 + EP #13784

youkaichao merged 3 commits into main from fix_layer_num_experts
tlrmchlsmth
tlrmchlsmth Fix FP8 + EP
a0d358d6
tlrmchlsmth tlrmchlsmth requested a review from mgoin mgoin 1 year ago
tlrmchlsmth tlrmchlsmth requested a review from robertgshaw2-redhat robertgshaw2-redhat 1 year ago
github-actions
tlrmchlsmth compressed_tensors_moe
6c430ed4
mgoin
mgoin approved these changes on 2025-02-24
mgoin mgoin added bug
mgoin mgoin added quantization
mgoin mgoin added ready
tlrmchlsmth be explicit about global/local
af4ffa8a
mgoin
mgoin approved these changes on 2025-02-24
tlrmchlsmth tlrmchlsmth enabled auto-merge (squash) 1 year ago
disabled auto-merge 1 year ago
Manually disabled by user
youkaichao youkaichao merged 1e15aaef into main 1 year ago
youkaichao youkaichao deleted the fix_layer_num_experts branch 1 year ago

Login to write a write a comment.

Login via GitHub

Assignees
No one assigned
Labels
Milestone