vllm
[Bugfix][Quantization] Fix FP8 + EP
#13784
Merged
Go
Login via GitHub
Home
Pricing
FAQ
Install
Login
via GitHub
Overview
Commits
3
Changes
View On
GitHub
[Bugfix][Quantization] Fix FP8 + EP
#13784
youkaichao
merged 3 commits into
main
from
fix_layer_num_experts
Fix FP8 + EP
a0d358d6
tlrmchlsmth
requested a review
from
mgoin
1 year ago
tlrmchlsmth
requested a review
from
robertgshaw2-redhat
1 year ago
compressed_tensors_moe
6c430ed4
mgoin
approved these changes on 2025-02-24
mgoin
added
bug
mgoin
added
quantization
mgoin
added
ready
be explicit about global/local
af4ffa8a
mgoin
approved these changes on 2025-02-24
tlrmchlsmth
enabled auto-merge (squash)
1 year ago
disabled auto-merge
1 year ago
Manually disabled by user
youkaichao
merged
1e15aaef
into main
1 year ago
youkaichao
deleted the fix_layer_num_experts branch
1 year ago
Login to write a write a comment.
Login via GitHub
Reviewers
mgoin
robertgshaw2-redhat
Assignees
No one assigned
Labels
bug
ready
Milestone
No milestone
Login to write a write a comment.
Login via GitHub