fix cuda ut fail for transformers v5.0 #1357
fix cuda ut fail for transformers v5.0
23caeec1
set default eval batch size to avoid None case
f48306fb
fix cpu llama4 issue
ba01ce87
fix typo
ff9cef42
use monkey_patch to resolve no_init_weights issue of auto_gptq
50779390
xin3he
marked this pull request as draft 204 days ago
add skip for transformers issue
0b297c4f
add skip for moe and add mock_cuda_capability for fp8 model input
72e1cecb
update expected outputs for transformers 5.0.0 and refine model name …
1d2b0d12
xin3he
marked this pull request as ready for review 204 days ago
Merge branch 'main' into hengguo/fix_v5_cuda_ut
f1d94e41
xin3he
requested changes
on 2026-01-28
remove default batch size
076948bb
keep moe failure waiting for #1345
07ced708
fix
7c753c98
fix gptqmodel
f2b4af24
merge
ede38216
codescan
4c81a3df
update
5e6fada3
fix spwan error of vllm
dfd4a111
[pre-commit.ci] auto fixes from pre-commit.com hooks
5a35fd16
Merge branch 'main' into hengguo/fix_v5_cuda_ut
53bf7f3f
xin3he
approved these changes
on 2026-01-29
update lm_eval requirement
112d7baa
xin3he
merged
eb2ccdc3
into main 202 days ago
xin3he
deleted the hengguo/fix_v5_cuda_ut branch 202 days ago
Assignees
No one assigned
Login to write a write a comment.
Login via GitHub