[Misc] Consolidate Audio tests into multimodal common generation tests #18214
refactor vlm test runner to support audio input
12ba7b41
add mixed modality input
c0486f62
fix model loading
a1f9ef8c
Merge remote-tracking branch 'upstream/main' into qwen25-omni-test
6e5aaf25
fix tests
f5b23c4f
fix aspect ratio tests
f1a06a08
add helper class
c43e3ce1
rename
ee1d37bd
make mypy happy
a04edd06
expose video data
4d996c1a
remove runner mm_key
00ca64dc
add video to mix modality test
9a55fa5a
Merge remote-tracking branch 'origin/qwen25-omni-test' into audio-test
cf27d8ec
integrate audio tests
36903f65
add audio test
8a5981a7
clean up ultravox test
63d23b90
align vllm and hf outputs
c42edb4e
fix audio prompt building
b1b52511
code format
b01d02b3
fix types
c042ccf7
correct annotations
bccb7b4b
Merge branch 'vllm-project:main' into audio-test
cfd48e86
Isotr0py
deleted the audio-test branch 1 year ago
Assignees
No one assigned
Labels
ready
multi-modality
Login to write a write a comment.
Login via GitHub