llama.cpp
model: GraniteSpeech5ForCTC (Turbo CTC)
#29446
Open

model: GraniteSpeech5ForCTC (Turbo CTC) #29446

gabe-l-hart wants to merge 36 commits into ggml-org:master from gabe-l-hart:GraniteSpeechCTC
gabe-l-hart
gabe-l-hart gabe-l-hart requested a review 4 days ago
gabe-l-hart gabe-l-hart requested a review 4 days ago
gabe-l-hart gabe-l-hart requested a review from CISC CISC 4 days ago
gabe-l-hart gabe-l-hart requested a review from ggerganov ggerganov 4 days ago
gabe-l-hart gabe-l-hart requested a review 4 days ago
github-actions github-actions added model
github-actions github-actions added server
github-actions github-actions added mtmd
github-actions github-actions added conversion
pwilkin
ggml-gh-bot
gabe-l-hart gabe-l-hart force pushed from 9761d878 to ac132eb8 1 day ago
gabe-l-hart feat: Add constants, tensor_mapping, and gguf_writer helpers for Gran…
35bf7444
gabe-l-hart feat: Add option to skip n_embd_text check for MMPROJ models without …
2cbd789f
gabe-l-hart feat: Add chkhsh exception for granite-speech-5
e23a7d98
gabe-l-hart feat: Add encoder and mmproj FE conversion for granite-speech-5
ebc12ed0
gabe-l-hart feat(api): Add public llama_n_outputs
5210decd
gabe-l-hart feat: Automatically raise ubatch size for encoder-only models to fit …
1e4f98f0
gabe-l-hart feat: Compute encoder output size explicitly to allow for time downsa…
7988022c
gabe-l-hart feat: Make subsample_layers a per-layer factor rather than an indicator
d13fe6d8
gabe-l-hart feat: Add llama_hparams handling for per-layer subsample_factor
0451f2d3
gabe-l-hart feat: Explicitly map pretokenizer to default
20326b6e
gabe-l-hart feat: Add CTC encoder hparams block
6ef1a902
gabe-l-hart feat: Add c++ constants for arch, GGUF kvs, and tensors
ffc5b134
gabe-l-hart feat: Add llama_model tensor pointers
4cc1b67b
gabe-l-hart feat: Add granite-speech-5 model impl
ead6674e
gabe-l-hart fix: Typo in granite_speech-5
3b6117a0
gabe-l-hart fix: Remove unused python-side hparam constants
43ae45d9
gabe-l-hart fix: Use native output_dim instead of counted n_vocab for output tensors
afcb8291
gabe-l-hart feat: Add interface in common for running ctc forward pass and decoding
56fdcd3d
gabe-l-hart refactor(convert): granite-speech-5 -> granite_speech_5
5748fc83
gabe-l-hart refactor(c++): granite-speech-5 -> granite_speech_5
716df2df
gabe-l-hart feat: Add transformer-less frontend clip model for granite-speech-5
983a44dc
gabe-l-hart feat: Add clip model and hparam plumbing for granite-speech-5 frontend
286fb5bb
gabe-l-hart feat: Add granite-speech-5 preprocessor
3a537720
gabe-l-hart feat(server): Add server task and response type for transcribe
d01cbc61
gabe-l-hart feat(server): Wire through encoder-only transcriptions handling
511da47c
gabe-l-hart fix: Use the actual unpadded number of tokens for output
f37a0409
gabe-l-hart fix(convert): Don't insert an extra blank in the vocabulary
6a92302a
gabe-l-hart Revert "fix: Use native output_dim instead of counted n_vocab for out…
1b843565
gabe-l-hart fix: Always use llama_n_tokens in common_ctc_greedy_decode
1e497b50
gabe-l-hart Revert "fix: Use the actual unpadded number of tokens for output"
6131ff97
gabe-l-hart feat: Add CTC encoder params to llama-model-saver
3e2cbe39
gabe-l-hart test: Add CTC params for G5 speech in test-llama-archs
023e0345
gabe-l-hart fix: Serialize new tensors for CTC out_mid in llama-model-saver
6aeca0ab
gabe-l-hart test: Skip g-speech-5 for tests that run tokens through
f0259423
gabe-l-hart fix: Allow non-sampling models through logits memory-free check after…
08d6392c
gabe-l-hart gabe-l-hart force pushed from ac132eb8 to 08d6392c 4 hours ago
github-actions github-actions added testing
gabe-l-hart fix: Fix bad HF link in base.py
61563269

Login to write a write a comment.

Login via GitHub

Reviewers
Assignees
No one assigned
Labels
Milestone