perf: avoid spaCy statistical pipeline for word tokenization #4408
perf: tokenize words without running spaCy pipeline
3a88532e
perf: add text partition benchmark
ab5e888b
test: cover tokenizer-only behavior
efe77c7f
docs: add tokenizer performance changelog
722584c0
build(version): bump to 0.25.2-dev0
956fadbb
cragwolfe
approved these changes
on 2026-07-22
fix benchmark aggregate median
7fb89c06
fix parallel metrics evaluation test
8e06546a
Assignees
No one assigned
Login to write a write a comment.
Login via GitHub