It intelligently segments text into meaningful semantic chunks. Could be useful for RAG systems as text-chunking module.
-
mirth/chonky_distilbert_base_uncased_1
Token Classification • 66.4M • Updated • 43.4k • • 15 -
mirth/chonky_mmbert_small_multilingual_1
Token Classification • 0.1B • Updated • 174 • 23 -
mirth/chonky_modernbert_base_1
Token Classification • 0.1B • Updated • 8.29k • • 6 -
mirth/chonky_modernbert_large_1
Token Classification • 0.4B • Updated • 2.08k • • 2