view article Article The state-of-the-art in open-source AI for Swiss legal tasks joelniklaus • Jul 14 • 9
ibm-granite/granite-embedding-311m-multilingual-r2 Feature Extraction • 0.3B • Updated May 18 • 76.5k • • 125
CohereLabs/cohere-transcribe-03-2026 Automatic Speech Recognition • 2B • Updated Jun 10 • 499k • • 1.09k
mistralai/Voxtral-Mini-4B-Realtime-2602 Automatic Speech Recognition • 4B • Updated Mar 11 • 2.33M • 952
llm-semantic-router/mmbert-embed-32k-2d-matryoshka Sentence Similarity • 0.3B • Updated Mar 6 • 15.7k • • 18
NanoBEIR 🍺 Collection A collection of smaller versions of BEIR datasets with 50 queries and up to 10K documents each. • 13 items • Updated Sep 11, 2024 • 27
mdbr-leaf-embedding Collection A collection of compact, high performance text-embedding models trained using our proposed LEAF framework, see https://arxiv.org/abs/2509.12539 • 4 items • Updated Sep 29, 2025 • 7