CVSS-X: A Multilingual Speech-to-Speech Translation Corpus for 28 Languages Paper • 2609.13413 • Published 23 days ago
BabyBabelLM: A Multilingual Benchmark of Developmentally Plausible Training Data Paper • 2510.10159 • Published Oct 11, 2025 • 3
Do Satellites See Commuters? A Critical Benchmark of Vision Foundation Models Paper • 2609.00661 • Published Sep 1
view post Post 823 How does billing work for Hugging Face Inference Providers? If I already have Hugging Face GPU credits, can I use them, or do I need to add a credit card and pay separately? See translation 3 replies · ➕ 3 3 + Reply
Building Bridges: A Dataset for Evaluating Gender-Fair Machine Translation into German Paper • 2406.06131 • Published Jun 10, 2024
Movie Facts and Fibs (MF$^2$): A Benchmark for Long Movie Understanding Paper • 2506.06275 • Published Jun 6, 2025
Ouvia: A User-centered Framework for Measuring Usability of Speech Translation in Real-World Communication Scenarios Paper • 2606.06177 • Published Jun 4
AMALIA Technical Report: A Fully Open Source Large Language Model for European Portuguese Paper • 2603.26511 • Published Mar 27
AnyMo: Geometry-Aware Setup-Agnostic Modeling of Human Motion in the Wild Paper • 2605.22715 • Published May 21 • 2
TrajPrism: A Multi-Task Benchmark for Language-Grounded Urban Trajectory Understanding Paper • 2605.10782 • Published May 11
TrajDLM: Topology-Aware Block Diffusion Language Model for Trajectory Generation Paper • 2605.10020 • Published May 11 • 2
Blending LLMs into Cascaded Speech Translation: KIT's Offline Speech Translation System for IWSLT 2024 Paper • 2406.16777 • Published Jun 24, 2024 • 1
Open ASR Leaderboard: Towards Reproducible and Transparent Multilingual and Long-Form Speech Recognition Evaluation Paper • 2510.06961 • Published Oct 8, 2025 • 14
BOE-XSUM: Extreme Summarization in Clear Language of Spanish Legal Decrees and Notifications Paper • 2509.24908 • Published Sep 29, 2025 • 3
KIT's Offline Speech Translation and Instruction Following Submission for IWSLT 2025 Paper • 2505.13036 • Published May 19, 2025
ViCocktail: Automated Multi-Modal Data Collection for Vietnamese Audio-Visual Speech Recognition Paper • 2506.04635 • Published Jun 5, 2025