News
Updates from the Slovak Language AI Hub and the Slovak NLP ecosystem.
Follow what's happening in Slovak NLP - new resources, research results, and community activities.
What it takes to build a speech-to-text benchmark for Slovak
If you want to know how well speech-to-text models work for English, you have your answer in a minute. For Slovak, until recently you had no answer at all. Multilingual evaluations either skip Slovak entirely or include it as a fraction of a percent of the test data, far too little to say anything reliable.
The First Slovak ASR Benchmark Is Now Available on SLAIH
We are releasing the first dedicated automatic speech recognition benchmark for Slovak on SLAIH.
SkMTEB: The First Benchmark for Slovak Embeddings
SkMTEB — the first comprehensive benchmark for text embeddings in Slovak — has been accepted at ACL 2026, one of the most prestigious international conferences in natural language processing. The benchmark covers 31 datasets across 7 task types and establishes a standard comparable to MTEB for English. The work is led by a team from the Kempelen Institute of Intelligent Technologies (KInIT), the Technical University of Košice (TUKE), and Comenius University Bratislava (UK), together with other partners from the Slovak AI ecosystem.
Launching Slovak Language AI Hub
We are launching Slovak Language AI Hub - the central place for datasets, models, and benchmarks for the Slovak language. Our goal is to improve the availability of resources and support the development of NLP in Slovakia.
What's New in the SLAIH Asset Catalog? Five Resources Worth Exploring
One of the main goals of SLAIH is to make it easier to discover high-quality resources for Slovak language AI. In this edition, we'd like to highlight a few recent additions that we believe deserve your attention. Whether you're building an LLM application, training an ASR system, or experimenting with multilingual models, these resources are well worth exploring.