Breakthrough for the Lithuanian Language in AI: 10 000-Hour Speech Corpus Developed
The development of the Large Lithuanian Language Speech Corpus LIEPA-3 has been completed in Lithuania. Researchers from Vilnius University, Vytautas Magnus University, and the Institute of the Lithuanian Language have collected and annotated 10,000 hours of Lithuanian speech recordings, equivalent to more than one year of continuous speech. It is the largest Lithuanian speech dataset ever created for artificial intelligence technologies.







