# Speech Recognition & AI Research | Digisensus

> Digisensus research on speech recognition, large language models and AI: pre-registered studies, released with their code, datasets and models.

Research

# We share our research in ASR, LLMs and other AI technologies

Digisensus Research publishes studies on speech recognition, large language models and the other technologies behind our products.

Studies

## Published studies

Study ltd26 · 5 October 2026

### [Lithuanian speech recognition: effects of dialect training and transcript spelling](/lithuanian-dialect-speech-recognition/)

Does adding dialect recordings improve Lithuanian speech recognition, and should their transcripts use dialect or standard spelling? A controlled fine-tuning study on the LIEPA-3 corpus with 18 training runs, a pre-registered protocol and an independent customer-service test.

-   Dialect WER falls from 41.62% to 31.30% with 82 h of dialect speech; other speech of the same length reaches only 40.04%
-   The first 25 hours give about two thirds of the gain
-   Customer-service WER falls from 44.89% to 39.06% with standard-spelling dialect transcripts

Released: Code (MIT), three datasets and six fine-tuned models (CC BY 4.0), and the full report as PDF.

[Read the full report](/lithuanian-dialect-speech-recognition/) [PDF](/assets/research/lithuanian-dialect-asr-study-2026-10-05.pdf)

## Build on it

The datasets, models and code from every study are on the Open source page, with their licenses.

[Open source](/open-source/) [GitHub](https://github.com/Digisensus/) [🤗Hugging Face](https://huggingface.co/Digisensus)

---

- Canonical: https://digisensus.com/research/
- Lithuanian version: https://digisensus.com/lt/tyrimai/
- Book a demo: https://calendar.app.google/MM4rtjEa8ctThvoZ8
- Contact: saulius@digisensus.com · +370 620 69969
