Radio and Television Speech Corpus

Research on public spoken language. This corpus is an empirical research instrument for the project “Lithuanian Language: Ideals, Ideologies and Identity Shifts,” intended to analyze how public language has changed over the last five decades. It consists of transcripts of 63 hours of radio and television programmes or excerpts, totaling approximately 351,270 words, annotated with linguistic features relevant to the research, programme genres, and speaker types.

Details

Name c000920
Identifier C000920
Homepage https://dataportal.gov.lt/lt/datasets/c000920
Publisher 111955023
Creator 188600177
Licence cc_by_4_0
Abbreviation Radio and Television Speech Corpus

Datasets

This catalog has no datasets yet.