Welcome to the VLO!
Use the search bar below to start searching through hundreds of thousands of language resources, or continue to browse everything and use facets to narrow down to your area of interest or discover new resources.
See all records Learn more Take a quick tourUse the categories below to limit the search results to those matching the selected value(s).
Show more facetsThese levels provide an indication of the degree to which resources and tools are publicly accessible. Please check the specific conditions on any resource or tool that you end up using.
This corpus contains the audio recordings of all actors who use the SmartKom system; it covers the audio recordings (no …
This corpus contains the audio recordings of all actors who use the SmartKom system; it covers the audio recordings (no video) and annotations of all three original SmartKom corpora Public, Mobile and Home. Naive users were asked to test a 'prototype' for a market study not knowing that the system was in fact controlle…
De Stichting Nederlands Dagboekarchief verzamelt en beheert (ongepubliceerde) dagboeken, reisdagboeken, memoires, brieve…
De Stichting Nederlands Dagboekarchief verzamelt en beheert (ongepubliceerde) dagboeken, reisdagboeken, memoires, brieven en poëzie-albums uit het hele Nederlandse taalgebied en maakt deze toegankelijk voor wetenschap en onderwijs en voor particulier onderzoek. De collectie is in eigendom en beheer van de Stichting Ned…
The songs in this collection were recorded and annotated as part of the project 'Metre and Melody in Dinka Speech and So…
The songs in this collection were recorded and annotated as part of the project 'Metre and Melody in Dinka Speech and Song', a project carried out by researchers from the University of Edinburgh and the School of Oriental and African Studies in London, and funded by the UK Arts and Humanities Research Council as part o…
The MOCHA database was compiled as part of the Engineering and Physical Sciences Research Council grant number:GR/L78680…
The MOCHA database was compiled as part of the Engineering and Physical Sciences Research Council grant number:GR/L78680 : "Speech recognition using articulatory data." It features a set of 460 short sentences designed to include the main connected speech processes in English (e.g. assimilations, weak forms ...). All r…
This project will deliver detailed documentation of two undescribed Papuan languages from an almost completely unknown f…
This project will deliver detailed documentation of two undescribed Papuan languages from an almost completely unknown family in Southern New Guinea, plus more basic materials on two others. The project embeds a young German PhD student (Döhler) in a team including a seasoned field linguist (Evans) and a post-doctoral …
The CI_2 corpora contain synchronous speech recordings of 48 cochlear implant users (CI) and 48 speakers without hearing…
The CI_2 corpora contain synchronous speech recordings of 48 cochlear implant users (CI) and 48 speakers without hearing impairment (control group, KG). The data were analyzed in Veronika Neumeyer's dissertation "Akustische Analysen der Sprachproduktion von CI-Trägern" (2015). CI_2_VOT contains recordings used for the …
This corpus contains recordings of 162 speakers while being sober and intoxicated. Beginning with version 3, this corpus…
This corpus contains recordings of 162 speakers while being sober and intoxicated. Beginning with version 3, this corpus edition also contains an emuR compatible database version of the corpus (with a minor bugfix in the database in version 3.1).; Speech data collection of alcoholized speakers of German, age 21-75.
Verbmobil 2 contains the speech of 401 speakers participating in 810 recordings. The emotional tagged recordings are not…
Verbmobil 2 contains the speech of 401 speakers participating in 810 recordings. The emotional tagged recordings are not part of this edition but are collected inthe corpus 'BAS VMEmo'. The total VM2 corpus amounts to 17.6GB of data containing 58961 conversational turns distributed on 39 CD-R. VM2 contains dialogs in G…
The CI_2 corpora contain German speech recordings of 48 cochlear implant users (CI) and 48 speakers without hearing impa…
The CI_2 corpora contain German speech recordings of 48 cochlear implant users (CI) and 48 speakers without hearing impairment (control group, KG). The data were analyzed in Veronika Neumeyer's dissertation "Akustische Analysen der Sprachproduktion von CI-Trägern" (2015). CI_2_Vowels contains recordings used for the an…