LibriSpeech
E1312480
UNEXPLORED
LibriSpeech is a large-scale corpus of read English audiobooks widely used as a standard benchmark dataset for training and evaluating automatic speech recognition systems.
All labels observed (1)
| Label | Occurrences |
|---|---|
| LibriSpeech canonical | 1 |
How this entity was disambiguated
This entity first appeared as the object of triple T18205230 — resolving that mention is where its identity was fixed. The disambiguator weighed these candidate entities and picked the highlighted one (or “None”, minting a new entity). This is how homonymy is resolved: the same surface form can point to different entities.
NED1
Entity disambiguation (via context triple)
gpt-5-mini-2025-08-07
Target entity: LibriSpeech Context triple: [HuBERT, evaluationBenchmark, LibriSpeech]
-
A.
Common Voice dataset
The Common Voice dataset is a large, open-source multilingual speech corpus created by Mozilla to support and democratize voice recognition research and technology.
-
B.
CORDE corpus
The CORDE corpus is a large historical Spanish language corpus compiled by the Royal Spanish Academy, used for studying the evolution and usage of Spanish over time.
-
C.
The Language Archive
The Language Archive is a poignant stage play by Julia Cho that explores love, communication, and the limits of language through the story of a linguist struggling to understand the people closest to him.
-
D.
Reading Aloud
"Reading Aloud" is a 19th-century painting by Albert Joseph Moore that exemplifies his refined Aesthetic Movement style, depicting harmoniously arranged female figures in a serene, decorative interior.
-
E.
Versoix
Versoix is a Swiss municipality on the shores of Lake Geneva, known as a residential suburb of Geneva with lakeside promenades and a mix of urban and natural landscapes.
- F. None of above. chosen
- G. Unsure - the case is ambiguous/there is not enough information to decide.
NED2
Entity disambiguation (via description)
gpt-5-mini-2025-08-07
Target entity: LibriSpeech Target entity description: LibriSpeech is a large-scale corpus of read English audiobooks widely used as a standard benchmark dataset for training and evaluating automatic speech recognition systems.
-
A.
Common Voice dataset
The Common Voice dataset is a large, open-source multilingual speech corpus created by Mozilla to support and democratize voice recognition research and technology.
-
B.
CORDE corpus
The CORDE corpus is a large historical Spanish language corpus compiled by the Royal Spanish Academy, used for studying the evolution and usage of Spanish over time.
-
C.
The Language Archive
The Language Archive is a poignant stage play by Julia Cho that explores love, communication, and the limits of language through the story of a linguist struggling to understand the people closest to him.
-
D.
Reading Aloud
"Reading Aloud" is a 19th-century painting by Albert Joseph Moore that exemplifies his refined Aesthetic Movement style, depicting harmoniously arranged female figures in a serene, decorative interior.
-
E.
Versoix
Versoix is a Swiss municipality on the shores of Lake Geneva, known as a residential suburb of Geneva with lakeside promenades and a mix of urban and natural landscapes.
- F. None of above. chosen
Referenced by (1)
Full triples — surface form annotated when it differs from this entity's canonical label.