Triple
T21783283
| Position | Surface form | Disambiguated ID | Type / Status |
|---|---|---|---|
| Subject | Oxford Dictionaries American English corpus |
E537770
|
entity |
| Predicate | instanceOf |
P0
|
FINISHED |
| Object | American English corpus |
C15286
|
CONCEPT FINISHED |
How this triple was built (1 step)
Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.
CD
Concept disambiguation
gpt-5-mini-2025-08-07
Target class: American English corpus Context triple: [Oxford Dictionaries American English corpus, instanceOf, American English corpus]
-
A.
linguistic corpus
chosen
A linguistic corpus is a large, structured collection of authentic texts or transcribed speech used for analyzing language patterns, usage, and structure.
-
B.
Spanish language corpus
A Spanish language corpus is a structured, large-scale collection of Spanish texts (written and/or spoken) compiled to support linguistic analysis, language research, and natural language processing applications.
-
C.
reference corpus
A reference corpus is a large, structured collection of texts compiled to represent a particular language, genre, or domain, used as an authoritative basis for linguistic analysis, comparison, and research.
-
D.
variety of American English
A variety of American English is a distinct, systematically patterned form of English used in the United States, characterized by particular phonological, lexical, grammatical, and pragmatic features associated with specific regions, social groups, or contexts.
-
E.
linguistic archive
A linguistic archive is a curated, long-term repository that collects, preserves, and provides access to language data and related documentation in various formats for research, revitalization, and educational purposes.
- F. None of above.
Provenance (1 batch)
The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.
| Step | Stage | Batch ID | Status | When |
|---|---|---|---|---|
| creating | Elicitation | batch_69e0c47198f881908cb0d237266c10e9 |
completed | April 16, 2026, 11:13 a.m. |
Created at: April 16, 2026, 6:52 p.m.