Triple

T21783283
Position Surface form Disambiguated ID Type / Status
Subject Oxford Dictionaries American English corpus E537770 entity
Predicate instanceOf P0 FINISHED
Object American English corpus C15286 CONCEPT FINISHED

How this triple was built (1 step)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

CD Concept disambiguation gpt-5-mini-2025-08-07
Target class: American English corpus
Context triple: [Oxford Dictionaries American English corpus, instanceOf, American English corpus]
  • A. linguistic corpus chosen
    A linguistic corpus is a large, structured collection of authentic texts or transcribed speech used for analyzing language patterns, usage, and structure.
  • B. Spanish language corpus
    A Spanish language corpus is a structured, large-scale collection of Spanish texts (written and/or spoken) compiled to support linguistic analysis, language research, and natural language processing applications.
  • C. reference corpus
    A reference corpus is a large, structured collection of texts compiled to represent a particular language, genre, or domain, used as an authoritative basis for linguistic analysis, comparison, and research.
  • D. variety of American English
    A variety of American English is a distinct, systematically patterned form of English used in the United States, characterized by particular phonological, lexical, grammatical, and pragmatic features associated with specific regions, social groups, or contexts.
  • E. linguistic archive
    A linguistic archive is a curated, long-term repository that collects, preserves, and provides access to language data and related documentation in various formats for research, revitalization, and educational purposes.
  • F. None of above.

Provenance (1 batch)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69e0c47198f881908cb0d237266c10e9 completed April 16, 2026, 11:13 a.m.
Created at: April 16, 2026, 6:52 p.m.