Triple

T22672235
Position Surface form Disambiguated ID Type / Status
Subject Sipakapense language E560247 entity
Predicate closelyRelatedTo P37 FINISHED
Object Tektiteko language NE NERFINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Tektiteko language | Statement: [Sipakapense language, closelyRelatedTo, Tektiteko language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Tektiteko language
Context triple: [Sipakapense language, closelyRelatedTo, Tektiteko language]
  • A. Tektiteko language chosen
    The Tektiteko language is a Mayan language spoken by the Tektiteko people of Guatemala and Mexico, closely related to Mam and part of the broader Mamean branch.
  • B. Teke-Kega language
    The Teke-Kega language is a Bantu language spoken by the Teke people of Central Africa, primarily in the Republic of the Congo and surrounding regions.
  • C. Jakaltek language
    The Jakaltek language is a Mayan language spoken primarily by the Jakaltek (Popti’) people of northwestern Guatemala and parts of southern Mexico.
  • D. Terik language
    The Terik language is a Southern Nilotic language spoken by the Terik people of western Kenya, closely associated with neighboring Kalenjin groups such as the Kipsigis.
  • E. Kitharaka language
    The Kitharaka language is a Bantu language spoken by the Tharaka people of Kenya, closely related to neighboring Kamba and Meru varieties.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (2 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69e2454bfd00819099115715a22cb057 completed April 17, 2026, 2:35 p.m.
NER Named-entity recognition batch_69f17820a8088190bc0ce907adf95863 completed April 29, 2026, 3:16 a.m.
Created at: April 17, 2026, 3:10 p.m.