Triple

T8169995
Position Surface form Disambiguated ID Type / Status
Subject Bantu W languages E190791 entity
Predicate hasMember P10 FINISHED
Object Suku language
Suku language is a Bantu language spoken primarily by the Suku people in the Democratic Republic of the Congo.
E715992 NE FINISHED

How this triple was built (4 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Suku language | Statement: [Bantu W languages, hasMember, Suku language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Suku language
Context triple: [Bantu W languages, hasMember, Suku language]
  • A. Suwawa language
    The Suwawa language is an Austronesian language spoken by the Suwawa people of northern Sulawesi, Indonesia, and is part of the Gorontalo–Mongondow subgroup.
  • B. Sawu language
    The Sawu language is an Austronesian language spoken primarily on Savu (Sawu) Island in eastern Indonesia.
  • C. Wuvulu-Aua language
    The Wuvulu-Aua language is an Oceanic language spoken on the Wuvulu and Aua islands of Papua New Guinea, known for its complex verbal morphology and distinctive phonological features.
  • D. Murle language
    The Murle language is an Eastern Sudanic language spoken primarily by the Murle people of South Sudan.
  • E. Kisukuma language
    Kisukuma is a major Bantu language spoken primarily by the Sukuma people in northwestern Tanzania.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NEDg Description generation gpt-5.1
Instruction
Generate a one-sentence description of the target entity. 
You are given a context triple in the form (subject, predicate, object), where the object is the target entity. 
# Instructions
Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. 
Avoid repeating the information from the triple, unless really essential.
# Response Format
Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: Suku language
Triple: [Bantu W languages, hasMember, Suku language]
Generated description
Suku language is a Bantu language spoken primarily by the Suku people in the Democratic Republic of the Congo.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: Suku language
Target entity description: Suku language is a Bantu language spoken primarily by the Suku people in the Democratic Republic of the Congo.
  • A. Suwawa language
    The Suwawa language is an Austronesian language spoken by the Suwawa people of northern Sulawesi, Indonesia, and is part of the Gorontalo–Mongondow subgroup.
  • B. Sawu language
    The Sawu language is an Austronesian language spoken primarily on Savu (Sawu) Island in eastern Indonesia.
  • C. Wuvulu-Aua language
    The Wuvulu-Aua language is an Oceanic language spoken on the Wuvulu and Aua islands of Papua New Guinea, known for its complex verbal morphology and distinctive phonological features.
  • D. Murle language
    The Murle language is an Eastern Sudanic language spoken primarily by the Murle people of South Sudan.
  • E. Kisukuma language
    Kisukuma is a major Bantu language spoken primarily by the Sukuma people in northwestern Tanzania.
  • F. None of above. chosen

Provenance (5 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69ca82c1c0a08190bf8692b4d91a03ca completed March 30, 2026, 2:03 p.m.
NER Named-entity recognition batch_69cb4803de688190960438aa059d163b completed March 31, 2026, 4:05 a.m.
NED1 Entity disambiguation (via context triple) batch_69ccbf542c388190b99fe4f0c6b7b946 completed April 1, 2026, 6:46 a.m.
NEDg Description generation batch_69ccc312a8608190b899394752ef375f completed April 1, 2026, 7:02 a.m.
NED2 Entity disambiguation (via description) batch_69ccd83115fc8190a3e276bed0a00926 completed April 1, 2026, 8:32 a.m.
Created at: March 30, 2026, 5:39 p.m.