Triple

T404491
Position Surface form Disambiguated ID Type / Status
Subject Republic of the Congo E9354 entity
Predicate recognizedNationalLanguage P236 FINISHED
Object Kituba
Kituba is a widely spoken Bantu-based creole language of Central Africa, serving as a major lingua franca in the Republic of the Congo and surrounding regions.
E51306 NE FINISHED

How this triple was built (4 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Kituba | Statement: [Republic of the Congo, recognizedNationalLanguage, Kituba]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Kituba
Context triple: [Republic of the Congo, recognizedNationalLanguage, Kituba]
  • A. Kichwa
    Kichwa is a Quechuan indigenous language variety widely spoken by Andean communities in Ecuador and neighboring regions.
  • B. Kwéyòl
    Kwéyòl is a French-based Creole language spoken primarily in the Lesser Antilles, notably in Saint Lucia and Dominica.
  • C. Kiswah
    Kiswah is the ornate black cloth embroidered with Quranic verses that traditionally drapes and adorns the Kaaba in Mecca.
  • D. Mandinka
    Mandinka is a major Mande language spoken primarily in The Gambia, Senegal, Guinea-Bissau, and neighboring West African countries by the Mandinka people.
  • E. Sranan Tongo
    Sranan Tongo is an English- and Dutch-influenced creole language originating in Suriname, widely used as a lingua franca among its diverse ethnic communities.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NEDg Description generation gpt-5.1
Instruction
Generate a one-sentence description of the target entity. 
You are given a context triple in the form (subject, predicate, object), where the object is the target entity. 
# Instructions
Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. 
Avoid repeating the information from the triple, unless really essential.
# Response Format
Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: Kituba
Triple: [Republic of the Congo, recognizedNationalLanguage, Kituba]
Generated description
Kituba is a widely spoken Bantu-based creole language of Central Africa, serving as a major lingua franca in the Republic of the Congo and surrounding regions.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: Kituba
Target entity description: Kituba is a widely spoken Bantu-based creole language of Central Africa, serving as a major lingua franca in the Republic of the Congo and surrounding regions.
  • A. Kichwa
    Kichwa is a Quechuan indigenous language variety widely spoken by Andean communities in Ecuador and neighboring regions.
  • B. Kwéyòl
    Kwéyòl is a French-based Creole language spoken primarily in the Lesser Antilles, notably in Saint Lucia and Dominica.
  • C. Kiswah
    Kiswah is the ornate black cloth embroidered with Quranic verses that traditionally drapes and adorns the Kaaba in Mecca.
  • D. Mandinka
    Mandinka is a major Mande language spoken primarily in The Gambia, Senegal, Guinea-Bissau, and neighboring West African countries by the Mandinka people.
  • E. Sranan Tongo
    Sranan Tongo is an English- and Dutch-influenced creole language originating in Suriname, widely used as a lingua franca among its diverse ethnic communities.
  • F. None of above. chosen

Provenance (5 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69a2e8004cb88190b92ed1add6abf41a completed Feb. 28, 2026, 1:05 p.m.
NER Named-entity recognition batch_69a2ee2c2d1881909aa1ccfaa4172b38 completed Feb. 28, 2026, 1:31 p.m.
NED1 Entity disambiguation (via context triple) batch_69a413f594848190bc73e37f30684a37 completed March 1, 2026, 10:24 a.m.
NEDg Description generation batch_69a4144b91f88190b57876fe5b71712f completed March 1, 2026, 10:26 a.m.
NED2 Entity disambiguation (via description) batch_69a414c79bd081908717ff6368fdafc1 completed March 1, 2026, 10:28 a.m.
Created at: Feb. 28, 2026, 1:08 p.m.