Triple

T22569910
Position Surface form Disambiguated ID Type / Status
Subject Mahakiranti (proposed) E558050 entity
Predicate includesLanguage P2177 FINISHED
Object Thami language
The Thami language is a lesser-known Sino-Tibetan language spoken by the Thami ethnic community in Nepal.
E1543631 NE FINISHED

How this triple was built (4 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Thami language | Statement: [Mahakiranti (proposed), includesLanguage, Thami language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Thami language
Context triple: [Mahakiranti (proposed), includesLanguage, Thami language]
  • A. Thakali language
    The Thakali language is a Tibeto-Burman language spoken by the Thakali people of Nepal’s Kali Gandaki region, known for its distinct dialects and close relation to neighboring Himalayan languages.
  • B. Thari language
    Thari language is an Indo-Aryan language spoken primarily by the Thari (Thar) people of Pakistan’s Thar Desert region.
  • C. Tembe language
    The Tembe language is an indigenous Tupi-Guarani language spoken by the Tembé people of northern Brazil.
  • D. Teguima language
    The Teguima language is an extinct Uto-Aztecan language once spoken by the Opata people of northern Mexico.
  • E. Thakri language
    Thakri is a lesser-known Southern Indo-Aryan language spoken by regional communities in the Indian subcontinent.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NEDg Description generation gpt-5.1
Instruction
Generate a one-sentence description of the target entity. 
You are given a context triple in the form (subject, predicate, object), where the object is the target entity. 
# Instructions
Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. 
Avoid repeating the information from the triple, unless really essential.
# Response Format
Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: Thami language
Triple: [Mahakiranti (proposed), includesLanguage, Thami language]
Generated description
The Thami language is a lesser-known Sino-Tibetan language spoken by the Thami ethnic community in Nepal.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: Thami language
Target entity description: The Thami language is a lesser-known Sino-Tibetan language spoken by the Thami ethnic community in Nepal.
  • A. Thakali language
    The Thakali language is a Tibeto-Burman language spoken by the Thakali people of Nepal’s Kali Gandaki region, known for its distinct dialects and close relation to neighboring Himalayan languages.
  • B. Thari language
    Thari language is an Indo-Aryan language spoken primarily by the Thari (Thar) people of Pakistan’s Thar Desert region.
  • C. Tembe language
    The Tembe language is an indigenous Tupi-Guarani language spoken by the Tembé people of northern Brazil.
  • D. Teguima language
    The Teguima language is an extinct Uto-Aztecan language once spoken by the Opata people of northern Mexico.
  • E. Thakri language
    Thakri is a lesser-known Southern Indo-Aryan language spoken by regional communities in the Indian subcontinent.
  • F. None of above. chosen

Provenance (5 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69e11e5ae4ac8190b1f503457603d969 completed April 16, 2026, 5:37 p.m.
NER Named-entity recognition batch_69f15fad35448190b51a3dd639ca8568 completed April 29, 2026, 1:32 a.m.
NED1 Entity disambiguation (via context triple) batch_6a0b2d75f8fc8190af7282848379184d completed May 18, 2026, 3:17 p.m.
NEDg Description generation batch_6a0b364801cc81908204a937c1099728 completed May 18, 2026, 3:54 p.m.
NED2 Entity disambiguation (via description) batch_6a0b37a1ecc08190ac89e862582833a0 completed May 18, 2026, 4 p.m.
Created at: April 16, 2026, 8:52 p.m.