Triple

T8802422
Position Surface form Disambiguated ID Type / Status
Subject Proto-Southern Dravidian E209443 entity
Predicate ancestorOf P369 FINISHED
Object Toda
Toda is a Southern Dravidian language spoken by the Toda people of the Nilgiri Hills in southern India, known for its highly complex phonology and small speaker population.
E759157 NE FINISHED

How this triple was built (4 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Toda | Statement: [Proto-Southern Dravidian, ancestorOf, Toda]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Toda
Context triple: [Proto-Southern Dravidian, ancestorOf, Toda]
  • A. Toda
    Toda is a subgroup of the Seediq, an Indigenous people of Taiwan known for their distinct language and cultural traditions.
  • B. Tiba
    Tiba is a modern planned city in Egypt’s Luxor Governorate, developed to accommodate population growth and support regional economic and urban expansion.
  • C. Toma
    Toma is a major Mande language spoken primarily in Guinea and neighboring West African countries.
  • D. Mijas
    Mijas is a picturesque municipality in the province of Málaga in southern Spain, known for its whitewashed village, coastal resorts, and location along the Costa del Sol.
  • E. Tomia
    Tomia is an island in Indonesia’s Wakatobi archipelago, renowned for its pristine coral reefs and exceptional scuba diving and snorkeling sites.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NEDg Description generation gpt-5.1
Instruction
Generate a one-sentence description of the target entity. 
You are given a context triple in the form (subject, predicate, object), where the object is the target entity. 
# Instructions
Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. 
Avoid repeating the information from the triple, unless really essential.
# Response Format
Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: Toda
Triple: [Proto-Southern Dravidian, ancestorOf, Toda]
Generated description
Toda is a Southern Dravidian language spoken by the Toda people of the Nilgiri Hills in southern India, known for its highly complex phonology and small speaker population.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: Toda
Target entity description: Toda is a Southern Dravidian language spoken by the Toda people of the Nilgiri Hills in southern India, known for its highly complex phonology and small speaker population.
  • A. Toda
    Toda is a subgroup of the Seediq, an Indigenous people of Taiwan known for their distinct language and cultural traditions.
  • B. Tiba
    Tiba is a modern planned city in Egypt’s Luxor Governorate, developed to accommodate population growth and support regional economic and urban expansion.
  • C. Toma
    Toma is a major Mande language spoken primarily in Guinea and neighboring West African countries.
  • D. Mijas
    Mijas is a picturesque municipality in the province of Málaga in southern Spain, known for its whitewashed village, coastal resorts, and location along the Costa del Sol.
  • E. Tomia
    Tomia is an island in Indonesia’s Wakatobi archipelago, renowned for its pristine coral reefs and exceptional scuba diving and snorkeling sites.
  • F. None of above. chosen

Provenance (5 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69ca836320e48190b5cf585b90a322c4 completed March 30, 2026, 2:06 p.m.
NER Named-entity recognition batch_69cc5fbb5b108190a9f889d40aa20521 completed March 31, 2026, 11:58 p.m.
NED1 Entity disambiguation (via context triple) batch_69cf6f799f00819089159da177c816e9 completed April 3, 2026, 7:42 a.m.
NEDg Description generation batch_69cf708dbd54819099efa4b5729d6298 completed April 3, 2026, 7:47 a.m.
NED2 Entity disambiguation (via description) batch_69cf7163e2088190bf252896cc4036b2 completed April 3, 2026, 7:51 a.m.
Created at: March 30, 2026, 6:44 p.m.