Triple

T8170839
Position Surface form Disambiguated ID Type / Status
Subject Eastern Neo-Brahmi script E190813 entity
Predicate relatedTo P37 FINISHED
Object Tirhuta script E51541 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Tirhuta script | Statement: [Eastern Neo-Brahmi script, relatedTo, Tirhuta script]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Tirhuta script
Context triple: [Eastern Neo-Brahmi script, relatedTo, Tirhuta script]
  • A. Tirhuta script chosen
    Tirhuta script is a traditional Brahmic writing system historically used for the Maithili language of the Mithila region in India and Nepal.
  • B. Tigalari script
    The Tigalari script is a historical South Indian writing system used primarily to write Tulu and Sanskrit, closely related to other southern Brahmic scripts.
  • C. Khojki script
    The Khojki script is a historical writing system used primarily by the Nizari Ismaili community of South Asia to record religious and literary texts in languages such as Sindhi and Gujarati.
  • D. Ruqʿah script
    Ruqʿah script is a simple, highly legible Arabic handwriting style commonly used for everyday writing and official documents in the Arab world.
  • E. Madnhaya script
    The Madnhaya script is a modern cursive form of the Syriac alphabet used primarily by Assyrian and Chaldean Christian communities in liturgical and literary contexts.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69ca82c1c0a08190bf8692b4d91a03ca completed March 30, 2026, 2:03 p.m.
NER Named-entity recognition batch_69cb48056d0c819094575090a41e0083 completed March 31, 2026, 4:05 a.m.
NED1 Entity disambiguation (via context triple) batch_69ccbf5cfd588190b12ef9b5799ffd88 completed April 1, 2026, 6:46 a.m.
Created at: March 30, 2026, 5:39 p.m.