Triple

T13428689
Position Surface form Disambiguated ID Type / Status
Subject Ahom language E313549 entity
Predicate hasScriptUnicodeBlock P1445 FINISHED
Object Ahom Unicode block E657085 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Ahom Unicode block | Statement: [Ahom language, hasScriptUnicodeBlock, Ahom Unicode block]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Ahom Unicode block
Context triple: [Ahom language, hasScriptUnicodeBlock, Ahom Unicode block]
  • A. Ahom script chosen
    The Ahom script is an abugida historically used to write the Tai Ahom language of the Ahom people in what is now Assam, India.
  • B. Myanmar Unicode block
    The Myanmar Unicode block is a range of Unicode code points that encodes characters used for writing the Burmese language and several related scripts of Myanmar, including Karen.
  • C. Tai Nüa script
    The Tai Nüa script is an abugida used primarily by the Tai Nüa (Dai) people of China and Southeast Asia to write the Tai Nüa language.
  • D. Chakma script
    Chakma script is an abugida used primarily by the Chakma people of Bangladesh and India to write the Chakma language and related liturgical texts.
  • E. Myanmar Extended-A
    Myanmar Extended-A is a Unicode block that provides additional characters used primarily for writing the Shan script and related languages of Myanmar.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69d806ad0c44819088833ae1ec9e9690 completed April 9, 2026, 8:06 p.m.
NER Named-entity recognition batch_69dbaed304ac8190a8021f749de8164c completed April 12, 2026, 2:40 p.m.
NED1 Entity disambiguation (via context triple) batch_69f730883cb48190add9469c48dc3e89 completed May 3, 2026, 11:24 a.m.
Created at: April 9, 2026, 9:40 p.m.