Triple

T4784118
Position Surface form Disambiguated ID Type / Status
Subject Somali people E106434 entity
Predicate usesScript P1587 FINISHED
Object Borama script E207229 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Borama script | Statement: [Somali people, usesScript, Borama script]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Borama script
Context triple: [Somali people, usesScript, Borama script]
  • A. Borama script chosen
    The Borama script is an indigenous writing system historically used by some Somali communities to represent the Somali language before the adoption of more widespread scripts.
  • B. Tigalari script
    The Tigalari script is a historical South Indian writing system used primarily to write Tulu and Sanskrit, closely related to other southern Brahmic scripts.
  • C. Tirhuta script
    Tirhuta script is a traditional Brahmic writing system historically used for the Maithili language of the Mithila region in India and Nepal.
  • D. Lontara script
    The Lontara script is an indigenous writing system traditionally used by the Bugis and Makassarese peoples of South Sulawesi, Indonesia, to write their Austronesian languages.
  • E. Khojki script
    The Khojki script is a historical writing system used primarily by the Nizari Ismaili community of South Asia to record religious and literary texts in languages such as Sindhi and Gujarati.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69bd43f4a9588190bf73e20bc27c03cc completed March 20, 2026, 12:56 p.m.
NER Named-entity recognition batch_69bd65ae49ec81908f16248d22d1155f completed March 20, 2026, 3:20 p.m.
NED1 Entity disambiguation (via context triple) batch_69be43dbdec88190817845e7930a18f6 completed March 21, 2026, 7:08 a.m.
Created at: March 20, 2026, 1:22 p.m.