Triple

T6716980
Position Surface form Disambiguated ID Type / Status
Subject Yami language E153293 entity
Predicate alternativeName P39 FINISHED
Object Tao (Yami) language E153293 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Tao (Yami) language | Statement: [Yami language, alternativeName, Tao (Yami) language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Tao (Yami) language
Context triple: [Yami language, alternativeName, Tao (Yami) language]
  • A. Yami language chosen
    The Yami language is an Austronesian language spoken by the Tao (Yami) people of Orchid Island in Taiwan, known for preserving many archaic features within the Batanic subgroup.
  • B. Tanema language
    Tanema is a nearly extinct Oceanic language once spoken on Vanikoro Island in the Temotu Province of the Solomon Islands.
  • C. Anyin language
    The Anyin language is a Niger-Congo language spoken primarily in Côte d'Ivoire and Ghana by the Anyin people, closely related to Baoulé and other Central Tano languages.
  • D. Tyap language
    Tyap language is a Plateau language of the Niger-Congo family spoken predominantly by the Atyap people in southern Kaduna State, Nigeria.
  • E. Towa language
    Towa is a Native American language spoken by the Towa (Jemez) people of New Mexico and is part of the Puebloan language family.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69c68809b4608190a2509ddb5ab87f05 completed March 27, 2026, 1:37 p.m.
NER Named-entity recognition batch_69c6d125db3c8190aad28919226a16da completed March 27, 2026, 6:49 p.m.
NED1 Entity disambiguation (via context triple) batch_69c700993128819081614ccfa68d7320 completed March 27, 2026, 10:11 p.m.
Created at: March 27, 2026, 2:07 p.m.