Triple

T5379736
Position Surface form Disambiguated ID Type / Status
Subject Tuscarora language E113051 entity
Predicate relatedTo P37 FINISHED
Object Seneca language E279195 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Seneca language | Statement: [Tuscarora language, relatedTo, Seneca language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Seneca language
Context triple: [Tuscarora language, relatedTo, Seneca language]
  • A. Seneca language chosen
    The Seneca language is an Iroquoian language traditionally spoken by the Seneca people, one of the nations of the Haudenosaunee (Iroquois) Confederacy, and is the focus of ongoing revitalization efforts.
  • B. Tuvinian language
    The Tuvinian language is a Turkic language spoken primarily in the Tuva Republic of Russia, known for its rich oral traditions and use in Tuvan throat singing culture.
  • C. Sentinelese language
    The Sentinelese language is the undocumented and unclassified tongue spoken by the isolated Sentinelese people of North Sentinel Island in the Andaman archipelago.
  • D. Pamona language
    The Pamona language is an Austronesian language spoken by the Pamona people of central Sulawesi, Indonesia.
  • E. Defaka language
    The Defaka language is a highly endangered Niger-Congo language spoken by a small community in Nigeria’s Niger Delta region.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69bd4436a1988190af18dcff7fd306b4 completed March 20, 2026, 12:57 p.m.
NER Named-entity recognition batch_69bd86ce56c88190a66b3852416edccb completed March 20, 2026, 5:41 p.m.
NED1 Entity disambiguation (via context triple) batch_69bf2949cd9881908ca0d8fdf1642f71 completed March 21, 2026, 11:27 p.m.
Created at: March 20, 2026, 2:03 p.m.