Triple

T5008791
Position Surface form Disambiguated ID Type / Status
Subject Lezgian E112561 entity
Predicate closelyRelatedTo P37 FINISHED
Object Tabasaran language E222232 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Tabasaran language | Statement: [Lezgian, closelyRelatedTo, Tabasaran language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Tabasaran language
Context triple: [Lezgian, closelyRelatedTo, Tabasaran language]
  • A. Damara language
    The Damara language is a Khoe (Central Khoisan) language spoken primarily by the Damara people of Namibia.
  • B. Piipaash language
    The Piipaash language is a Native American language of the Yuman family traditionally spoken by the Piipaash (Maricopa) people of the lower Colorado River region in the southwestern United States.
  • C. Hoanya language
    The Hoanya language is an extinct Austronesian language once spoken by the Hoanya people of western Taiwan and classified among the indigenous Formosan languages.
  • D. Amuesha language
    The Amuesha language, also known as Yanesha', is an Arawakan language spoken by the Yanesha' people of the central Peruvian Amazon.
  • E. Khwarshi language chosen
    The Khwarshi language is a Northeast Caucasian (Nakh-Daghestanian) language spoken by a small ethnic group in Dagestan, Russia, known for its complex phonology and rich case system.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69bd4433d0b08190877e83959ef40d81 completed March 20, 2026, 12:57 p.m.
NER Named-entity recognition batch_69bd72eb05f881908d7dc3d7cd07b2ae completed March 20, 2026, 4:16 p.m.
NED1 Entity disambiguation (via context triple) batch_69be9269e72881908ea49a77a83b8958 completed March 21, 2026, 12:43 p.m.
Created at: March 20, 2026, 1:35 p.m.