Triple

T257224
Position Surface form Disambiguated ID Type / Status
Subject Yalta E5461 entity
Predicate languageUsed P238 FINISHED
Object Crimean Tatar E6711 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Crimean Tatar | Statement: [Yalta, languageUsed, Crimean Tatar]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Crimean Tatar
Context triple: [Yalta, languageUsed, Crimean Tatar]
  • A. Crimean Tatars chosen
    Crimean Tatars are a Turkic ethnic group indigenous to the Crimean Peninsula, with a distinct language, culture, and history marked by periods of autonomy, repression, and diaspora.
  • B. Balkan Turkish
    Balkan Turkish is a regional variety of the Turkish language spoken by Turkish communities across several Balkan countries, influenced by local languages and cultures.
  • C. Turks
    Turks are a Turkic ethnic group primarily associated with Turkey and other regions of the former Ottoman Empire, including parts of the Balkans.
  • D. Ottoman Turkish
    Ottoman Turkish was the administrative and literary language of the Ottoman Empire, blending Turkish with extensive Arabic and Persian influences and written in a variant of the Arabic script.
  • E. Turkmen language
    The Turkmen language is a Turkic language spoken primarily in Turkmenistan and surrounding regions, closely related to Turkish and other Oghuz languages.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69a2580a64ac8190ad76e34bb0715b5e completed Feb. 28, 2026, 2:50 a.m.
NER Named-entity recognition batch_69a25d5884c88190a349d7593b688921 completed Feb. 28, 2026, 3:13 a.m.
NED1 Entity disambiguation (via context triple) batch_69a37958242c81909d114fba70b4211c completed Feb. 28, 2026, 11:25 p.m.
Created at: Feb. 28, 2026, 2:55 a.m.