Triple

T19870992
Position Surface form Disambiguated ID Type / Status
Subject Nicobarese E477515 entity
Predicate hasPart P35 FINISHED
Object Camorta language NE NERFINISHED

How this triple was built (3 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Camorta language | Statement: [Nicobarese, hasPart, Camorta language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Camorta language
Context triple: [Nicobarese, hasPart, Camorta language]
  • A. Damara language
    The Damara language is a Khoe (Central Khoisan) language spoken primarily by the Damara people of Namibia.
  • B. Tuvinian language
    The Tuvinian language is a Turkic language spoken primarily in the Tuva Republic of Russia, known for its rich oral traditions and use in Tuvan throat singing culture.
  • C. Lasgerdi language
    The Lasgerdi language is an Iranian language spoken in parts of north-central Iran and classified within the Semnani branch of Northwestern Iranian languages.
  • D. Tabarchino language
    The Tabarchino language is a Ligurian-based Romance dialect spoken by the Tabarchini community, primarily on the islands of San Pietro and Sant’Antioco in Sardinia, Italy.
  • E. Argobba language
    The Argobba language is an endangered Ethiosemitic language spoken by the Argobba people of Ethiopia, closely related to Harari and other languages of the Harar region.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: Camorta language
Target entity description: The Camorta language is an Austroasiatic language spoken by the Nicobarese people on Camorta Island in India’s Nicobar Islands.
  • A. Damara language
    The Damara language is a Khoe (Central Khoisan) language spoken primarily by the Damara people of Namibia.
  • B. Tuvinian language
    The Tuvinian language is a Turkic language spoken primarily in the Tuva Republic of Russia, known for its rich oral traditions and use in Tuvan throat singing culture.
  • C. Lasgerdi language
    The Lasgerdi language is an Iranian language spoken in parts of north-central Iran and classified within the Semnani branch of Northwestern Iranian languages.
  • D. Tabarchino language
    The Tabarchino language is a Ligurian-based Romance dialect spoken by the Tabarchini community, primarily on the islands of San Pietro and Sant’Antioco in Sardinia, Italy.
  • E. Argobba language
    The Argobba language is an endangered Ethiosemitic language spoken by the Argobba people of Ethiopia, closely related to Harari and other languages of the Harar region.
  • F. None of above. chosen

Provenance (2 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69d8e51e7d948190aedbcd6c30361c39 completed April 10, 2026, 11:55 a.m.
NER Named-entity recognition batch_69e658a3d2b08190ad81914d4860df0e completed April 20, 2026, 4:47 p.m.
Created at: April 10, 2026, 1:51 p.m.