Triple

T20170522
Position Surface form Disambiguated ID Type / Status
Subject Oroqen E491945 entity
Predicate traditionalLanguage P6149 FINISHED
Object Oroqen language NE NERFINISHED

How this triple was built (3 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Oroqen language | Statement: [Oroqen, traditionalLanguage, Oroqen language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Oroqen language
Context triple: [Oroqen, traditionalLanguage, Oroqen language]
  • A. Oroqen dialect
    The Oroqen dialect is a regional variety of the Tungusic Evenki language spoken primarily by the Oroqen people of northeastern China.
  • B. Orok language
    The Orok language is a critically endangered Tungusic language spoken by the Orok (Uilta) people of Sakhalin Island in Russia.
  • C. Oron language
    The Oron language is a Niger-Congo language spoken by the Oron people of southeastern Nigeria, particularly in coastal areas of Akwa Ibom State.
  • D. Oro language
    The Oro language is the traditional indigenous language spoken by the Oro people of Papua New Guinea.
  • E. Nyoro language
    The Nyoro language is a Bantu language spoken primarily by the Banyoro people in western Uganda.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: Oroqen language
Target entity description: The Oroqen language is a critically endangered Tungusic language spoken by the Oroqen people of northeastern China and parts of Siberia.
  • A. Oroqen dialect
    The Oroqen dialect is a regional variety of the Tungusic Evenki language spoken primarily by the Oroqen people of northeastern China.
  • B. Orok language
    The Orok language is a critically endangered Tungusic language spoken by the Orok (Uilta) people of Sakhalin Island in Russia.
  • C. Oron language
    The Oron language is a Niger-Congo language spoken by the Oron people of southeastern Nigeria, particularly in coastal areas of Akwa Ibom State.
  • D. Oro language
    The Oro language is the traditional indigenous language spoken by the Oro people of Papua New Guinea.
  • E. Nyoro language
    The Nyoro language is a Bantu language spoken primarily by the Banyoro people in western Uganda.
  • F. None of above. chosen

Provenance (2 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69da6266c6888190bc1a3ecf24814d34 completed April 11, 2026, 3:01 p.m.
NER Named-entity recognition batch_69e66847ed9481908e6b23b399fa7005 completed April 20, 2026, 5:54 p.m.
Created at: April 11, 2026, 11:35 p.m.