Triple

T21138574
Position Surface form Disambiguated ID Type / Status
Subject Richard Salomon E520873 entity
Predicate researchFocus P31 FINISHED
Object Gāndhārī language NE NERFINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Gāndhārī language | Statement: [Richard Salomon, researchFocus, Gāndhārī language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Gāndhārī language
Context triple: [Richard Salomon, researchFocus, Gāndhārī language]
  • A. Bactrian language
    The Bactrian language is an extinct Eastern Iranian language once spoken in the ancient region of Bactria, known from inscriptions and manuscripts written in a modified Greek script.
  • B. Parthian language
    Parthian language was an ancient Northwestern Iranian language once spoken in the Parthian Empire, known primarily from inscriptions and Manichaean texts.
  • C. Gandhari Prakrit chosen
    Gandhari Prakrit is an ancient Middle Indo-Aryan language of northwestern South Asia, known from Buddhist texts and inscriptions written in the Kharoṣṭhī script.
  • D. Sogdian language
    The Sogdian language was an Eastern Iranian language once widely used along the Silk Road, especially in trade and religious communities of Central Asia.
  • E. Dardic languages
    Dardic languages are a group of Indo-Aryan languages spoken primarily in the mountainous regions of northern Pakistan, India, and eastern Afghanistan.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (2 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69e0b50b53048190ae34e8abbe3c5ada completed April 16, 2026, 10:08 a.m.
NER Named-entity recognition batch_69e7235d1c788190a28577a753532b2a completed April 21, 2026, 7:12 a.m.
Created at: April 16, 2026, 2:57 p.m.