Triple
T10920908
| Position | Surface form | Disambiguated ID | Type / Status |
|---|---|---|---|
| Subject | Tai Viet script |
E257942
|
entity |
| Predicate | usedFor |
P98
|
FINISHED |
| Object | Tai Dam language |
E313544
|
NE FINISHED |
How this triple was built (2 steps)
Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.
NER
Named-entity recognition
gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Tai Dam language | Statement: [Tai Viet script, usedFor, Tai Dam language]
NED1
Entity disambiguation (via context triple)
gpt-5-mini-2025-08-07
Target entity: Tai Dam language Context triple: [Tai Viet script, usedFor, Tai Dam language]
-
A.
Tai Dam language
chosen
The Tai Dam language is a Southwestern Tai language spoken primarily by the Tai Dam (Black Tai) people in parts of Vietnam, Laos, Thailand, and China.
-
B.
Damana language
The Damana language is an indigenous Chibchan tongue spoken by the Wiwa people of the Sierra Nevada de Santa Marta region in northern Colombia.
-
C.
Padam language
Padam language is a Tani (Sino-Tibetan) language of northeastern India spoken by the Padam subgroup of the Mishing/Adi peoples of Arunachal Pradesh and Assam.
-
D.
Damara language
The Damara language is a Khoe (Central Khoisan) language spoken primarily by the Damara people of Namibia.
-
E.
Tanema language
Tanema is a nearly extinct Oceanic language once spoken on Vanikoro Island in the Temotu Province of the Solomon Islands.
- F. None of above.
- G. Unsure - the case is ambiguous/there is not enough information to decide.
Provenance (3 batches)
The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.
| Step | Stage | Batch ID | Status | When |
|---|---|---|---|---|
| creating | Elicitation | batch_69d6aa864ed88190818280ab6791d065 |
completed | April 8, 2026, 7:20 p.m. |
| NER | Named-entity recognition | batch_69d77082a1488190850a4409339c3e1e |
completed | April 9, 2026, 9:25 a.m. |
| NED1 | Entity disambiguation (via context triple) | batch_69e23bcee80481909a9ec8a03bc5266d |
completed | April 17, 2026, 1:55 p.m. |
Created at: April 8, 2026, 9:22 p.m.