Triple

T30769248
Position Surface form Disambiguated ID Type / Status
Subject Semelai E783464 entity
Predicate hasAlternativeName P39 FINISHED
Object Semelai language
The Semelai language is an Austroasiatic language spoken by the Semelai indigenous people of Peninsular Malaysia, primarily in the central and southern regions.
E1931593 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Semelai language | Statement: [Semelai, hasAlternativeName, Semelai language]
NEDg Description generation gpt-5.1
Instruction
Generate a one-sentence description of the target entity. 
You are given a context triple in the form (subject, predicate, object), where the object is the target entity. 
# Instructions
Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. 
Avoid repeating the information from the triple, unless really essential.
# Response Format
Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: Semelai language
Triple: [Semelai, hasAlternativeName, Semelai language]
Generated description
The Semelai language is an Austroasiatic language spoken by the Semelai indigenous people of Peninsular Malaysia, primarily in the central and southern regions.

Provenance (5 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69f224b1519081908b9db003fd2073e0 completed April 29, 2026, 3:33 p.m.
NER Named-entity recognition batch_69f68fbffc3481909fc9b762f2cc5dd4 completed May 2, 2026, 11:58 p.m.
NED1 Entity disambiguation (via context triple) batch_6a28b098d7c08190a5a149c3a8f56363 completed June 10, 2026, 12:32 a.m.
NEDg Description generation batch_6a28b339bb848190b2ace288998c2bf4 completed June 10, 2026, 12:43 a.m.
NED2 Entity disambiguation (via description) batch_6a28b47c1c008190b45a0d6dd6bbda00 completed June 10, 2026, 12:49 a.m.
Created at: April 29, 2026, 8:40 p.m.