Triple
T20880586
| Position | Surface form | Disambiguated ID | Type / Status |
|---|---|---|---|
| Subject | Javanese dialect continuum |
E514133
|
entity |
| Predicate | hasDialect |
P4251
|
FINISHED |
| Object | Pekalongan Javanese |
—
|
NE NERFINISHED |
How this triple was built (3 steps)
Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.
NER
Named-entity recognition
gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Pekalongan Javanese | Statement: [Javanese dialect continuum, hasDialect, Pekalongan Javanese]
NED1
Entity disambiguation (via context triple)
gpt-5-mini-2025-08-07
Target entity: Pekalongan Javanese Context triple: [Javanese dialect continuum, hasDialect, Pekalongan Javanese]
-
A.
Banyumasan Javanese
Banyumasan Javanese is a regional variety of the Javanese language spoken in the western part of Central Java, Indonesia, known for its distinct phonology, vocabulary, and more conservative linguistic features compared to standard Javanese.
-
B.
Cirebon Javanese
Cirebon Javanese is a regional variety of the Javanese language spoken around Cirebon on Java’s north coast, characterized by its distinct phonology and vocabulary influenced by Sundanese and coastal trading cultures.
-
C.
Banten Javanese
Banten Javanese is a regional variety of the Javanese language spoken primarily in the Banten province of western Java, Indonesia, characterized by its distinct phonological and lexical features.
-
D.
Middle Javanese
Middle Javanese is a historical stage of the Javanese language that developed after Old Javanese and served as a key literary and cultural medium in Java during the late medieval period.
-
E.
Surabayan Javanese
Surabayan Javanese is a regional variety of the Javanese language spoken in and around the city of Surabaya in East Java, Indonesia, characterized by its distinctive accent and vocabulary.
- F. None of above. chosen
- G. Unsure - the case is ambiguous/there is not enough information to decide.
NED2
Entity disambiguation (via description)
gpt-5-mini-2025-08-07
Target entity: Pekalongan Javanese Target entity description: Pekalongan Javanese is a regional variety of the Javanese language spoken around the city and regency of Pekalongan in Central Java, Indonesia, characterized by its own distinctive phonological and lexical features.
-
A.
Banyumasan Javanese
Banyumasan Javanese is a regional variety of the Javanese language spoken in the western part of Central Java, Indonesia, known for its distinct phonology, vocabulary, and more conservative linguistic features compared to standard Javanese.
-
B.
Cirebon Javanese
Cirebon Javanese is a regional variety of the Javanese language spoken around Cirebon on Java’s north coast, characterized by its distinct phonology and vocabulary influenced by Sundanese and coastal trading cultures.
-
C.
Banten Javanese
Banten Javanese is a regional variety of the Javanese language spoken primarily in the Banten province of western Java, Indonesia, characterized by its distinct phonological and lexical features.
-
D.
Middle Javanese
Middle Javanese is a historical stage of the Javanese language that developed after Old Javanese and served as a key literary and cultural medium in Java during the late medieval period.
-
E.
Surabayan Javanese
Surabayan Javanese is a regional variety of the Javanese language spoken in and around the city of Surabaya in East Java, Indonesia, characterized by its distinctive accent and vocabulary.
- F. None of above. chosen
Provenance (2 batches)
The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.
| Step | Stage | Batch ID | Status | When |
|---|---|---|---|---|
| creating | Elicitation | batch_69e0b4f733f081908a401c0b7beb0b9f |
completed | April 16, 2026, 10:07 a.m. |
| NER | Named-entity recognition | batch_69e6c67974348190bd3484032c0d7b31 |
completed | April 21, 2026, 12:36 a.m. |
Created at: April 16, 2026, 12:45 p.m.