Triple
T20880585
| Position | Surface form | Disambiguated ID | Type / Status |
|---|---|---|---|
| Subject | Javanese dialect continuum |
E514133
|
entity |
| Predicate | hasDialect |
P4251
|
FINISHED |
| Object | Tegal Javanese |
—
|
NE NERFINISHED |
How this triple was built (3 steps)
Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.
NER
Named-entity recognition
gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Tegal Javanese | Statement: [Javanese dialect continuum, hasDialect, Tegal Javanese]
NED1
Entity disambiguation (via context triple)
gpt-5-mini-2025-08-07
Target entity: Tegal Javanese Context triple: [Javanese dialect continuum, hasDialect, Tegal Javanese]
-
A.
Banyumasan Javanese
Banyumasan Javanese is a regional variety of the Javanese language spoken in the western part of Central Java, Indonesia, known for its distinct phonology, vocabulary, and more conservative linguistic features compared to standard Javanese.
-
B.
Central Javanese
Central Javanese is a major regional variety of the Javanese language spoken primarily in the central part of Java, known for its influential literary tradition and role as a cultural and linguistic standard.
-
C.
Middle Javanese
Middle Javanese is a historical stage of the Javanese language that developed after Old Javanese and served as a key literary and cultural medium in Java during the late medieval period.
-
D.
Cirebon Javanese
Cirebon Javanese is a regional variety of the Javanese language spoken around Cirebon on Java’s north coast, characterized by its distinct phonology and vocabulary influenced by Sundanese and coastal trading cultures.
-
E.
Banten Javanese
Banten Javanese is a regional variety of the Javanese language spoken primarily in the Banten province of western Java, Indonesia, characterized by its distinct phonological and lexical features.
- F. None of above. chosen
- G. Unsure - the case is ambiguous/there is not enough information to decide.
NED2
Entity disambiguation (via description)
gpt-5-mini-2025-08-07
Target entity: Tegal Javanese Target entity description: Tegal Javanese is a regional variety of the Javanese language spoken around the city of Tegal in Central Java, known for its distinctive pronunciation and vocabulary compared to standard Javanese.
-
A.
Banyumasan Javanese
Banyumasan Javanese is a regional variety of the Javanese language spoken in the western part of Central Java, Indonesia, known for its distinct phonology, vocabulary, and more conservative linguistic features compared to standard Javanese.
-
B.
Central Javanese
Central Javanese is a major regional variety of the Javanese language spoken primarily in the central part of Java, known for its influential literary tradition and role as a cultural and linguistic standard.
-
C.
Middle Javanese
Middle Javanese is a historical stage of the Javanese language that developed after Old Javanese and served as a key literary and cultural medium in Java during the late medieval period.
-
D.
Cirebon Javanese
Cirebon Javanese is a regional variety of the Javanese language spoken around Cirebon on Java’s north coast, characterized by its distinct phonology and vocabulary influenced by Sundanese and coastal trading cultures.
-
E.
Banten Javanese
Banten Javanese is a regional variety of the Javanese language spoken primarily in the Banten province of western Java, Indonesia, characterized by its distinct phonological and lexical features.
- F. None of above. chosen
Provenance (2 batches)
The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.
| Step | Stage | Batch ID | Status | When |
|---|---|---|---|---|
| creating | Elicitation | batch_69e0b4f733f081908a401c0b7beb0b9f |
completed | April 16, 2026, 10:07 a.m. |
| NER | Named-entity recognition | batch_69e6c67974348190bd3484032c0d7b31 |
completed | April 21, 2026, 12:36 a.m. |
Created at: April 16, 2026, 12:45 p.m.