Triple

T20880585
Position Surface form Disambiguated ID Type / Status
Subject Javanese dialect continuum E514133 entity
Predicate hasDialect P4251 FINISHED
Object Tegal Javanese NE NERFINISHED

How this triple was built (3 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Tegal Javanese | Statement: [Javanese dialect continuum, hasDialect, Tegal Javanese]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Tegal Javanese
Context triple: [Javanese dialect continuum, hasDialect, Tegal Javanese]
  • A. Banyumasan Javanese
    Banyumasan Javanese is a regional variety of the Javanese language spoken in the western part of Central Java, Indonesia, known for its distinct phonology, vocabulary, and more conservative linguistic features compared to standard Javanese.
  • B. Central Javanese
    Central Javanese is a major regional variety of the Javanese language spoken primarily in the central part of Java, known for its influential literary tradition and role as a cultural and linguistic standard.
  • C. Middle Javanese
    Middle Javanese is a historical stage of the Javanese language that developed after Old Javanese and served as a key literary and cultural medium in Java during the late medieval period.
  • D. Cirebon Javanese
    Cirebon Javanese is a regional variety of the Javanese language spoken around Cirebon on Java’s north coast, characterized by its distinct phonology and vocabulary influenced by Sundanese and coastal trading cultures.
  • E. Banten Javanese
    Banten Javanese is a regional variety of the Javanese language spoken primarily in the Banten province of western Java, Indonesia, characterized by its distinct phonological and lexical features.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: Tegal Javanese
Target entity description: Tegal Javanese is a regional variety of the Javanese language spoken around the city of Tegal in Central Java, known for its distinctive pronunciation and vocabulary compared to standard Javanese.
  • A. Banyumasan Javanese
    Banyumasan Javanese is a regional variety of the Javanese language spoken in the western part of Central Java, Indonesia, known for its distinct phonology, vocabulary, and more conservative linguistic features compared to standard Javanese.
  • B. Central Javanese
    Central Javanese is a major regional variety of the Javanese language spoken primarily in the central part of Java, known for its influential literary tradition and role as a cultural and linguistic standard.
  • C. Middle Javanese
    Middle Javanese is a historical stage of the Javanese language that developed after Old Javanese and served as a key literary and cultural medium in Java during the late medieval period.
  • D. Cirebon Javanese
    Cirebon Javanese is a regional variety of the Javanese language spoken around Cirebon on Java’s north coast, characterized by its distinct phonology and vocabulary influenced by Sundanese and coastal trading cultures.
  • E. Banten Javanese
    Banten Javanese is a regional variety of the Javanese language spoken primarily in the Banten province of western Java, Indonesia, characterized by its distinct phonological and lexical features.
  • F. None of above. chosen

Provenance (2 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69e0b4f733f081908a401c0b7beb0b9f completed April 16, 2026, 10:07 a.m.
NER Named-entity recognition batch_69e6c67974348190bd3484032c0d7b31 completed April 21, 2026, 12:36 a.m.
Created at: April 16, 2026, 12:45 p.m.