Triple

T20880584
Position Surface form Disambiguated ID Type / Status
Subject Javanese dialect continuum E514133 entity
Predicate hasDialect P4251 FINISHED
Object Mataraman Javanese NE NERFINISHED

How this triple was built (3 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Mataraman Javanese | Statement: [Javanese dialect continuum, hasDialect, Mataraman Javanese]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Mataraman Javanese
Context triple: [Javanese dialect continuum, hasDialect, Mataraman Javanese]
  • A. Banyumasan Javanese
    Banyumasan Javanese is a regional variety of the Javanese language spoken in the western part of Central Java, Indonesia, known for its distinct phonology, vocabulary, and more conservative linguistic features compared to standard Javanese.
  • B. Surabayan Javanese
    Surabayan Javanese is a regional variety of the Javanese language spoken in and around the city of Surabaya in East Java, Indonesia, characterized by its distinctive accent and vocabulary.
  • C. Javanese
    The Javanese are the largest ethnic group in Indonesia, primarily inhabiting the island of Java and known for their rich cultural traditions, language, and influence on Indonesian politics and arts.
  • D. Middle Javanese
    Middle Javanese is a historical stage of the Javanese language that developed after Old Javanese and served as a key literary and cultural medium in Java during the late medieval period.
  • E. Cirebon Javanese
    Cirebon Javanese is a regional variety of the Javanese language spoken around Cirebon on Java’s north coast, characterized by its distinct phonology and vocabulary influenced by Sundanese and coastal trading cultures.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: Mataraman Javanese
Target entity description: Mataraman Javanese is a regional variety of the Javanese language spoken mainly in the western and southwestern parts of East Java, influenced by both Central Javanese and local cultural traditions.
  • A. Banyumasan Javanese
    Banyumasan Javanese is a regional variety of the Javanese language spoken in the western part of Central Java, Indonesia, known for its distinct phonology, vocabulary, and more conservative linguistic features compared to standard Javanese.
  • B. Surabayan Javanese
    Surabayan Javanese is a regional variety of the Javanese language spoken in and around the city of Surabaya in East Java, Indonesia, characterized by its distinctive accent and vocabulary.
  • C. Javanese
    The Javanese are the largest ethnic group in Indonesia, primarily inhabiting the island of Java and known for their rich cultural traditions, language, and influence on Indonesian politics and arts.
  • D. Middle Javanese
    Middle Javanese is a historical stage of the Javanese language that developed after Old Javanese and served as a key literary and cultural medium in Java during the late medieval period.
  • E. Cirebon Javanese
    Cirebon Javanese is a regional variety of the Javanese language spoken around Cirebon on Java’s north coast, characterized by its distinct phonology and vocabulary influenced by Sundanese and coastal trading cultures.
  • F. None of above. chosen

Provenance (2 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69e0b4f733f081908a401c0b7beb0b9f completed April 16, 2026, 10:07 a.m.
NER Named-entity recognition batch_69e6c67974348190bd3484032c0d7b31 completed April 21, 2026, 12:36 a.m.
Created at: April 16, 2026, 12:45 p.m.