Triple

T16735137
Position Surface form Disambiguated ID Type / Status
Subject Multan District E406696 entity
Predicate majorLanguage P207 FINISHED
Object Saraiki E14560 NE FINISHED

Named-entity recognition

Before disambiguation, gpt-5-mini classified whether the object phrase is a named entity — the step behind the object's NE type shown above.

Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Saraiki | Statement: [Multan District, majorLanguage, Saraiki]

Disambiguation candidates (1 decision)

The exact options the model was shown at each disambiguation step, with the option it chose highlighted — the evidence behind this triple's disambiguated ids.

NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Saraiki
Context triple: [Multan District, majorLanguage, Saraiki]
  • A. Saraiki chosen
    Saraiki is an Indo-Aryan language spoken primarily in central and southern Pakistan, especially in the southern Punjab region.
  • B. Abbottabadi Hindko
    Abbottabadi Hindko is a regional variety of the Hindko language spoken primarily in and around the city of Abbottabad in northern Pakistan.
  • C. Peshawari Hindko
    Peshawari Hindko is a regional variety of the Hindko language spoken primarily in and around the city of Peshawar in Pakistan.
  • D. Pahari-Pothwari
    Pahari-Pothwari is an Indo-Aryan language variety spoken primarily in the Pothohar Plateau of northern Pakistan and adjoining regions, often considered transitional between Punjabi and Hindko.
  • E. Khowar
    Khowar is an Indo-Aryan language primarily spoken in the Chitral region of Pakistan and parts of neighboring areas, known for its rich oral tradition and distinct phonology.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

Stage Batch ID Job type Status
creating batch_69d8838ffb088190a0b11149929006bf elicitation completed
NER batch_69e39c39c570819088723d59242c5c5c ner completed
NED1 batch_6a009d4c723c8190ad92628f4164d11c ned_source_triple completed
Created at: April 10, 2026, 5:20 a.m.