Triple

T16510071
Position Surface form Disambiguated ID Type / Status
Subject Multan Division E401036 entity
Predicate languageUsed P238 FINISHED
Object Saraiki E14560 NE FINISHED

Named-entity recognition

Before disambiguation, gpt-5-mini classified whether the object phrase is a named entity — the step behind the object's NE type shown above.

Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Saraiki | Statement: [Multan Division, languageUsed, Saraiki]

Disambiguation candidates (1 decision)

The exact options the model was shown at each disambiguation step, with the option it chose highlighted — the evidence behind this triple's disambiguated ids.

NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Saraiki
Context triple: [Multan Division, languageUsed, Saraiki]
  • A. Saraiki chosen
    Saraiki is an Indo-Aryan language spoken primarily in central and southern Pakistan, especially in the southern Punjab region.
  • B. Abbottabadi Hindko
    Abbottabadi Hindko is a regional variety of the Hindko language spoken primarily in and around the city of Abbottabad in northern Pakistan.
  • C. Peshawari Hindko
    Peshawari Hindko is a regional variety of the Hindko language spoken primarily in and around the city of Peshawar in Pakistan.
  • D. Pahari-Pothwari
    Pahari-Pothwari is an Indo-Aryan language variety spoken primarily in the Pothohar Plateau of northern Pakistan and adjoining regions, often considered transitional between Punjabi and Hindko.
  • E. Khowar
    Khowar is an Indo-Aryan language primarily spoken in the Chitral region of Pakistan and parts of neighboring areas, known for its rich oral tradition and distinct phonology.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

Stage Batch ID Job type Status
creating batch_69d88381f6148190819958a038be990e elicitation completed
NER batch_69e32e54f7508190804bbae4c9bc8fe3 ner completed
NED1 batch_6a00608038288190925586dfb7e64689 ned_source_triple completed
Created at: April 10, 2026, 5:14 a.m.