Triple

T176481
Position Surface form Disambiguated ID Type / Status
Subject Punjabi language E3585 entity
Predicate closelyRelatedTo P37 FINISHED
Object Lahnda languages E14560 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Lahnda languages | Statement: [Punjabi language, closelyRelatedTo, Lahnda languages]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Lahnda languages
Context triple: [Punjabi language, closelyRelatedTo, Lahnda languages]
  • A. Nuristani languages
    Nuristani languages are a small, distinct group of Indo-Iranian languages spoken primarily in the remote Nuristan region of eastern Afghanistan.
  • B. Utian languages
    The Utian languages are a small group of Native American languages once spoken in central California, traditionally including the Miwok and Costanoan (Ohlone) language branches.
  • C. Yola language
    The Yola language was an extinct West Germanic language once spoken in County Wexford, Ireland, that preserved many archaic features derived from early English settlers.
  • D. Tocharian languages
    The Tocharian languages were an extinct branch of the Indo-European family once spoken in the Tarim Basin of Central Asia, known from early medieval manuscripts and notable for their archaic linguistic features.
  • E. Saraiki chosen
    Saraiki is an Indo-Aryan language spoken primarily in central and southern Pakistan, especially in the southern Punjab region.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69a25374990081909766d30c79a18e0e completed Feb. 28, 2026, 2:31 a.m.
NER Named-entity recognition batch_69a258fd278481908ad4498e03f38e2f completed Feb. 28, 2026, 2:54 a.m.
NED1 Entity disambiguation (via context triple) batch_69a2f0b4f1708190b766e1d9b43038ed completed Feb. 28, 2026, 1:42 p.m.
Created at: Feb. 28, 2026, 2:39 a.m.