Triple

T7743211
Position Surface form Disambiguated ID Type / Status
Subject Pwo Karen language E175559 entity
Predicate closelyRelatedTo P37 FINISHED
Object Pa-O language E179082 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Pa-O language | Statement: [Pwo Karen language, closelyRelatedTo, Pa-O language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Pa-O language
Context triple: [Pwo Karen language, closelyRelatedTo, Pa-O language]
  • A. Pa’O language chosen
    The Pa’O language is a Tibeto-Burman language spoken primarily by the Pa’O (Taungthu) people of Myanmar, especially in Shan and Kayin States.
  • B. Paunaka language
    The Paunaka language is an endangered Arawakan language traditionally spoken by the Paunaka people of eastern Bolivia.
  • C. Patamona language
    The Patamona language is an indigenous Cariban language spoken by the Patamona people of the Guiana Highlands in Guyana and northern Brazil.
  • D. Opata language
    The Opata language is an extinct Uto-Aztecan language once spoken by the Opata people of northern Mexico, particularly in the present-day state of Sonora.
  • E. Tai Khamyang language
    The Tai Khamyang language is an endangered Southwestern Tai language spoken by the Khamyang ethnic community in parts of Northeast India, notable for its close relation to other Tai languages of the region and its ongoing revitalization efforts.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69c6995f9c60819092e386192bd63c6f completed March 27, 2026, 2:51 p.m.
NER Named-entity recognition batch_69c70388d58081909aad2c03b4501e78 completed March 27, 2026, 10:24 p.m.
NED1 Entity disambiguation (via context triple) batch_69c8be48d61c8190aba1e5f23d7cb1be completed March 29, 2026, 5:53 a.m.
Created at: March 27, 2026, 4:07 p.m.