Triple

T9791811
Position Surface form Disambiguated ID Type / Status
Subject Central Asian American E237623 entity
Predicate languageSpoken P151 FINISHED
Object Uighur E458717 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Uighur | Statement: [Central Asian American, languageSpoken, Uighur]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Uighur
Context triple: [Central Asian American, languageSpoken, Uighur]
  • A. Uyghurs
    The Uyghurs are a Turkic-speaking, predominantly Muslim ethnic group native to the Xinjiang region of northwest China, with a distinct culture, language, and history.
  • B. Uyghur language chosen
    The Uyghur language is a Turkic language spoken primarily by the Uyghur people in China’s Xinjiang region, written in several scripts and serving as a major language of Central Asia.
  • C. Chagatai Turkic
    Chagatai Turkic is a historical Turkic literary language that served as a major cultural and administrative lingua franca in Central Asia, especially under Turkic-Mongol and Timurid rule.
  • D. Tangut
    Tangut is an extinct Tibeto-Burman language once used in the Western Xia dynasty, best known today for its large and complex logographic writing system.
  • E. Uyghur Arabic alphabet
    The Uyghur Arabic alphabet is a Perso-Arabic–based script adapted to represent the sounds of the Uyghur language, historically used by Uyghur communities in Central Asia.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69ca84dc04488190b9c91193976c0960 completed March 30, 2026, 2:12 p.m.
NER Named-entity recognition batch_69cda3456bac819097afb20ce9081dd7 completed April 1, 2026, 10:59 p.m.
NED1 Entity disambiguation (via context triple) batch_69d1c43141cc81908c45a3f5d5103564 completed April 5, 2026, 2:08 a.m.
Created at: March 30, 2026, 8:28 p.m.