Triple

T13505038
Position Surface form Disambiguated ID Type / Status
Subject tk (macrolanguage) E320993 entity
Predicate usesAlphabet P7160 FINISHED
Object Turkmen Latin alphabet E313455 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Turkmen Latin alphabet | Statement: [tk (macrolanguage), usesAlphabet, Turkmen Latin alphabet]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Turkmen Latin alphabet
Context triple: [tk (macrolanguage), usesAlphabet, Turkmen Latin alphabet]
  • A. Turkmen Cyrillic alphabet
    The Turkmen Cyrillic alphabet is a Cyrillic-based writing system formerly used for the Turkmen language, particularly during the Soviet era before its replacement by a Latin-based script.
  • B. Turkmen alphabet chosen
    The Turkmen alphabet is the standardized script used to write the Turkmen language, currently based on a modified Latin script adopted after the Soviet era.
  • C. Turkmen Arabic alphabet
    The Turkmen Arabic alphabet is a historical writing system based on the Arabic script that was formerly used to write the Turkmen language before being replaced by Latin- and Cyrillic-based alphabets.
  • D. Kazakh Latin alphabet
    The Kazakh Latin alphabet is a modern script based on the Latin writing system that has been adopted for writing the Kazakh language as part of Kazakhstan’s language reform and modernization efforts.
  • E. Uyghur Latin alphabet
    The Uyghur Latin alphabet is a romanized writing system developed for the Uyghur language, used primarily in digital communication and linguistic transcription.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69d807629d6c8190998f1b9bb12d2ed0 completed April 9, 2026, 8:09 p.m.
NER Named-entity recognition batch_69dbaf810e248190a060481004503f96 completed April 12, 2026, 2:43 p.m.
NED1 Entity disambiguation (via context triple) batch_69f77f8239c481909faf5a9c403b55f2 completed May 3, 2026, 5:01 p.m.
Created at: April 9, 2026, 9:43 p.m.