Triple

T12555976
Position Surface form Disambiguated ID Type / Status
Subject Karen script E295216 entity
Predicate hasUnicodeBlock P1445 FINISHED
Object Myanmar Unicode block
The Myanmar Unicode block is a range of Unicode code points that encodes characters used for writing the Burmese language and several related scripts of Myanmar, including Karen.
E990796 NE FINISHED

How this triple was built (4 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Myanmar Unicode block | Statement: [Karen script, hasUnicodeBlock, Myanmar Unicode block]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Myanmar Unicode block
Context triple: [Karen script, hasUnicodeBlock, Myanmar Unicode block]
  • A. Myanmar Extended-B
    Myanmar Extended-B is a Unicode block that provides additional characters used in writing Burmese and related languages beyond those covered in the main Myanmar block.
  • B. Burmese script
    The Burmese script is an abugida used for writing the Burmese language and several other languages of Myanmar, characterized by its rounded letters and Indic origins.
  • C. CJK Unified Ideographs
    CJK Unified Ideographs is a standardized set of Chinese, Japanese, and Korean logographic characters encoded in Unicode to unify and represent Han-based writing systems across East Asia.
  • D. Cypriot Syllabary block
    The Cypriot Syllabary block is a range of Unicode characters representing the ancient syllabic writing system used on Cyprus for writing Greek and other languages.
  • E. Tai Nüa script
    The Tai Nüa script is an abugida used primarily by the Tai Nüa (Dai) people of China and Southeast Asia to write the Tai Nüa language.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NEDg Description generation gpt-5.1
Instruction
Generate a one-sentence description of the target entity. 
You are given a context triple in the form (subject, predicate, object), where the object is the target entity. 
# Instructions
Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. 
Avoid repeating the information from the triple, unless really essential.
# Response Format
Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: Myanmar Unicode block
Triple: [Karen script, hasUnicodeBlock, Myanmar Unicode block]
Generated description
The Myanmar Unicode block is a range of Unicode code points that encodes characters used for writing the Burmese language and several related scripts of Myanmar, including Karen.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: Myanmar Unicode block
Target entity description: The Myanmar Unicode block is a range of Unicode code points that encodes characters used for writing the Burmese language and several related scripts of Myanmar, including Karen.
  • A. Myanmar Extended-B
    Myanmar Extended-B is a Unicode block that provides additional characters used in writing Burmese and related languages beyond those covered in the main Myanmar block.
  • B. Burmese script
    The Burmese script is an abugida used for writing the Burmese language and several other languages of Myanmar, characterized by its rounded letters and Indic origins.
  • C. CJK Unified Ideographs
    CJK Unified Ideographs is a standardized set of Chinese, Japanese, and Korean logographic characters encoded in Unicode to unify and represent Han-based writing systems across East Asia.
  • D. Cypriot Syllabary block
    The Cypriot Syllabary block is a range of Unicode characters representing the ancient syllabic writing system used on Cyprus for writing Greek and other languages.
  • E. Tai Nüa script
    The Tai Nüa script is an abugida used primarily by the Tai Nüa (Dai) people of China and Southeast Asia to write the Tai Nüa language.
  • F. None of above. chosen

Provenance (5 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69d6ad9cac2c81908e8a7bed82d1e21d completed April 8, 2026, 7:33 p.m.
NER Named-entity recognition batch_69d95490d2708190857f0cb9b8dd6a30 completed April 10, 2026, 7:50 p.m.
NED1 Entity disambiguation (via context triple) batch_69f655895c9c819082f41e79906567c6 completed May 2, 2026, 7:50 p.m.
NEDg Description generation batch_69f6599e89708190bc1d38e9702d7626 completed May 2, 2026, 8:07 p.m.
NED2 Entity disambiguation (via description) batch_69f65a414118819095049600f1ad6d63 completed May 2, 2026, 8:10 p.m.
Created at: April 8, 2026, 11:47 p.m.