Triple

T21718389
Position Surface form Disambiguated ID Type / Status
Subject Unicode 6.1 E536089 entity
Predicate addsScript P47142 FINISHED
Object Chakma NE NERFINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Chakma | Statement: [Unicode 6.1, addsScript, Chakma]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Chakma
Context triple: [Unicode 6.1, addsScript, Chakma]
  • A. Chakma language
    The Chakma language is an Indo-Aryan language spoken primarily by the Chakma people of Bangladesh and northeastern India.
  • B. Chakma script chosen
    Chakma script is an abugida used primarily by the Chakma people of Bangladesh and India to write the Chakma language and related liturgical texts.
  • C. Ahom language
    Ahom language is an extinct Tai language once spoken by the Ahom people of Assam in northeastern India, now preserved mainly in religious and historical manuscripts.
  • D. Chakma people
    The Chakma people are an indigenous ethnic group of South and Southeast Asia, primarily inhabiting the Chittagong Hill Tracts of Bangladesh and neighboring regions of India and Myanmar, known for their distinct language, Theravada Buddhist traditions, and rich cultural heritage.
  • E. Assamese
    Assamese is an Eastern Indo-Aryan language primarily spoken in the Indian state of Assam and recognized as one of the official languages of India.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (2 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69e0c46c6dd88190a595375fa6ebd701 completed April 16, 2026, 11:13 a.m.
NER Named-entity recognition batch_69efd96cc58081908dda09819041b888 completed April 27, 2026, 9:47 p.m.
Created at: April 16, 2026, 6:47 p.m.