Triple

T9031857
Position Surface form Disambiguated ID Type / Status
Subject Chakhesang Naga E216390 entity
Predicate hasLanguage P15 FINISHED
Object Khezha language
The Khezha language is a Sino-Tibetan language spoken primarily by the Chakhesang Naga community in northeastern India.
E772405 NE FINISHED

How this triple was built (4 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Khezha language | Statement: [Chakhesang Naga, hasLanguage, Khezha language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Khezha language
Context triple: [Chakhesang Naga, hasLanguage, Khezha language]
  • A. Ezha language
    The Ezha language is an Afroasiatic Semitic language spoken by the Ezha people of Ethiopia as part of the Gurage linguistic group.
  • B. Hezhen language
    The Hezhen language is a critically endangered Tungusic language spoken by the Hezhen (Nanai) people of northeastern China along the Amur and Ussuri rivers.
  • C. Kaxabu language
    The Kaxabu language is an indigenous Formosan language of Taiwan spoken by the Kaxabu people and considered highly endangered.
  • D. Khortha language
    Khortha is an Eastern Indo-Aryan language spoken primarily in the Indian state of Jharkhand, where it serves as a major regional lingua franca among various communities.
  • E. Kaera language
    The Kaera language is a Papuan language spoken by a small community on Pantar Island in eastern Indonesia.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NEDg Description generation gpt-5.1
Instruction
Generate a one-sentence description of the target entity. 
You are given a context triple in the form (subject, predicate, object), where the object is the target entity. 
# Instructions
Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. 
Avoid repeating the information from the triple, unless really essential.
# Response Format
Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: Khezha language
Triple: [Chakhesang Naga, hasLanguage, Khezha language]
Generated description
The Khezha language is a Sino-Tibetan language spoken primarily by the Chakhesang Naga community in northeastern India.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: Khezha language
Target entity description: The Khezha language is a Sino-Tibetan language spoken primarily by the Chakhesang Naga community in northeastern India.
  • A. Ezha language
    The Ezha language is an Afroasiatic Semitic language spoken by the Ezha people of Ethiopia as part of the Gurage linguistic group.
  • B. Hezhen language
    The Hezhen language is a critically endangered Tungusic language spoken by the Hezhen (Nanai) people of northeastern China along the Amur and Ussuri rivers.
  • C. Kaxabu language
    The Kaxabu language is an indigenous Formosan language of Taiwan spoken by the Kaxabu people and considered highly endangered.
  • D. Khortha language
    Khortha is an Eastern Indo-Aryan language spoken primarily in the Indian state of Jharkhand, where it serves as a major regional lingua franca among various communities.
  • E. Kaera language
    The Kaera language is a Papuan language spoken by a small community on Pantar Island in eastern Indonesia.
  • F. None of above. chosen

Provenance (5 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69ca83d10b608190b2b2f8e0a7faaf14 completed March 30, 2026, 2:08 p.m.
NER Named-entity recognition batch_69cc6a9f2c7481909b4a272183f20585 completed April 1, 2026, 12:45 a.m.
NED1 Entity disambiguation (via context triple) batch_69cfdbc9c6e08190aa71d84316afc6d5 completed April 3, 2026, 3:24 p.m.
NEDg Description generation batch_69cfdcb95b508190b9d5562f4248e074 completed April 3, 2026, 3:28 p.m.
NED2 Entity disambiguation (via description) batch_69cfdd3293888190a9cd5e0621c2fe10 completed April 3, 2026, 3:30 p.m.
Created at: March 30, 2026, 7:08 p.m.