Triple

T2756879
Position Surface form Disambiguated ID Type / Status
Subject Jharkhand E61121 entity
Predicate regionalLanguage P237 FINISHED
Object Khortha
Khortha is an Indo-Aryan language spoken primarily in the Indian state of Jharkhand, where it serves as a major regional vernacular among local communities.
E295178 NE FINISHED

How this triple was built (4 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Khortha | Statement: [Jharkhand, regionalLanguage, Khortha]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Khortha
Context triple: [Jharkhand, regionalLanguage, Khortha]
  • A. Krakhuna
    Krakhuna is a Georgian white grape variety from the Imereti region, known for producing aromatic, full-bodied wines with pronounced acidity.
  • B. Kiphire
    Kiphire is a town and administrative district headquarters in the northeastern Indian state of Nagaland, known for its hilly terrain and proximity to Mount Saramati.
  • C. Harauti
    Harauti is an Indo-Aryan dialect of the Rajasthani language spoken primarily in the Hadoti region of Rajasthan, India.
  • D. Khoni
    Khoni is a small town in western Georgia’s Imereti region, known for its historical churches and surrounding natural landscapes.
  • E. Kungara
    Kungara is an alternative name for the Fur language, a Nilo-Saharan language spoken primarily by the Fur people of western Sudan.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NEDg Description generation gpt-5.1
Instruction
Generate a one-sentence description of the target entity. 
You are given a context triple in the form (subject, predicate, object), where the object is the target entity. 
# Instructions
Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. 
Avoid repeating the information from the triple, unless really essential.
# Response Format
Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: Khortha
Triple: [Jharkhand, regionalLanguage, Khortha]
Generated description
Khortha is an Indo-Aryan language spoken primarily in the Indian state of Jharkhand, where it serves as a major regional vernacular among local communities.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: Khortha
Target entity description: Khortha is an Indo-Aryan language spoken primarily in the Indian state of Jharkhand, where it serves as a major regional vernacular among local communities.
  • A. Krakhuna
    Krakhuna is a Georgian white grape variety from the Imereti region, known for producing aromatic, full-bodied wines with pronounced acidity.
  • B. Kiphire
    Kiphire is a town and administrative district headquarters in the northeastern Indian state of Nagaland, known for its hilly terrain and proximity to Mount Saramati.
  • C. Harauti
    Harauti is an Indo-Aryan dialect of the Rajasthani language spoken primarily in the Hadoti region of Rajasthan, India.
  • D. Khoni
    Khoni is a small town in western Georgia’s Imereti region, known for its historical churches and surrounding natural landscapes.
  • E. Kungara
    Kungara is an alternative name for the Fur language, a Nilo-Saharan language spoken primarily by the Fur people of western Sudan.
  • F. None of above. chosen

Provenance (5 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69ab4b7a85bc819094a349b84beb1f2c completed March 6, 2026, 9:47 p.m.
NER Named-entity recognition batch_69abdb8a292c8190ab3982434805241a completed March 7, 2026, 8:02 a.m.
NED1 Entity disambiguation (via context triple) batch_69afbbdf650c8190baa020143b51f94b completed March 10, 2026, 6:36 a.m.
NEDg Description generation batch_69afbc67c39c8190b5932c0e23595f64 completed March 10, 2026, 6:38 a.m.
NED2 Entity disambiguation (via description) batch_69afbd2d8a2c8190896a9154ebbd8bab completed March 10, 2026, 6:41 a.m.
Created at: March 6, 2026, 9:56 p.m.