Triple

T1250906
Position Surface form Disambiguated ID Type / Status
Subject Telugu script E26871 entity
Predicate writingSystemFor P454 FINISHED
Object Koya language E111783 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Koya language | Statement: [Telugu script, writingSystemFor, Koya language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Koya language
Context triple: [Telugu script, writingSystemFor, Koya language]
  • A. Koya language chosen
    Koya language is a South-Central Dravidian language spoken by the Koya tribal communities in central and southern India.
  • B. Kawaiisu language
    Kawaiisu language is an endangered Uto-Aztecan language traditionally spoken by the Kawaiisu people of southern California.
  • C. Chimariko language
    The Chimariko language is an extinct Native American language once spoken in northwestern California, often classified within the proposed Hokan language family.
  • D. Yana language
    The Yana language is an extinct Native American language once spoken in northern California, notable for its complex verb morphology and documentation by linguist Edward Sapir.
  • E. Yola language
    The Yola language was an extinct West Germanic language once spoken in County Wexford, Ireland, that preserved many archaic features derived from early English settlers.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69a49487a9c48190ba9b05348fd1b53f completed March 1, 2026, 7:33 p.m.
NER Named-entity recognition batch_69a4bf84c73c8190bbb14265cd7ab6ae completed March 1, 2026, 10:36 p.m.
NED1 Entity disambiguation (via context triple) batch_69ac93c66bf881908f4b63548341178e completed March 7, 2026, 9:08 p.m.
Created at: March 1, 2026, 7:47 p.m.