Triple

T1918602
Position Surface form Disambiguated ID Type / Status
Subject Northwestern Iranian languages E40074 entity
Predicate hasHistoricalMember P18194 FINISHED
Object Tumshuqese language E214833 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Tumshuqese language | Statement: [Northwestern Iranian languages, hasHistoricalMember, Tumshuqese language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Tumshuqese language
Context triple: [Northwestern Iranian languages, hasHistoricalMember, Tumshuqese language]
  • A. Tumshuqese language chosen
    The Tumshuqese language is an extinct Middle Iranian language once spoken in the Tarim Basin region of present-day Xinjiang, China, known primarily from Buddhist and administrative manuscripts.
  • B. Amuesha language
    The Amuesha language, also known as Yanesha', is an Arawakan language spoken by the Yanesha' people of the central Peruvian Amazon.
  • C. Kumzari language
    The Kumzari language is an endangered Southwestern Iranian language spoken primarily by the Kumzari people in the Musandam Peninsula of Oman.
  • D. Esselen language
    The Esselen language is an extinct and poorly documented Native American language once spoken by the Esselen people of coastal central California.
  • E. Ghomara language
    The Ghomara language is a lesser-known Berber language spoken by the Ghomara people in northern Morocco.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69a8864298748190a2f2fd34f7ef8d77 completed March 4, 2026, 7:21 p.m.
NER Named-entity recognition batch_69abb7c51c2881908054760c624dd577 completed March 7, 2026, 5:29 a.m.
NED1 Entity disambiguation (via context triple) batch_69adfbae5760819083c046d0941513de completed March 8, 2026, 10:43 p.m.
Created at: March 4, 2026, 7:35 p.m.