Triple

T20382450
Position Surface form Disambiguated ID Type / Status
Subject Daur E497871 entity
Predicate language P15 FINISHED
Object Daur language NE NERFINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Daur language | Statement: [Daur, language, Daur language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Daur language
Context triple: [Daur, language, Daur language]
  • A. Daur language chosen
    The Daur language is a Mongolic language spoken primarily by the Daur ethnic group in northeastern China, notable for preserving several archaic features of the Mongolic family.
  • B. Ulch language
    The Ulch language is a critically endangered Tungusic language spoken by the Ulch people in the Russian Far East, primarily along the lower Amur River.
  • C. Enets language
    Enets language is a critically endangered Samoyedic language of the Uralic family spoken by a small Indigenous community in northern Siberia, Russia.
  • D. Aka-Kol language
    The Aka-Kol language is an extinct Ongan language once spoken by the indigenous Great Andamanese people of the Andaman Islands in the Bay of Bengal.
  • E. Altai language
    The Altai language is a Turkic language spoken primarily in Russia’s Altai Republic and surrounding regions by the indigenous Altai people.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (2 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69e0b4a5b7908190a972e4e7e698ae94 completed April 16, 2026, 10:06 a.m.
NER Named-entity recognition batch_69e678b208ec8190a09bf5fd947a2a02 completed April 20, 2026, 7:04 p.m.
Created at: April 16, 2026, 11:27 a.m.