Triple

T3648562
Position Surface form Disambiguated ID Type / Status
Subject Finno-Ugric languages E77361 entity
Predicate hasMember P10 FINISHED
Object Khanty language E357085 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Khanty language | Statement: [Finno-Ugric languages, hasMember, Khanty language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Khanty language
Context triple: [Finno-Ugric languages, hasMember, Khanty language]
  • A. Khanty language chosen
    The Khanty language is a Uralic language spoken by the Khanty people of western Siberia, closely related to Mansi and traditionally used in the Khanty-Mansi Autonomous Okrug of Russia.
  • B. Enets language
    Enets language is a critically endangered Samoyedic language of the Uralic family spoken by a small Indigenous community in northern Siberia, Russia.
  • C. Nenets language
    The Nenets language is a Uralic Samoyedic language spoken by the Nenets people of northern Arctic Russia.
  • D. Selkup language
    The Selkup language is a critically endangered Uralic (Samoyedic) language spoken by the indigenous Selkup people of western Siberia in Russia.
  • E. Komi-Permyak language
    The Komi-Permyak language is a Uralic language spoken by the Komi-Permyak people in Russia’s Perm Krai, closely related to other Permic languages and written in a Cyrillic-based script.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69ad85de1b988190a45f8dbfebc806fc completed March 8, 2026, 2:21 p.m.
NER Named-entity recognition batch_69adc38c22548190a271a69fb832a5a8 completed March 8, 2026, 6:44 p.m.
NED1 Entity disambiguation (via context triple) batch_69b4c38f989c8190befc64db51041a53 completed March 14, 2026, 2:10 a.m.
Created at: March 8, 2026, 3:24 p.m.