Triple

T10127311
Position Surface form Disambiguated ID Type / Status
Subject Karakalpak language E226246 entity
Predicate languageFamilyAncestor P30710 FINISHED
Object Proto-Turkic language E96950 NE FINISHED

How this triple was built (3 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Proto-Turkic language | Statement: [Karakalpak language, languageFamilyAncestor, Proto-Turkic language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Proto-Turkic language
Context triple: [Karakalpak language, languageFamilyAncestor, Proto-Turkic language]
  • A. Proto-Turkic chosen
    Proto-Turkic is the reconstructed common ancestor language of all Turkic languages, from which branches like Southwestern Turkic later evolved.
  • B. Proto-Mongolic language
    Proto-Mongolic language is the reconstructed common ancestor of the Mongolic language family, hypothesized through comparative linguistic methods.
  • C. Proto-Ugric language
    Proto-Ugric language is a hypothesized prehistoric ancestor of the Ugric branch of the Uralic language family, reconstructed through comparative linguistic methods.
  • D. Oghuz Turkic language
    Oghuz Turkic language is a major branch of the Turkic language family that includes modern languages such as Turkish, Azerbaijani, and Turkmen.
  • E. Proto-Uralic language
    Proto-Uralic language is the reconstructed common ancestor of the Uralic language family, from which languages like Finnish, Hungarian, and Estonian are believed to have descended.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
PD Predicate disambiguation gpt-5-mini-2025-08-07
Target predicate: languageFamilyAncestor
Context triple: [Karakalpak language, languageFamilyAncestor, Proto-Turkic language]
  • A. languageFamilyBranchOf
    Indicates that one language family branch is a sub-group or subdivision within a larger language family.
  • B. languageFamilyAssociated
    Indicates that there is an association or connection between a language and a particular language family.
  • C. languageFamily
    Indicates that two or more languages belong to the same genealogical language family or linguistic lineage.
  • D. inLanguageFamily
    Indicates that two languages belong to the same linguistic family or classification.
  • E. derivedFromLanguageFamily chosen
    Indicates that one language originates from, or historically descends from, a particular language family.
  • F. None of above.

Provenance (4 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69ca843057b48190a86730167f5d6b98 completed March 30, 2026, 2:09 p.m.
NER Named-entity recognition batch_69cdd2eef7388190b95ffd02814f2d1f completed April 2, 2026, 2:22 a.m.
NED1 Entity disambiguation (via context triple) batch_69d2cc69a5c88190ab7b108e1aab20ba completed April 5, 2026, 8:56 p.m.
PD Predicate disambiguation batch_69cd4ba1d360819087698d04a53cc87e completed April 1, 2026, 4:45 p.m.
Created at: March 30, 2026, 9:05 p.m.