Triple

T3185951
Position Surface form Disambiguated ID Type / Status
Subject South Asian ethnic groups E66698 entity
Predicate includesMajorLanguageFamily P46508 FINISHED
Object Indo-Aryan languages LITERAL FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Indo-Aryan languages | Statement: [South Asian ethnic groups, includesMajorLanguageFamily, Indo-Aryan languages]
PD Predicate disambiguation gpt-5-mini-2025-08-07
Target predicate: includesMajorLanguageFamily
Context triple: [South Asian ethnic groups, includesMajorLanguageFamily, Indo-Aryan languages]
  • A. isOneOfLargestLanguageFamilies
    Indicates that the language family ranks among the largest in terms of number of languages, speakers, or geographic spread.
  • B. areOfficialLanguageFamilyOf
    Indicates that one or more language families hold official status within, or are formally recognized as official for, a particular entity (such as a country or region).
  • C. languageFamily
    Indicates that two or more languages belong to the same genealogical language family or linguistic lineage.
  • D. languageFamilyOf
    Indicates that one entity is the language family to which the other entity (a specific language) belongs.
  • E. hasWritingSystemForMajorLanguage
    Indicates that there exists a writing system used to represent a major language associated with the given entity.
  • F. None of above. chosen

Provenance (4 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69ad8587c1bc8190a2595f2c22ee1001 completed March 8, 2026, 2:19 p.m.
NER Named-entity recognition batch_69ada6c32ce88190a231be18d38ec5ba completed March 8, 2026, 4:41 p.m.
PD Predicate disambiguation batch_69ad9e04290481909092ddfbe6fdaabc completed March 8, 2026, 4:04 p.m.
PDg Predicate description generation batch_69ada148e9108190b363dd0f1a94ac8e completed March 8, 2026, 4:18 p.m.
Created at: March 8, 2026, 3:06 p.m.