Triple

T21296162
Position Surface form Disambiguated ID Type / Status
Subject Angkola people E524927 entity
Predicate language P15 FINISHED
Object Angkola language NE NERFINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Angkola language | Statement: [Angkola people, language, Angkola language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Angkola language
Context triple: [Angkola people, language, Angkola language]
  • A. Angkola language chosen
    Angkola language is an Austronesian language of the Batak group spoken primarily by the Angkola people in North Sumatra, Indonesia.
  • B. Annang language
    The Annang language is a Niger-Congo language spoken primarily by the Annang people of southern Nigeria, closely associated with the Ibibio-Efik linguistic cluster.
  • C. Adang language
    Adang language is a Papuan language spoken by the Adang people on Alor Island in Indonesia’s Alor archipelago.
  • D. Angkuic languages
    The Angkuic languages are a subgroup of Austroasiatic languages spoken primarily in parts of Myanmar, China, and neighboring regions, known for their complex phonology and close relation to other Palaungic languages.
  • E. Anuak language
    The Anuak language is a Nilotic language spoken primarily by the Anuak people of western Ethiopia and eastern South Sudan.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (2 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69e0b517e6748190850d6f6ddf323d69 completed April 16, 2026, 10:08 a.m.
NER Named-entity recognition batch_69e7385858ec8190bdc9c5cdcb8d4507 completed April 21, 2026, 8:42 a.m.
Created at: April 16, 2026, 4:04 p.m.