Triple

T6474910
Position Surface form Disambiguated ID Type / Status
Subject Northern Bantoid E146047 entity
Predicate hasSubgroup P747 FINISHED
Object Tikar language
The Tikar language is a Bantoid language spoken primarily by the Tikar people of central Cameroon.
E596043 NE FINISHED

How this triple was built (4 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Tikar language | Statement: [Northern Bantoid, hasSubgroup, Tikar language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Tikar language
Context triple: [Northern Bantoid, hasSubgroup, Tikar language]
  • A. Tigak language
    The Tigak language is an Austronesian language spoken by the Tigak people of New Ireland Province in Papua New Guinea.
  • B. Tiv language
    Tiv language is a Southern Bantoid language of the Benue–Congo family spoken predominantly by the Tiv people of central Nigeria and parts of Cameroon.
  • C. Tat language
    Tat language is an endangered Southwestern Iranian language spoken primarily by the Tat people of Azerbaijan and neighboring regions, distinct from but related to Judeo-Tat.
  • D. Tamyen language
    The Tamyen language is an extinct Ohlone (Costanoan) Native American language once spoken in the Santa Clara Valley region of California.
  • E. Towa language
    Towa is a Native American language spoken by the Towa (Jemez) people of New Mexico and is part of the Puebloan language family.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NEDg Description generation gpt-5.1
Instruction
Generate a one-sentence description of the target entity. 
You are given a context triple in the form (subject, predicate, object), where the object is the target entity. 
# Instructions
Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. 
Avoid repeating the information from the triple, unless really essential.
# Response Format
Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: Tikar language
Triple: [Northern Bantoid, hasSubgroup, Tikar language]
Generated description
The Tikar language is a Bantoid language spoken primarily by the Tikar people of central Cameroon.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: Tikar language
Target entity description: The Tikar language is a Bantoid language spoken primarily by the Tikar people of central Cameroon.
  • A. Tigak language
    The Tigak language is an Austronesian language spoken by the Tigak people of New Ireland Province in Papua New Guinea.
  • B. Tiv language
    Tiv language is a Southern Bantoid language of the Benue–Congo family spoken predominantly by the Tiv people of central Nigeria and parts of Cameroon.
  • C. Tat language
    Tat language is an endangered Southwestern Iranian language spoken primarily by the Tat people of Azerbaijan and neighboring regions, distinct from but related to Judeo-Tat.
  • D. Tamyen language
    The Tamyen language is an extinct Ohlone (Costanoan) Native American language once spoken in the Santa Clara Valley region of California.
  • E. Towa language
    Towa is a Native American language spoken by the Towa (Jemez) people of New Mexico and is part of the Puebloan language family.
  • F. None of above. chosen

Provenance (5 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69c008fec7408190af7b146dc63d9750 completed March 22, 2026, 3:21 p.m.
NER Named-entity recognition batch_69c06a341360819082f2b5496a1a68b0 completed March 22, 2026, 10:16 p.m.
NED1 Entity disambiguation (via context triple) batch_69c653a595b881909e5d3cb781ad5ad4 completed March 27, 2026, 9:53 a.m.
NEDg Description generation batch_69c6553c17bc81908719ecc7db9e3960 completed March 27, 2026, 10 a.m.
NED2 Entity disambiguation (via description) batch_69c655f4ee5c81909620e732b72ee694 completed March 27, 2026, 10:03 a.m.
Created at: March 22, 2026, 4:50 p.m.