Triple

T1556358
Position Surface form Disambiguated ID Type / Status
Subject Central Banda languages E33212 entity
Predicate glottologName P6521 FINISHED
Object Central Banda
Central Banda is a group of closely related Ubangian languages spoken primarily in the Central African Republic and neighboring regions.
E178208 NE FINISHED

How this triple was built (4 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Central Banda | Statement: [Central Banda languages, glottologName, Central Banda]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Central Banda
Context triple: [Central Banda languages, glottologName, Central Banda]
  • A. Tamba
    Tamba is a city located in Hyogo Prefecture, Japan, known for its rural landscapes, traditional pottery, and historical sites.
  • B. Karanga
    Karanga is a major dialect of the Shona language spoken primarily in southern Zimbabwe, known for its distinct phonological and lexical features.
  • C. Kuanua
    Kuanua is an Austronesian language spoken primarily by the Tolai people of East New Britain in Papua New Guinea.
  • D. Waingapu
    Waingapu is the main urban and economic center of the Indonesian island of Sumba, serving as a key hub for transportation and regional administration.
  • E. Madura
    Madura is an island off the northeastern coast of Java in Indonesia, known for its distinct Madurese culture and traditional bull races.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NEDg Description generation gpt-5.1
Instruction
Generate a one-sentence description of the target entity. 
You are given a context triple in the form (subject, predicate, object), where the object is the target entity. 
# Instructions
Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. 
Avoid repeating the information from the triple, unless really essential.
# Response Format
Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: Central Banda
Triple: [Central Banda languages, glottologName, Central Banda]
Generated description
Central Banda is a group of closely related Ubangian languages spoken primarily in the Central African Republic and neighboring regions.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: Central Banda
Target entity description: Central Banda is a group of closely related Ubangian languages spoken primarily in the Central African Republic and neighboring regions.
  • A. Tamba
    Tamba is a city located in Hyogo Prefecture, Japan, known for its rural landscapes, traditional pottery, and historical sites.
  • B. Karanga
    Karanga is a major dialect of the Shona language spoken primarily in southern Zimbabwe, known for its distinct phonological and lexical features.
  • C. Kuanua
    Kuanua is an Austronesian language spoken primarily by the Tolai people of East New Britain in Papua New Guinea.
  • D. Waingapu
    Waingapu is the main urban and economic center of the Indonesian island of Sumba, serving as a key hub for transportation and regional administration.
  • E. Madura
    Madura is an island off the northeastern coast of Java in Indonesia, known for its distinct Madurese culture and traditional bull races.
  • F. None of above. chosen

Provenance (5 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69a885ef9cf48190b0af0f5ce3d02231 completed March 4, 2026, 7:20 p.m.
NER Named-entity recognition batch_69a908704d208190937af41c6454df4e completed March 5, 2026, 4:37 a.m.
NED1 Entity disambiguation (via context triple) batch_69ad370e10248190b060a0209b979ef9 completed March 8, 2026, 8:45 a.m.
NEDg Description generation batch_69ad3b03780881909ca372552cdc4ff3 completed March 8, 2026, 9:01 a.m.
NED2 Entity disambiguation (via description) batch_69ad3b6bbc9c8190964ad173cfca5698 completed March 8, 2026, 9:03 a.m.
Created at: March 4, 2026, 7:27 p.m.