Triple

T15819708
Position Surface form Disambiguated ID Type / Status
Subject Sierra Totonac E383572 entity
Predicate languageFamily P1047 FINISHED
Object Totonacan E241837 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Totonacan | Statement: [Sierra Totonac, languageFamily, Totonacan]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Totonacan
Context triple: [Sierra Totonac, languageFamily, Totonacan]
  • A. Totonac
    Totonac is an indigenous language family of eastern Mexico, traditionally spoken by the Totonac people primarily in the states of Veracruz and Puebla.
  • B. Huastec
    Huastec is a Mayan language spoken by the Huastec people primarily in northeastern Mexico, especially in parts of Veracruz and neighboring states.
  • C. Popoloca
    Popoloca is an indigenous language of central Mexico belonging to the Oto-Manguean family and spoken by the Popoloca people of Puebla.
  • D. Huasteca Nahuatl
    Huasteca Nahuatl is a modern variety of the Nahuatl language spoken by the Huastec Nahua people in northeastern Mexico.
  • E. Totonac languages chosen
    Totonac languages are an indigenous language family of eastern Mexico spoken primarily by the Totonac people in the states of Veracruz, Puebla, and Hidalgo.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69d86da2858c819090cc8481e7207b6e completed April 10, 2026, 3:25 a.m.
NER Named-entity recognition batch_69e0c4a6e6748190acb0791bd465587f completed April 16, 2026, 11:14 a.m.
NED1 Entity disambiguation (via context triple) batch_69ffa939b4608190ac411b2f2f61e19e completed May 9, 2026, 9:38 p.m.
Created at: April 10, 2026, 4:49 a.m.