Triple

T2593651
Position Surface form Disambiguated ID Type / Status
Subject Southeastern Woodlands E58180 entity
Predicate includesPeople P17131 FINISHED
Object Taensa
The Taensa are a Native American people historically associated with the lower Mississippi Valley region of what is now the southeastern United States.
E280860 NE FINISHED

How this triple was built (4 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Taensa | Statement: [Southeastern Woodlands, includesPeople, Taensa]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Taensa
Context triple: [Southeastern Woodlands, includesPeople, Taensa]
  • A. Taalunie
    Taalunie is an international institution that coordinates and promotes the Dutch language and literature across Dutch-speaking regions.
  • B. Lisnaskea
    Lisnaskea is a small market town in County Fermanagh, Northern Ireland, known for serving as a local commercial and community hub for the surrounding rural area.
  • C. Riasti
    Riasti is a regional dialect of the Saraiki language spoken primarily in parts of southern Punjab, Pakistan.
  • D. Tuineje
    Tuineje is a coastal municipality on the island of Fuerteventura in Spain’s Canary Islands, known for its rural landscapes, beaches, and traditional Canarian culture.
  • E. Tumpa
    Tumpa is a song featured on the album "Legend of the Sun Virgin."
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NEDg Description generation gpt-5.1
Instruction
Generate a one-sentence description of the target entity. 
You are given a context triple in the form (subject, predicate, object), where the object is the target entity. 
# Instructions
Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. 
Avoid repeating the information from the triple, unless really essential.
# Response Format
Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: Taensa
Triple: [Southeastern Woodlands, includesPeople, Taensa]
Generated description
The Taensa are a Native American people historically associated with the lower Mississippi Valley region of what is now the southeastern United States.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: Taensa
Target entity description: The Taensa are a Native American people historically associated with the lower Mississippi Valley region of what is now the southeastern United States.
  • A. Taalunie
    Taalunie is an international institution that coordinates and promotes the Dutch language and literature across Dutch-speaking regions.
  • B. Lisnaskea
    Lisnaskea is a small market town in County Fermanagh, Northern Ireland, known for serving as a local commercial and community hub for the surrounding rural area.
  • C. Riasti
    Riasti is a regional dialect of the Saraiki language spoken primarily in parts of southern Punjab, Pakistan.
  • D. Tuineje
    Tuineje is a coastal municipality on the island of Fuerteventura in Spain’s Canary Islands, known for its rural landscapes, beaches, and traditional Canarian culture.
  • E. Tumpa
    Tumpa is a song featured on the album "Legend of the Sun Virgin."
  • F. None of above. chosen

Provenance (5 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69ab4ac019c8819094add11c46706e32 completed March 6, 2026, 9:44 p.m.
NER Named-entity recognition batch_69abd83299708190993b79daaffcc9a1 completed March 7, 2026, 7:48 a.m.
NED1 Entity disambiguation (via context triple) batch_69af83bee4908190b5e446ddbf4e8889 completed March 10, 2026, 2:36 a.m.
NEDg Description generation batch_69af8434f61c81909bffb3f06acb733b completed March 10, 2026, 2:38 a.m.
NED2 Entity disambiguation (via description) batch_69af84b260b881909bbd3d2825f9dea7 completed March 10, 2026, 2:40 a.m.
Created at: March 6, 2026, 9:49 p.m.