Triple

T18368582
Position Surface form Disambiguated ID Type / Status
Subject Kathlamet language E440117 entity
Predicate partOf P40 FINISHED
Object Plateau linguistic area NE NERFINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Plateau linguistic area | Statement: [Kathlamet language, partOf, Plateau linguistic area]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Plateau linguistic area
Context triple: [Kathlamet language, partOf, Plateau linguistic area]
  • A. Plateau linguistic area chosen
    The Plateau linguistic area is a region of the North American Plateau where diverse Indigenous languages, often from different families, share common structural features due to long-term contact and interaction.
  • B. Plains linguistic area
    The Plains linguistic area is a region of North America where diverse Indigenous languages, including the Caddoan family, have converged and shared structural features through long-term contact.
  • C. Yuman linguistic area
    The Yuman linguistic area is a region of the southwestern United States and northwestern Mexico characterized by a group of closely related Yuman languages spoken by several Indigenous peoples.
  • D. Derajat linguistic area
    The Derajat linguistic area is a dialect region in western Punjab and adjacent areas where closely related varieties of Punjabi and Saraiki, including Derawali, are spoken.
  • E. Andean linguistic area
    The Andean linguistic area is a region of the central Andes where diverse languages have converged to share common structural features through long-term contact and interaction.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (2 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69d8b918221c8190a9f7b563d64ac677 completed April 10, 2026, 8:47 a.m.
NER Named-entity recognition batch_69e51750d3dc8190b153046c1171ee2b completed April 19, 2026, 5:56 p.m.
Created at: April 10, 2026, 10:38 a.m.