Triple

T1431869
Position Surface form Disambiguated ID Type / Status
Subject Central–Eastern Oceanic languages E30464 entity
Predicate includesLanguage P2177 FINISHED
Object Mortlockese language E146933 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Mortlockese language | Statement: [Central–Eastern Oceanic languages, includesLanguage, Mortlockese language]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Mortlockese language
Context triple: [Central–Eastern Oceanic languages, includesLanguage, Mortlockese language]
  • A. Mortlockese language chosen
    Mortlockese is an Austronesian language of the Chuukic branch spoken primarily in the Mortlock Islands of Chuuk State in the Federated States of Micronesia.
  • B. Eonavian language
    Eonavian language is a Romance language variety spoken in the western coastal region of Asturias, Spain, sharing features with both Galician and Asturian.
  • C. Avikam language
    The Avikam language is a Kwa language spoken by the Avikam people of southern Côte d'Ivoire.
  • D. Tobian language
    The Tobian language is a Micronesian language spoken primarily on Tobi Island in Palau, known for its small speaker population and close relation to other Carolinean languages.
  • E. Hadza language
    The Hadza language is an isolate spoken by the Hadza people of northern Tanzania, notable for its extensive use of click consonants and lack of clear genetic affiliation to other language families.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69a498fc69ec8190b61722bd4b67c4d2 completed March 1, 2026, 7:52 p.m.
NER Named-entity recognition batch_69a4c4ddbe208190a68cb000a6970d17 completed March 1, 2026, 10:59 p.m.
NED1 Entity disambiguation (via context triple) batch_69ad016bf2608190a675cbd42e474082 completed March 8, 2026, 4:56 a.m.
Created at: March 1, 2026, 8 p.m.