Triple

T3936138
Position Surface form Disambiguated ID Type / Status
Subject Pucikwar E90916 entity
Predicate classification P87 FINISHED
Object Andamanese language family E80892 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Andamanese language family | Statement: [Pucikwar, classification, Andamanese language family]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Andamanese language family
Context triple: [Pucikwar, classification, Andamanese language family]
  • A. Great Andamanese languages chosen
    The Great Andamanese languages are a small, nearly extinct group of indigenous languages once spoken by the Great Andamanese peoples of the Andaman Islands in the Bay of Bengal.
  • B. Nicobarese languages
    The Nicobarese languages are a group of Austroasiatic languages spoken by the indigenous Nicobarese people of India’s Nicobar Islands in the eastern Indian Ocean.
  • C. Penutian languages
    Penutian languages are a proposed family of Native American languages spoken primarily in the western United States, noted for their controversial genetic relationships and inclusion of several distinct regional language groups.
  • D. Moru–Madi languages
    The Moru–Madi languages are a subgroup of related Central Sudanic languages spoken primarily in South Sudan, Uganda, and the Democratic Republic of the Congo.
  • E. Timor–Babar languages
    The Timor–Babar languages are a subgroup of Austronesian languages spoken primarily on Timor and nearby islands in eastern Indonesia, noted for their complex phonologies and diverse grammatical structures.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69aed95f26e0819094b0e71974543a19 completed March 9, 2026, 2:29 p.m.
NER Named-entity recognition batch_69aeedcd29148190a98e4549c9ed8888 completed March 9, 2026, 3:57 p.m.
NED1 Entity disambiguation (via context triple) batch_69b5338d9a7c8190ac5960eab5ae2ada completed March 14, 2026, 10:08 a.m.
Created at: March 9, 2026, 3:23 p.m.