Triple

T14395773
Position Surface form Disambiguated ID Type / Status
Subject Shtokavian dialect E356943 entity
Predicate usedInStandardFormOf P23250 FINISHED
Object Standard Croatian E29128 NE FINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Standard Croatian | Statement: [Shtokavian dialect, usedInStandardFormOf, Standard Croatian]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Standard Croatian
Context triple: [Shtokavian dialect, usedInStandardFormOf, Standard Croatian]
  • A. Croatian chosen
    Croatian is a South Slavic language primarily spoken in Croatia and recognized as one of the official languages of the European Union.
  • B. Serbo-Croatian
    Serbo-Croatian is a South Slavic language historically spoken across the former Yugoslavia, encompassing the standardized varieties now known as Serbian, Croatian, Bosnian, and Montenegrin.
  • C. Department of Standard Croatian Language
    The Department of Standard Croatian Language is a specialized unit that researches, codifies, and advises on the norms and usage of the contemporary Croatian standard language.
  • D. Shtokavian dialect
    The Shtokavian dialect is the South Slavic dialectal base from which the standard forms of Bosnian, Croatian, Serbian, and Montenegrin developed.
  • E. Croatian Latin alphabet
    The Croatian Latin alphabet is the standardized Latin-based writing system used for the Croatian language, employing a set of letters with specific diacritics to represent its phonemic inventory.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69d827927c988190ad98bb0360981783 completed April 9, 2026, 10:26 p.m.
NER Named-entity recognition batch_69de90826f908190b3969af9b7cf922f completed April 14, 2026, 7:07 p.m.
NED1 Entity disambiguation (via context triple) batch_69fd551cbdb08190a9ea53e607f2555b completed May 8, 2026, 3:14 a.m.
Created at: April 10, 2026, 1:17 a.m.