Triple
T14395773
| Position | Surface form | Disambiguated ID | Type / Status |
|---|---|---|---|
| Subject | Shtokavian dialect |
E356943
|
entity |
| Predicate | usedInStandardFormOf |
P23250
|
FINISHED |
| Object | Standard Croatian |
E29128
|
NE FINISHED |
How this triple was built (2 steps)
Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.
NER
Named-entity recognition
gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Standard Croatian | Statement: [Shtokavian dialect, usedInStandardFormOf, Standard Croatian]
NED1
Entity disambiguation (via context triple)
gpt-5-mini-2025-08-07
Target entity: Standard Croatian Context triple: [Shtokavian dialect, usedInStandardFormOf, Standard Croatian]
-
A.
Croatian
chosen
Croatian is a South Slavic language primarily spoken in Croatia and recognized as one of the official languages of the European Union.
-
B.
Serbo-Croatian
Serbo-Croatian is a South Slavic language historically spoken across the former Yugoslavia, encompassing the standardized varieties now known as Serbian, Croatian, Bosnian, and Montenegrin.
-
C.
Department of Standard Croatian Language
The Department of Standard Croatian Language is a specialized unit that researches, codifies, and advises on the norms and usage of the contemporary Croatian standard language.
-
D.
Shtokavian dialect
The Shtokavian dialect is the South Slavic dialectal base from which the standard forms of Bosnian, Croatian, Serbian, and Montenegrin developed.
-
E.
Croatian Latin alphabet
The Croatian Latin alphabet is the standardized Latin-based writing system used for the Croatian language, employing a set of letters with specific diacritics to represent its phonemic inventory.
- F. None of above.
- G. Unsure - the case is ambiguous/there is not enough information to decide.
Provenance (3 batches)
The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.
| Step | Stage | Batch ID | Status | When |
|---|---|---|---|---|
| creating | Elicitation | batch_69d827927c988190ad98bb0360981783 |
completed | April 9, 2026, 10:26 p.m. |
| NER | Named-entity recognition | batch_69de90826f908190b3969af9b7cf922f |
completed | April 14, 2026, 7:07 p.m. |
| NED1 | Entity disambiguation (via context triple) | batch_69fd551cbdb08190a9ea53e607f2555b |
completed | May 8, 2026, 3:14 a.m. |
Created at: April 10, 2026, 1:17 a.m.