Triple
T1040516
| Position | Surface form | Disambiguated ID | Type / Status |
|---|---|---|---|
| Subject | Bulgarian Sign Language (written form) |
E22458
|
entity |
| Predicate | basedOn |
P98
|
FINISHED |
| Object | Bulgarian Cyrillic orthography |
E128744
|
NE FINISHED |
How this triple was built (2 steps)
Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.
NER
Named-entity recognition
gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Bulgarian Cyrillic orthography | Statement: [Bulgarian Sign Language (written form), basedOn, Bulgarian Cyrillic orthography]
NED1
Entity disambiguation (via context triple)
gpt-5-mini-2025-08-07
Target entity: Bulgarian Cyrillic orthography Context triple: [Bulgarian Sign Language (written form), basedOn, Bulgarian Cyrillic orthography]
-
A.
Bulgarian Cyrillic alphabet
chosen
The Bulgarian Cyrillic alphabet is the modern variant of the Cyrillic script used for writing the Bulgarian language and adapted for various related linguistic and signed systems.
-
B.
Serbian Cyrillic alphabet
The Serbian Cyrillic alphabet is the standardized Cyrillic script used for writing the Serbian language, consisting of 30 letters in a one-to-one correspondence with Serbian phonemes.
-
C.
Bulgarian language
Bulgarian is a South Slavic language spoken primarily in Bulgaria, notable for being the first Slavic language with a written literary tradition and for its distinctive grammatical features such as the loss of noun cases and the use of suffixed definite articles.
-
D.
Cyrillic script
The Cyrillic script is an alphabetic writing system used for many Slavic and other Eurasian languages, including Russian, Bulgarian, Serbian, and Ukrainian.
-
E.
Serbian Latin alphabet
The Serbian Latin alphabet is the standardized Latin-script writing system used for the Serbian language alongside the Serbian Cyrillic alphabet.
- F. None of above.
- G. Unsure - the case is ambiguous/there is not enough information to decide.
Provenance (3 batches)
The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.
| Step | Stage | Batch ID | Status | When |
|---|---|---|---|---|
| creating | Elicitation | batch_69a493d91478819094cc01fb65564bc1 |
completed | March 1, 2026, 7:30 p.m. |
| NER | Named-entity recognition | batch_69a4b82e4d2c81909ca1264852baf04d |
completed | March 1, 2026, 10:05 p.m. |
| NED1 | Entity disambiguation (via context triple) | batch_69ac5999d30881909bc9e2d8528b1b56 |
completed | March 7, 2026, 5 p.m. |
Created at: March 1, 2026, 7:41 p.m.