Triple

T27039620
Position Surface form Disambiguated ID Type / Status
Subject 马佳 E684450 entity
Predicate romanizationHanyuPinyin P41219 FINISHED
Object Mǎjiā NE NERFINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Mǎjiā | Statement: [马佳, romanizationHanyuPinyin, Mǎjiā]
PD Predicate disambiguation gpt-5-mini-2025-08-07
Target predicate: romanizationHanyuPinyin
Context triple: [马佳, romanizationHanyuPinyin, Mǎjiā]
  • A. ChinesePinyin chosen
    Indicates that one entity is the Chinese pinyin (romanized phonetic transcription) representation of another entity.
  • B. romanizationFrom
    Indicates that one entity is a romanized representation derived from the script or writing system of another entity.
  • C. romanizationType
    Indicates the specific system or method used to convert text from one writing system into its Roman (Latin) alphabet representation.
  • D. romanizationVariantOf
    Indicates that one written form is a different romanized representation of the same underlying word or expression as another.
  • E. mandarinReadingBopomofo
    Indicates the Bopomofo (Zhuyin) phonetic transcription used to represent the Mandarin pronunciation of a given expression or character.
  • F. None of above.

Provenance (3 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69ef148193c48190bb1a0cfae6a407c4 completed April 27, 2026, 7:47 a.m.
NER Named-entity recognition batch_69f6383625cc8190aa223d8ef655743c completed May 2, 2026, 5:45 p.m.
PD Predicate disambiguation batch_69f63709e4848190b5cf322e06b23fb6 completed May 2, 2026, 5:40 p.m.
Created at: April 27, 2026, 8:04 a.m.