Triple

T20160781
Position Surface form Disambiguated ID Type / Status
Subject Kayastha E491696 entity
Predicate hasSubgroup P747 FINISHED
Object Bengali Kayastha NE NERFINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Bengali Kayastha | Statement: [Kayastha, hasSubgroup, Bengali Kayastha]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Bengali Kayastha
Context triple: [Kayastha, hasSubgroup, Bengali Kayastha]
  • A. Kayastha chosen
    Kayastha is a prominent South Asian community traditionally associated with literacy, administration, and record-keeping roles, particularly in North India.
  • B. Ganguli
    Ganguli is a Bengali surname commonly associated with Indian families, notably featured in Jhumpa Lahiri’s novel "The Namesake" through the character Gogol Ganguli.
  • C. Maithil Brahmin
    Maithil Brahmin are a Hindu Brahmin community from the Mithila region of India and Nepal, traditionally known for their scholarship, ritual expertise, and preservation of Maithili culture.
  • D. Dogra Brahmin
    Dogra Brahmin is a Hindu Brahmin community from the Jammu region of India, known for its distinct Dogri language, customs, and religious traditions.
  • E. Bengali script
    Bengali script is an abugida used across eastern South Asia to write languages such as Bengali and Assamese, derived from the ancient Brahmi script.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (2 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69da6266c6888190bc1a3ecf24814d34 completed April 11, 2026, 3:01 p.m.
NER Named-entity recognition batch_69e667e43940819080f6a0b7331aaab0 completed April 20, 2026, 5:52 p.m.
Created at: April 11, 2026, 11:34 p.m.