Triple

T6970498
Position Surface form Disambiguated ID Type / Status
Subject Indian Institute of Technology Guwahati E161585 entity
Predicate hasResearchCenter P40 FINISHED
Object Centre for Linguistic Science and Technology
The Centre for Linguistic Science and Technology is a research hub at IIT Guwahati focused on language science, computational linguistics, and language technology development.
E633222 NE FINISHED

How this triple was built (4 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Centre for Linguistic Science and Technology | Statement: [Indian Institute of Technology Guwahati, hasResearchCenter, Centre for Linguistic Science and Technology]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Centre for Linguistic Science and Technology
Context triple: [Indian Institute of Technology Guwahati, hasResearchCenter, Centre for Linguistic Science and Technology]
  • A. Center for the Study of Language and Information
    The Center for the Study of Language and Information is a Stanford University-based interdisciplinary research center and academic publisher focused on linguistics, logic, cognitive science, and computational approaches to language and information.
  • B. Leiden University Centre for Linguistics
    Leiden University Centre for Linguistics is a research and teaching institute at Leiden University specializing in theoretical, descriptive, and applied linguistics across a wide range of languages and language families.
  • C. Language Technology Unit
    The Language Technology Unit is a specialized division focused on developing and promoting language technologies and computational tools for Indian languages under the Central Institute of Indian Languages.
  • D. Department of Language and Computer Science
    The Department of Language and Computer Science is a unit within the Institute of Croatian Language and Linguistics that focuses on the intersection of linguistics and computational methods for language analysis and processing.
  • E. Language Technologies Institute, Carnegie Mellon University
    The Language Technologies Institute at Carnegie Mellon University is a leading research and education center focused on areas such as natural language processing, machine learning for language, speech recognition, and related AI-driven language technologies.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NEDg Description generation gpt-5.1
Instruction
Generate a one-sentence description of the target entity. 
You are given a context triple in the form (subject, predicate, object), where the object is the target entity. 
# Instructions
Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. 
Avoid repeating the information from the triple, unless really essential.
# Response Format
Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: Centre for Linguistic Science and Technology
Triple: [Indian Institute of Technology Guwahati, hasResearchCenter, Centre for Linguistic Science and Technology]
Generated description
The Centre for Linguistic Science and Technology is a research hub at IIT Guwahati focused on language science, computational linguistics, and language technology development.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: Centre for Linguistic Science and Technology
Target entity description: The Centre for Linguistic Science and Technology is a research hub at IIT Guwahati focused on language science, computational linguistics, and language technology development.
  • A. Center for the Study of Language and Information
    The Center for the Study of Language and Information is a Stanford University-based interdisciplinary research center and academic publisher focused on linguistics, logic, cognitive science, and computational approaches to language and information.
  • B. Leiden University Centre for Linguistics
    Leiden University Centre for Linguistics is a research and teaching institute at Leiden University specializing in theoretical, descriptive, and applied linguistics across a wide range of languages and language families.
  • C. Language Technology Unit
    The Language Technology Unit is a specialized division focused on developing and promoting language technologies and computational tools for Indian languages under the Central Institute of Indian Languages.
  • D. Department of Language and Computer Science
    The Department of Language and Computer Science is a unit within the Institute of Croatian Language and Linguistics that focuses on the intersection of linguistics and computational methods for language analysis and processing.
  • E. Language Technologies Institute, Carnegie Mellon University
    The Language Technologies Institute at Carnegie Mellon University is a leading research and education center focused on areas such as natural language processing, machine learning for language, speech recognition, and related AI-driven language technologies.
  • F. None of above. chosen

Provenance (5 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69c68854a0d88190bc0bf82263f1afce completed March 27, 2026, 1:38 p.m.
NER Named-entity recognition batch_69c6db1649288190a52c7dab57b3c7dc completed March 27, 2026, 7:31 p.m.
NED1 Entity disambiguation (via context triple) batch_69c7619ebab88190916e3d68068ed71d completed March 28, 2026, 5:05 a.m.
NEDg Description generation batch_69c7630440508190a66f218fd912d732 completed March 28, 2026, 5:11 a.m.
NED2 Entity disambiguation (via description) batch_69c7639747b88190b3429817d53c5703 completed March 28, 2026, 5:13 a.m.
Created at: March 27, 2026, 2:30 p.m.