Triple

T22692132
Position Surface form Disambiguated ID Type / Status
Subject Sanger sequencing E561076 entity
Predicate contributedTo P37 FINISHED
Object Human Genome Project NE NERFINISHED

How this triple was built (2 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Human Genome Project | Statement: [Sanger sequencing, contributedTo, Human Genome Project]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Human Genome Project
Context triple: [Sanger sequencing, contributedTo, Human Genome Project]
  • A. Human Genome Project chosen
    The Human Genome Project was an international scientific research initiative that successfully mapped and sequenced the entire human DNA genome, revolutionizing genetics and biomedical research.
  • B. Personal Genome Project
    The Personal Genome Project is a pioneering open-science initiative that publicly shares the genomic and health data of volunteers to advance research and understanding of human genetics.
  • C. Edico Genome
    Edico Genome was a biotechnology company specializing in high-speed genomic data analysis through its DRAGEN bio-IT platform and FPGA-based acceleration technology.
  • D. ENCODE project
    The ENCODE project is a large-scale collaborative research initiative aimed at identifying and characterizing all functional elements in the human genome.
  • E. Human Genome Sequencing Center
    The Human Genome Sequencing Center is a major genomics research facility specializing in large-scale DNA sequencing and analysis.
  • F. None of above.
  • G. Unsure - the case is ambiguous/there is not enough information to decide.

Provenance (2 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69e2454d71b48190a1f80af9f82b6fcf completed April 17, 2026, 2:35 p.m.
NER Named-entity recognition batch_69f1789ba0148190891781d05ec64f3c completed April 29, 2026, 3:18 a.m.
Created at: April 17, 2026, 3:13 p.m.