Skip to content
Open access

HECTOR: A Web-Based Tool for Automated BRCA1/BRCA2 Variant Classification Under the ClinGen ENIGMA Specifications

Jul 2026 · medRxiv · 0 citations
Medicine

TL;DR

HECTOR provides a faithful, transparent implementation of the ENIGMA VCEP v1.2 specifications for BRCA1 and BRCA2, enabling rapid, standardized, and reproducible application of gene-specific variant classification guidelines while reducing the burden of manual curation.

Abstract

Background: The ClinGen ENIGMA BRCA1/BRCA2 Variant Curation Expert Panel (VCEP) has adapted the ACMG/AMP framework into gene-specific specifications. However, applying these specifications manually remains labour-intensive and prone to inconsistency, requiring integration of population, computational, functional, and clinical evidence through gene-specific decision trees and a points-based classification system. Methods: We developed HECTOR, a free web-based tool that implements the complete ENIGMA VCEP v1.2 specifications for BRCA1 and BRCA2. HECTOR automatically populates all evidence codes derivable from public data, routes curator-dependent evidence to a manual input layer and returns a transparent five-tier classification with code-level evidence. We validated HECTOR against two independent reference datasets: the 143-variant ENIGMA Evidence Repository, used as a clinical-grade reference standard, and 134 manually curated in-house variants of uncertain significance. HECTOR was then applied to the complete ClinVar BRCA1/BRCA2 catalogue (n = 34,077). Results: At the criterion level, HECTOR exactly reproduced 326 of 413 VCEP-assigned criteria (78.9%). The discordance arising predominantly from curator-dependent evidence rather than implementation errors whereas computationally accessible criteria showed perfect concordance. Across ClinVar, HECTOR classified 33,913 variants (99.5%). Agreement with definitive ClinVar classifications was 96.7% for pathogenic variants overall. Among variants for which HECTOR generated a definitive classification, directional concordance reached 99.7% for pathogenic and 99.9% for benign variants. HECTOR also resolved a substantial proportion of variants classified as uncertain (67.3%) or conflicting (88.7%), predominantly toward benign classifications. Conclusions: HECTOR provides a faithful, transparent implementation of the ENIGMA VCEP v1.2 specifications for BRCA1 and BRCA2, enabling rapid, standardized, and reproducible application of gene-specific variant classification guidelines while reducing the burden of manual curation.

Read PDF

Similar papers

Open access Aug 2026

megaMine: a scalable, rule-based framework for mining gene-cancer-drug evidence from biomedical literature

The rapid expansion of the oncology literature has outpaced manual curation of clinically relevant gene-cancer-drug associations and oncogenic driver evidence. Existing automated approaches often lack transparency or are difficult to scale across heterogeneous data sources. To address this gap, we developed megaMine, a...

Muhammad Junaid, K. Prazanowska, Ha-Eun Jeong et al. · 0 citations
Review Open access Aug 2026

OncoGenRAG: Evidence-Grounded Retrieval and BioBERT Classification for Precision Oncology Variant Interpretation

OncoGenRAG is a research framework that combines a parameter-efficiently fine-tuned BioBERT classifier with an entity-aware retrieval system over a curated, multi-source oncology knowledge base and provides a transparent design for evidence retrieval and abstention.

Amaan Arif, J. V. dos Santos · 0 citations
Open access Aug 2026

Variantscape: Large Language Model-Driven Mining of Biomedical Literature for Clinical Interpretation of Cancer Variants

Variantscape has the potential to support MTB workflows and translational research by rapidly revealing signals from underlying abstracts, and offers a practical resource for accelerating discovery and supporting precision oncology research and translation.

M. Wosny, A. Blindu, M. Boesch et al. · 0 citations
Open access Sep 2026

HeartVar: An LLM-Assisted Tool for Clinical Classification of Variants in Cardiovascular Disease Cohorts

Manual clinical DNA variant classification is the bottleneck of every clinical and research rare disease workflow. The process typically requires a curator to assemble evidence from numerous databases, weigh 28 criteria, reconcile competing evidence, and produce a defensible case for the final classification. Additiona...

Jamie-Lee M. Thompson, Debjani Das, Sally L. Dunwoodie et al. · 0 citations
Review Open access Aug 2026

A machine learning framework for predictive interpretation of variants of uncertain significance in hereditary cancer

This reproducible pipeline provides a clinically grounded computational approach to VUS triaging in precision oncology, with external validation supporting its generalizability to independent hereditary cancer gene datasets.

Nayeema Nizamuddin, Soham Biswas, Akshaykumar Zawar et al. · 0 citations
Open access Sep 2026

Bridging the data latency gap: automated extraction of genomic biomarkers from unstructured clinical documents to support real-world oncology data

Real-world oncology data are essential for clinical research and precision cancer care. However, genomic biomarkers are often embedded in scanned, unstructured clinical documents requiring manual abstraction before becoming available in cancer registries, delaying real-world evidence generation. This study evaluated an...

Qian-Yun Luo, Rui Zhang, Nikitha Vobugari et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.