This study explores whether human-written descriptions in Reactome can be used to infer the experts'defined global hierarchical structure and indicates that the global hierarchical structure of pathways can be inferred by experts textual metadata.
Abstract
Biological knowledgebases like Reactome provide high-quality pathways that include biological elements'relationships and textual descriptions (metadata). The quality of such pathways is granted by manual curation, that presents, however, significant scalability challenges. Lately, numerous NLP tools have been proposed to cope with this issue, leveraging textual information to automatically expand biological knowledgebases. However, little exploration has been done so far to assess whether relationships among textual descriptions mirror higher order biological relationships. This study explores whether human-written descriptions in Reactome can be used to infer the experts'defined global hierarchical structure. To test this, we extracted from Reactome the Homo Sapiens hierarchy of pathways and their reactions (Reactome Hierarchy), and used textual metadata to reconstruct a Semantic Hierarchy, combining a sentence transformer model (SPECTER2) with a modified agglomerative nesting algorithm and a graph reconstruction algorithm. Quantitative (Laplacian Spectral Distance and Bootstrapping) and qualitative (global topological metrics) analyses confirm our hypothesis and indicate that the global hierarchical structure of pathways can be inferred by experts textual metadata.
LLMBDC (Large Language Model for Biological Domains Oriented Clustering of Gene Ontology) provides a scalable, reproducible, and interpretable route to context-aware, system-level interpretation of GO enrichment results while preserving biological specificity.
AVA is introduced, a systematic framework for evaluating whether embeddings distinguish logic-sensitive relational semantics in ontologies and knowledge graphs, and reveals a persistent gap between linguistic representation learning and ontology-level discrimination, challenging the assumption that strong NLP benchmark...
Hamed Babaei Giglou, Jennifer D'Souza, S. Auer· 0 citations
A layered reliability framework is defined in which graph-based inference addresses knowledge incompleteness, retrieval-augmented prompt control mitigates instability, and ontology grounding reduces semantic ambiguity, providing a foundation for more reliable biomedical AI systems.
This work targets a KG for Sophocles’ Antigone that supports two coupled uses: structured retrieval, through integrity and competency questions expressed in SPARQL over dramatic structure and interpretive annotations; and interactive exploration, through a lightweight read client that navigates lines across languages,...
An ontology-guided framework that integrates a Knowledge Graph, an Ontology-Informed Retrieval Classifier, and a Large Language Model for interpretable mental health detection from social media text demonstrates that the KG–ORC cross-validation gate measurably improves predictive reliability over single component basel...
Amina Tahir, Ghulam Mustafa, Muhammad Tanvir Afzal et al.· Social Network Analysis and...· 0 citations
Findings show that model size alone is an insufficient selection criterion for OL and provide empirical guidance for reproducible LLM-assisted ontology engineering and indicate that architecture and model lineage can outweigh nominal parameter count.
Hamed Babaei Giglou, S. Auer, Jennifer D'Souza· 0 citations
With $2.1 million funding from Google.org, the open-source Public Transit Intelligence Hub will unify public transit monitoring, operations, and passenger communication.
Professor Sherry Turkle’s new book, “Artificial Intimacy,” offers a withering critique of chatbots and the antisocial dynamics she believes they encourage.
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.