Skip to content
Preprint

PatTree: a novel approach for automated creation of multimodal, graph-based patient representations for medical classification tasks

Aug 2026 · 0 citations · 73 references
Computer Science

TL;DR

PatTree is proposed, a graph-based, holistic representation of patients that can be derived from real-world clinical data through the automated structuring of multimodal clinical data that enables early-stage data integration without relying on pre-standardized inputs.

Abstract

Access to holistic, multimodal data improves the performance of Artificial Intelligence (AI) in medical classification tasks compared to utilizing single modalities or data sources. However, the inherent heterogeneity and complexity of clinical real-world data pose significant challenges to structured data analysis and AI application. This heterogeneity includes missing values, multiple time points, diverse modalities, and inconsistent formats and semantics. Data harmonization prior to data integration tackles this challenge but remains resource-intensive and error-prone, limiting the scalability and reproducibility of holistic, AI-driven decision support on clinical real-world data. We therefore propose PatTree, a graph-based, holistic representation of patients that can be derived from real-world clinical data through the automated structuring of multimodal clinical data. PatTree enables early-stage data integration without relying on pre-standardized inputs. While representing heterogeneous clinical data within a unified knowledge graph, PatTree preserves the semantic relationships between data elements across modalities and data sources, facilitating interoperability and machine-interpretable data access. Using a subset of the ADNI-1 cohort (n = 763), we demonstrate that classification of patients is directly feasible on PatTree reaching state-of-the-art classification performance. In the three-class classification task distinguishing Alzheimer's disease, mild cognitive impairment, and cognitively normal individuals, we achieve a balanced accuracy of 98.5% and an F$_1$ score of 0.987 on the held-out test set. Our results show that assumption-free, automated structuring of multimodal medical data can serve as a scalable foundation for clinical AI pipelines bypassing tedious data preparation and standardization.

View source

Similar papers

Preprint Aug 2026

H2: A Dual Hybrid Semantic Data Lake Architecture for Medical Data Harmonization with Human-In-the-Loop verified, LLM Driven Metadata Annotation System

Medical data, by its nature, exhibit a high degree of heterogeneity on multiple levels ranging from (a) different modalities like images, text and time series, (b) diverse tabular schemata introduced by institutions and (c) completely unstructured textual information data provided by healthcare professionals. Data lake...

I. Tzortzis, Georgia Kapetadimitri, Agapi Davradou et al. · 0 citations
Open access Aug 2026

Managing the Unmanageable: Multimodal Artificial Intelligence for Unstructured Data Management and Analysis

Today, data are no longer confined to numerical values arranged in row-by-column matrices or stored neatly within relational databases. One of the defining characteristics of big data is its high variety, encompassing unstructured and multimodal forms such as text, audio, images, and video. These data types dominate co...

Chong-Ho Yu, Nino Miljkovic, Zhaoyang Wang · 0 citations
Open access Sep 2026

Ontology-aware knowledge graph retrieval-augmented generation for clinical decision support

Effectively retrieving and interpreting the vast, diverse, and largely unstructured data contained within electronic health records (EHRs) present significant challenges for clinical decision support systems. Large language models (LLMs), when applied to complex healthcare datasets, frequently exhibit hallucinations, l...

Deepak Panneerselvam, Sasikala E · 0 citations
#machine learning Preprint Sep 2026

LatentVerse: A Framework for Understanding Shared and Modality-Specific Information in Multimodal Latent Representations

Latent embeddings have become a central data abstraction in modern machine learning, especially in biomedicine, where foundation models are increasingly used to encode multimodal data like clinical text, medical images, omics, and physiological signals. However, the utility and value of these representations depends on...

Majd Alafrange, S. Friedman, J. Kitonyo et al. · 0 citations
Aug 2026

Diagnosis classification in EMR data using latent representations and SNOMED-CT mapping for improved medical data integration.

A diagnosis classification model that automatically maps diagnosis spans in EMR data to the standardized clinical ontology Systematized Nomenclature of Medicine-Clinical Terms (SNOMED-CT) is developed, demonstrating strong potential for scalable and privacy-preserving medical concept normalization in real-world clinica...

Sungsu Oh, I. Han, Jae Il Lee et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.