Skip to content
Open access

A Reproducible Evaluation of Hybrid Spectral–Temporal Features for Four-Class Respiratory Sound Event Classification

Sep 2026 · Signals · 0 citations · 32 references

TL;DR

These findings apply to direct early concatenation with the shared 1D-CNN backbone and do not imply that feature fusion is generally ineffective, and do not imply that feature fusion is generally ineffective.

Abstract

Respiratory-sound event classification is challenged by non-stationarity, class imbalance, heterogeneous acquisition, and participant-correlated recordings. This study evaluates whether direct fusion of short-time Fourier transform (STFT), mel-frequency cepstral coefficient (MFCC), and wavelet-packet features improves a temporal one-dimensional convolutional neural network (1D-CNN), and whether temporal convolution offers an advantage over conventional classifiers. A radial basis function support vector machine (RBF-SVM) and random forest were included deliberately to separate the value of the engineered representation from classifier complexity. Experiments used 920 recordings and 6898 annotated cycles from the International Conference on Biomedical and Health Informatics (ICBHI) 2017 Respiratory Sound Database. The predefined 60:40 recording-level benchmark partition was retained; model selection used five-fold participant-grouped cross-validation, and 95% confidence intervals were estimated from 1000 participant-level bootstrap resamples. The complete hybrid 1D-CNN achieved a macro-averaged F1-score of 0.313. MFCC alone yielded 0.320, but the paired difference was not statistically resolved. The RBF-SVM and random forest achieved 0.401 and 0.378, respectively. These findings apply to direct early concatenation with the shared 1D-CNN backbone and do not imply that feature fusion is generally ineffective. The study provides leakage-aware baselines, controlled ablations, clustered uncertainty estimates, and frozen artifacts for reproducible comparison.

Read PDF

Similar papers

Open access Sep 2026

Reliability-aware transfer learning with BEATs for variable-length respiratory sound classification: a patient-disjoint evaluation in two public databases

Random or central placement of respiratory sounds within padded windows can cause length-based masks to exclude recorded audio. We developed an exact support-propagation interface for the pretrained bidirectional encoder representation from audio transformers (BEATs), mapping sample support through filterbank frames an...

Jie Niu, Peng Li · 0 citations
Open access Aug 2026

Hybrid SVM and CNN System for Cough Detection and Health Classification

This study introduces a hybrid architecture designed to address the challenges of automated cough sound analysis by combining unsupervised detection with supervised classification. The system integrates a One-Class Support Vector Machine (SVM) trained exclusively on positive cough recordings to establish an acoustic bo...

Fabien Mouomene Moffo, A. Noumsi, Joseph Mvogo et al. · 0 citations
#machine learning Preprint Aug 2026

Spectral Features Dominate BCG Respiratory-Event Detection: A Large-Scale Patient-Independent Comparison of Feature Groups in Sleep Apnea Patients

A literature-guided, patient-independent comparison of ten BCG feature groups using a 512-sensor capacitive pressure mat recorded simultaneously with respiratory polygraphy in 155 patients undergoing in-hospital evaluation for obstructive sleep apnea shows that a compact, interpretable subset of the full feature librar...

I. C. Jurado, Z. Bousraou, Lara Benning et al. · 0 citations
Open access Sep 2026

A convolutional attention transformer network for ECG beat classification

The findings suggest that the CAT Network can serve as an effective and reproducible framework for AAMI-compliant ECG beat classification, supporting downstream decision support and large-scale screening, while motivating future work on cross-database generalization, computational optimization for edge deployment, and...

Shanmukha Rao Narsupalli, Rajesh Kumar Pullagura, Rajeswara Rao Gangula · 0 citations
#edge computing Sep 2026

Hybrid feature engineering for resource-efficient Arrhythmia detection in electrocardiogram signals: An interpretable, separability-driven framework.

Deep neural networks dominate automated arrhythmia detection, yet their reported performance often relies on intra-patient evaluation, which allows models to exploit patient-specific morphology rather than learn disease-related patterns. Such shortcuts are unavailable in deployed edge monitors, motivating a fundamental...

Moirangthem Tiken Singh, Manibhushan Yaikhom, R. K. Prasad · 0 citations
Open access Aug 2026

HSD-Net: a dual-branch CNN-BiLSTM network with hybrid cepstral fusion for heart sound classification

The proposed HSD-Net provides an effective framework for automated heart sound analysis and offers a reliable technical foundation for developing clinical decision-support systems, with significant potential for application in early screening and telemedicine.

Caijian Hua, Ye Tian, Liu-Ying Li et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.