Skip to content
Open access

Human behavior alignment to improve the robustness of deep neural networks

Aug 2026 · Frontiers in Artificial Intelligence · Vol 9 · 0 citations · 24 references
Medicine

TL;DR

This work proposes BrainTrain, a framework to create more robust DNNs through human behavior alignment and shows its utility in the context of object recognition and proposes Similarity Driven Label Smoothing (SDLS), a regularization method that scales BrainTrain to applications where it is difficult or expensive to collect human behavioral data.

Abstract

Deep neural networks (DNNs) are known to produce erroneous results under real-world noisy inputs, presenting a major bottleneck to their use in applications where lives, safety, or significant resources are at stake. It has been commonly observed that humans are highly resilient to the noisy inputs that are challenging for DNNs. However, very few efforts have translated this observation into techniques to improve DNN robustness. We hypothesize that statistically aligning DNNs to human behavior during training could improve robustness. Based on this insight, we propose BrainTrain, a framework to create more robust DNNs through human behavior alignment and demonstrate its utility in the context of object recognition. BrainTrain captures human behavior in the form of a confusion matrix constructed from human responses to object recognition challenges and uses a composite loss function to co-optimize accuracy and human behavior alignment during stochastic gradient descent (SGD) based training. We also propose Similarity Driven Label Smoothing (SDLS), a regularization method that scales BrainTrain to applications where it is difficult or expensive to collect human behavioral data. DNNs trained with BrainTrain showed up to 26% higher accuracy under a wide range of noisy inputs and 2.1 times lower calibration error with negligible increase in training time. We also demonstrate that SDLS leads to improvements in noise robustness in ResNets trained on the ImageNet-1 K image classification dataset. We show that BrainTrain is complementary to conventional techniques like noise-added training and can provide further improvements over and above these techniques.

Read PDF

Similar papers

Open access Sep 2026

Predictive coding networks capture human neural representations missing in supervised DNNs

Neuroscientific learning theories propose that the brain acquires knowledge by constructing internal world models. Supervised learning, the dominant approach in deep neural networks (DNN), relies on external category labels, making it difficult to reconcile with biological learning. There is an increasing trend towards...

Dirk Gütlin, Denise Kittelmann, Ryszard Auksztulewicz · 0 citations
Open access Aug 2026

Systematic image perturbations reveal persistent gaps between human and machine vision

An image set that systematically untangles global shape, internal parts, and texture information is created, and human recognition behavior against >200 DNNs spanning diverse architectures, training diets, and training objectives is compared, revealing systematic and persistent differences between human and machine vis...

Mugihiko Kato, Biyu J. He · 1 citation
Open access Aug 2026

Brain alignment in deep neural networks emerges early and independently of object classification

Deep convolutional neural networks are leading models of biological vision, largely because of their strong brain alignment: their features predict neural responses better than earlier models. Yet they are believed to recognize objects differently, relying on texture where humans rely on shape and failing on perturbati...

H. Scholte, Niklas Müller, Julio Smidi et al. · 0 citations
Preprint Aug 2026

Relational Knowledge Distillation Brings DNN Representations Close Enough to Humans to Be Aligned Without Supervision

Linking the internal representations of deep neural networks (DNNs) to human mental representations is important for using DNNs as computational models of human vision. Existing DNN representations remain insufficiently similar to human mental representations, which are not directly observable and are therefore commonl...

Yuria Shimizu, Soh Takahashi, Takato Horii et al. · 0 citations
Preprint Aug 2026

SeFaR: Semantic Feature-aware Robustness Testing of Deep Neural Networks

Deep neural networks are increasingly deployed in safety-critical domains as perception modules, where failures are often caused due to rare and under-represented scenarios. This necessitates the need to evaluate the semantic robustness of perception models; conformance of behavior to high-level requirements over real-...

Nusrat Jahan Mozumder, Divya Gopinath, Corina S. Păsăreanu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.