Scholay

学术搜索 · AI 审稿 · LaTeX 协作

Residual bayesian attention networks for uncertainty quantification in regression tasks

作者:Youliang Chen, Wencan Guan, Rafig Azzam · 发表于:Scientific Reports · 年份:2025 · DOI:10.1038/s41598-025-24093-6 · 被引用次数:10 · 研究领域:Model Reduction and Neural Networks、Gaussian Processes and Bayesian Inference、Advanced Multi-Objective Optimization Algorithms

The demand for uncertainty quantification in modern sequence modeling tasks has prompted researchers to explore deep integration between Bayesian inference and Transformer architectures, but existing methods still face systematic engineering challenges in key technical aspects such as attention mechanism probabilization, residual connection uncertainty propagation, and epistemic-aleatoric uncertainty decoupling. This study proposes the Residual Bayesian Attention (RBA) framework, which achieves end-to-end probabilistic inference capabilities through three tightly coupled core components: Bayesian feedforward layers establish differentiable propagation mechanisms for parameter-level uncertainty, multi-layer residual Bayesian attention embeds radial basis function kernels into attention computation and introduces adaptive residual weights modeled by Beta distributions, and the Bayesian covariance construction module generates mathematically rigorous covariance representations through outer product operations and eigenvalue correction. Systematic evaluation on benchmark datasets covering six domains including engineering optimization, time series forecasting, and spatial modeling demonstrates that RBA achieves stable uncertainty quantification performance in medium-scale structured data scenarios, particularly exhibiting technical advantages in prediction interval calibration quality. Notably, through objective evaluation of challenging tasks such as complex physical systems, th...