QMF-MAP: A quantum multimodal fusion model for molecular activity prediction
Symbolic regression based on Kolmogorov-Arnold networks for multi-fidelity data fusion
Joint tensor self-representation and discriminative feature extraction for multi-view clustering
RAISE: Redundancy alleviation and informed structure enhancement for multimodal learning
Salience-guided counterfactual framework for multimodal sentiment analysis
Formulating and Unveiling In-context Learning over Graph Tasks
Leveraging visual foundation model priors for facade parsing from street view images
Evolutionary algorithms as information fusion architectures: A survey
Generative AI for wearable information fusion: Evidence mapping the reality gap in activity recognition and health monitoring
Beyond global, priors or multi-illumination supervision: Towards illumination-guided spatially adaptive white balancing for multi-illumination srgb images
TSCMixer: A multi-view learning and mixing framework for multivariate time series classification
Evolutionary dynamics of collective decision-making with local social influence on static and dynamic networks
Deep Augmented Inter-class Hashing for cross-modal retrieval
Incomplete disentangled multi-view representation learning with graph propagation
Multi-modal sentiment analysis with large language model and low-rank attention
Physics-consistent liquid state-space network for remaining useful life prediction with architecturally enforced degradation consistency
Layer-specific lipschitz modulation for fault-tolerant multimodal representation learning
Disambiguation-Enhanced robust partial label feature selection via fusion of feature-Specific priors and irrelevant-Label consistency
MWCE-GAN: Multi-way collaborative evolutionary generative adversarial network for trustworthy Alzheimer’s disease risk prediction based on imaging genetics data
A region representation for entity tagging and locating
Sparse-dense topology-aware fusion for PET-CT pulmonary lesion detection
Efficient multi-modal image fusion via flow-based model
Audio-guided visual selection for efficient multimodal fusion via a discrete variational autoencoder
PSAGAN : Prior-guided sample-adaptive graph attention network with cross-domain alignment for EEG-based emotion recognition
PreDyn-IDS: Pretrained lightweight open-set IDS for IoT via dynamic confidence and hierarchical feature fusion
Semantic-aware content alignment for multi-object video editing
Video entity linking in complex multimodal scenarios: High-Quality benchmark datasets and a general baseline framework
Identity suppression awareness for facial expression recognition via dual-path learning
Cross-modal spatiotemporal prediction in transportation: A survey of modeling paradigms, representations, alignment, and reasoning
Pairwise constraint weighted imputation for incomplete multi-view clustering with high missing rate
SPriG: Shared-only fusion with improvement-guided private gating for multimodal affective computing
IHEL: Multimodal remote sensing image classification via interactive heterogeneous expert learning
PRISM: A progressive disentanglement framework for robust subgraph learning in Alzheimer’s disease diagnosis
Consensus model fusing dynamic relationships and overlapping communities for non-cooperative behavior management in large-scale group decision-making
Semantic-continuous Mamba for space-spectral fusion
Hierarchical spatiotemporal fusion for full-body pose estimation using sparse wearable sensors
Empowering time series analysis with foundation models: A comprehensive survey
Enhanced Fourier-based network for Low-Light image enhancement
WormVIO: Deep visual-inertial odometry via hierarchical instinct-bias sensor fusion
GaitFuse: A hierarchical cross-modal alignment and uncertainty-aware fusion framework for multi-modal gait recognition
Lane-driven and goal-focused refinement transformer for multimodal trajectory prediction
Attribute-modulated geometric alignment attention for language-guided medical image segmentation
Nonnegative matrix factorization coupled with dynamic graph learning for clinical score prediction in neurodegenerative disease
Multimodal Helps Unimodal: Frequency-Aware Distillation for Cross-domain Few-shot Egocentric Action Recognition
Integrating graphs, large language models, and agents: Reasoning and retrieval
A learning-augmented singer model with online parameter adaptation for high-maneuver target tracking
Probabilistic object and label association algorithms for distributed multiobject tracking
Drift-suppressed visual localization via non-rigid rectification of vector road-network priors
CorrFuse: A cross-modal fusion framework for multimodal sentiment analysis
Dynamic deep subspace residual regularizer for hyperspectral image super-resolution