arXiv ScienceSearch

arXiv · 2604.08910

Lightweight and Generalizable Multi-Sensor Human Activity Recognition via Cascaded Fusion and Style-Augmented Decomposition

Abstract

Wearable Human Activity Recognition (WHAR) is a prominent research area within ubiquitous computing, whose core lies in effectively modeling intra- and inter-sensor spatio-temporal relationships from multi-modal time series data. Existing methods either suffer from high computational complexity due to attention-based fusion or lack robustness to data variations during feature extraction. To address these issues, we propose a lightweight and generalizable framework that retains the core "decomposition-extraction-fusion" paradigm while introducing two key innovations. First, we replace the computationally expensive Attention and Cross-Variable Fusion (CVF) modules with a Cascaded Fusion Block (CFB), which achieves efficient feature interaction without explicit attention weights through the operational process of "compression-recursion-concatenation-fusion". Second, we integrate a MixStyle-based data augmentation module before the Local Temporal Feature Extraction (LTFE) and Global Temporal Aggregation (GTA) stages. By mixing the mean and variance of different samples within a batch and introducing random coefficients to perturb the data distribution, the model's generalization ability is enhanced without altering the core information of the data. The proposed framework maintains sensor-level, variable-level, and channel-level independence during the decomposition phase, and achieves efficient feature fusion and robust feature extraction in subsequent processes. Experiments on two benchmark datasets (Realdisp, Skoda) demonstrate that our model outperforms state-of-the-art methods in both accuracy and macro-F1 score, while reducing computational overhead by more than 30\% compared to attention-based baselines. This work provides a practical solution for WHAR applications on resource-constrained wearable devices.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Wang Chenglong, Zhuo Yan, Ding Wenbo, Chen Xinlei. 2026-04-10. Lightweight and Generalizable Multi-Sensor Human Activity Recognition via Cascaded Fusion and Style-Augmented Decomposition. https://arxiv.org/abs/2604.08910

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Explanation Navigator: Rectifying Out-of-Scope Human Interpretations of Leaky AI Explanations through Conversational Guidance

As explanations of artificial intelligence systems proliferate, their recipients must grasp not only what they convey but also recognise what they cannot. We conducted an interview study with nine participants to examine how explainees reason when their information needs exceed the scope of available explanations. Participants often unwittingly confabulated explanatory insights when relevant information was missing from the explanations, not recognising the inherent limitations thereof. We characterise such explanations as leaky explanations -- simplifications that strive to hide complexity yet whose correct interpretation hinges on understanding of the concealed details. To address out-of-scope interpretations we propose Explanation Navigator: a conversational interaction framework that detects mismatches between users' information needs and explanations' content, elucidating pertinent yet implicit details and providing complementary explanations for unmet information needs. An online study with 316 participants showed that our approach allowed explainees to recognise and rectify confabulated explanatory insights, guiding them towards developing correct understanding.

cs.HC

Biased AI improves human performance but reduces perceived helpfulness

Artificial intelligence (AI) increasingly shapes how people think, engage, and evaluate information. To minimize risk, most current systems are designed to present as ideologically neutral with standardized output. Yet growing evidence suggests that these principles suppress cognitive engagement, impair human decision-making, and erode societal diversity. Here we test the opposite approach by deliberately injecting bias into AI assistants. In three randomized experiments with 5,000 participants, biased AI improved human performance relative to default and neutral AI in tasks ranging from misinformation evaluation and financial investment to graduate education. These gains carried a subjective cost. Participants systematically undervalued AI they believed to be biased and inflated the helpfulness of AI they believed to be neutral, regardless of the systems' actual behavior. Interacting with two AIs whose biases flanked the participant's own perspective preserved the performance gains while limiting the subjective cost and one-sided influence. Our findings reveal the strategic value of intentional bias in AI design. Rather than performing a single fair, reliable, and authoritative voice, AI that speaks from specific viewpoints triggers cognitive agency and elevates human-AI performance in judgment, decision-making, and problem-solving.

cs.HC

Learning Password Best Practices Through In-Task Instruction

Users often make security- and privacy-relevant decisions without a clear understanding of the rules that govern safe behavior. We introduce pedagogical friction, a design approach that inserts brief, instructional interactions at the moment of action. We evaluate this approach in the context of password creation, a familiar task with clear quality criteria. We conducted a randomized study with 128 participants across four interface conditions that varied the depth and interactivity of guidance. We assessed three outcomes: (1) rule compliance in a subsequent password task without guidance, (2) accuracy on survey questions tied to password rules, and (3) behavior-knowledge alignment, which captures whether participants who correctly followed a rule also recognized it on the survey. Across the guided conditions, participants corrected most rule violations in the follow-up task and showed high behavior-knowledge alignment. Survey results suggested clearer advantages for some rule types, especially symbol related questions. These results position pedagogical friction as a lightweight intervention for security- and privacy-critical interfaces.

cs.HC