arXiv · 2601.18305
SwipeGen: Bridging the Execution Gap in GUI Agents via Human-like Swipe Synthesis
Abstract
Despite numerous Graphical User Interface (GUI) agents claiming to automate user interaction tasks, to date, few achieve satisfactory interaction capability with human users in real-world scenarios. Through empirical analysis, this paper identifies the root cause of the limited interaction capability as the rigid swipe execution. In particular, unlike humans, who perform swipes with fine-grained control over trajectory, speed, and timing, existing agents can only conduct simplistic, deterministic swipe behaviors, leading to frequent failures on complicated user-like interaction tasks. Due to the lack of open-source human-like swipe training data, we propose SwipeGen, the first tool for synthesizing diverse and human-like swipe interactions, and SwipeBench, the first benchmark for evaluating agents' swipe interaction quality. Extensive experiments show that SwipeGen can improve the swipe execution success rate of existing agents by up to 2.46x. Our code, dataset, and model are available at https://github.com/TSKGHS17/SwipeGen.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Xuan Wang, Siyuan Su, Quantong Fu, Yongxiang Hu, Yangfan Zhou. 2026-01-26. SwipeGen: Bridging the Execution Gap in GUI Agents via Human-like Swipe Synthesis. https://doi.org/10.1145/3767308.3835803
Cite the original work for its findings. Save a collection to share your selection of sources.