arXiv · 2311.02506
Non-Hierarchical Transformers for Pedestrian Segmentation
Abstract
We propose a methodology to address the challenge of instance segmentation in autonomous systems, specifically targeting accessibility and inclusivity. Our approach utilizes a non-hierarchical Vision Transformer variant, EVA-02, combined with a Cascade Mask R-CNN mask head. Through fine-tuning on the AVA instance segmentation challenge dataset, we achieved a promising mean Average Precision (mAP) of 52.68\% on the test set. Our results demonstrate the efficacy of ViT-based architectures in enhancing vision capabilities and accommodating the unique needs of individuals with disabilities.
Explore related subjects
Keep this discovery
Amani Kiruga, Xi Peng. 2023-07-11. Non-Hierarchical Transformers for Pedestrian Segmentation. https://arxiv.org/abs/2311.02506
Cite the original work for its findings. Save a collection to share your selection of sources.