arXiv ScienceSearch

arXiv subjects

Abhishek Srinivas

Publications and source records attributed to Abhishek Srinivas.

2 recordsLinked to original sources

A Composition-Aware Pretraining Framework for Geospatial Foundation Models

Geospatial foundation models have emerged as state-of-the-art methods for downstream Earth observation tasks. However, existing pretraining methodologies process imagery through a single-concept lens, failing to capture the highly compositional nature of complex satellite scenes. We propose a composition-aware pretraining framework that explicitly encodes fractional land-cover mixtures. Each satellite image cell is mapped to a histogram representing its fractional land-cover distribution, which we term the "composition target". These targets serve as the primary prediction objective and are distilled into the backbone using Earth Mover's Distance. Experimental evaluation shows that composition-aware pretraining yields substantial gains on region-level understanding tasks requiring semantic similarity judgment, including zero-shot image retrieval and scene classification, while remaining competitive on tasks requiring fine-grained spatial precision, such as segmentation and object detection. With a 36.8M-parameter backbone, our framework outperforms SatMAE and Prithvi-EO-2.0, which contain 303M and 600M parameters, respectively, in most retrieval and scene classification settings. On the fine-grained ForestNet-12 dataset, a rigorous testbed for compositional discrimination, our method boosts baseline mAP@10 from 0.279 to 0.434, a 55.6% relative improvement, providing direct evidence for the effectiveness of explicit composition modeling. The code implementation can be found at https://github.com/05kashyap/GFM_Composition_Pretraining

cs.CV

Task-Invariant Learning of Continuous Joint Kinematics during Steady-State and Transient Ambulation Using Ultrasound Sensing

Natural control of limb motion is continuous and progressively adaptive to individual intent. While intuitive interfaces have the potential to rely on the neuromuscular input by the user for continuous adaptation, continuous volitional control of assistive devices that can generalize across various tasks has not been addressed. In this study, we propose a method to use spatiotemporal ultrasound features of the rectus femoris and vastus intermedius muscles of able-bodied individuals for task-invariant learning of continuous knee kinematics during steady-state and transient ambulation. The task-invariant learning paradigm was statistically evaluated against a task-specific paradigm for the steady-state (1) level-walk, (2) incline, (3) decline, (4) stair ascent, and (5) stair descent ambulation tasks. The transitions between steady-state stair ambulation and level-ground walking were also investigated. It was observed that the continuous knee kinematics can be learned using a task-invariant learning paradigm with statistically comparable accuracy to a task-specific paradigm. Statistical analysis further revealed that incorporating the temporal ultrasound features significantly improves the accuracy of continuous estimations (p < 0.05). The average root mean square errors (RMSEs) of knee angle and angular velocity estimation were 7.06{\deg} and 53.1{\deg}/sec, respectively, for the task-invariant learning compared to 6.00{\deg} and 51.8{\deg}/sec for the task-specific models. High accuracy of continuous task-invariant paradigms overcome the barrier of task-specific control schemes and motivate the implementation of direct volitional control of lower-limb assistive devices using ultrasound sensing, which may eventually enhance the intuitiveness and functionality of these devices towards a "free form" control approach.

cs.RO