arXiv ScienceSearch

arXiv subjects

Sagar Kumar

Publications and source records attributed to Sagar Kumar.

10 recordsLinked to original sources

Integrated Real-Time Motion Tracking and AI Analysis for Athletic Performance Optimization

Applying Human Pose Estimation (HPE) in real world environments remains a challenging task, this paper explores and surveys real time HPE approaches and their limitations in sports analysis for individuals, alongside developing a practical lightweight prototype for real world testing and usage. The older marker-based motion capture systems evolving to the modern accessible and adaptable markerless deep learning approaches, this survey explores the foundational architectures, which balance precision and efficiency. We also compare algorithmic frameworks (top-down, bottom-up, one-stage approaches, etc.) on practical deployment metrics such as inference latency, frame rate, mean per-joint position error, and temporal jitter to guide model selection process for sports application. As our prime contribution, we are proposing a modular, lightweight software prototype, which uses MediaPipe HPE framework with multiple exercise specific logic to deliver real-time insights and AI based feedback for non-expert users. We derive sports insights and providing feedback with minimal computational resources, while showcasing the performance and reliability metrics. In the end, we suggest other future research directions like combining sensors, and AR/VR. This work caters to researchers, engineers, sport scientists, etc., as both technical resource and a valid blueprint to implement a similar or improved real-time HPE analysis system for athletic performance enhancement or other purposes.

cs.HC

Failure of contextual invariance in large language models

Standard evaluation practices assume that large language model (LLM) outputs are stable when prompts are embedded in contextually equivalent discourses. Here, we test this assumption in the setting of gender inference. Using a controlled pronoun selection task, we introduce minimal, theoretically uninformative discourse context and find that this induces large, systematic shifts in model outputs. Correlations with cultural gender stereotypes, present in decontextualized settings, weaken or disappear once context is introduced, while theoretically irrelevant features, such as the gender of a pronoun for an unrelated referent, become the most informative predictors of model behavior. A Contextuality-by-Default analysis reveals that, in 19--52\% of cases across models, this dependence persists after accounting for all marginal effects of context on individual outputs and cannot be attributed to simple pronoun repetition. These findings show that LLM outputs violate contextual invariance even under near-identical syntactic formulations, with implications for bias benchmarking and deployment in high-stakes settings.

cs.CL

When Life Gives You AI, Will You Turn It Into A Market for Lemons? Understanding How Information Asymmetries About AI System Capabilities Affect Market Outcomes and Adoption

AI consumer markets are characterized by severe buyer-supplier market asymmetries. Complex AI systems can appear highly accurate while making costly errors or embedding hidden defects. While there have been regulatory efforts surrounding different forms of disclosure, large information gaps remain. This paper provides the first experimental evidence on the important role of information asymmetries and disclosure designs in shaping user adoption of AI systems. We systematically vary the density of low-quality AI systems and the depth of disclosure requirements in a simulated AI product market to gauge how people react to the risk of accidentally relying on a low-quality AI system. Then, we compare participants' choices to a rational Bayesian model, analyzing the degree to which partial information disclosure can improve AI adoption. Our results underscore the deleterious effects of information asymmetries on AI adoption, but also highlight the potential of partial disclosure designs to improve the overall efficiency of human decision-making.

cs.HC

Cultural evolution of human beauty standards

Beauty standards shape self-perception and health through social comparison and objectification, while exposure to idealized imagery exacerbates body-image concerns. Media and fashion are central arbiters of these ideals, yet long-term, quantitative, intersectional studies on how representation has changed remain scarce. We assembled a dataset of 793199 records spanning 25 years of advertising, magazine covers, runway shows, and editorials to quantify changes in anthropometric and demographic representation. We find a paradox in the evolution of beauty ideals: while representational diversity has increased, the median model physique remains stable. This is driven by selective plus-size inclusion at the upper tail, while the typical physique continues to diverge from the US population. Intersectionally, non-white models are 4.5 times more likely to be plus-size, indicating that progress in size inclusivity falls disproportionately on multiple underrepresented identities. Stratifying the industry via a data-driven prestige hierarchy, we find that thinness is overrepresented at the top tier. Finally, comparing two regulatory interventions we observe that numeric thresholds are more effective at reducing underweight appearances. Our results quantify the cultural evolution in media and fashion, revealing that inclusion has increased; however, gains are uneven and intersectionally concentrated on size and ethnicity, whereas the prevailing thin ideal remains largely unchanged.

physics.soc-ph

HyDeFuse: Provably Convergent Denoiser-Driven Hyperspectral Fusion

Hyperspectral (HS) images provide fine spectral resolution but have limited spatial resolution, whereas multispectral (MS) images capture finer spatial details but have fewer bands. HS-MS fusion aims to integrate HS and MS images to generate a single image with improved spatial and spectral resolution. This is commonly formulated as an inverse problem with a linear forward model. However, reconstructing high-quality images using the forward model alone is challenging, necessitating the use of regularization techniques. In this work, we investigate the paradigm of denoiser-driven regularization, where a powerful off-the-shelf denoiser is used for implicit regularization within an iterative algorithm. This has shown much promise but remains relatively underexplored in hyperspectral imaging. The technical challenge lies in designing hyperspectral denoisers that can guarantee convergence while strong denoisers can produce high-quality reconstructions, they may also cause instability or divergence. Specifically, we consider a denoiser-driven fusion algorithm, HyDeFuse, which leverages a class of pseudo-linear denoisers for implicit regularization. We demonstrate how the contraction mapping theorem can be applied to establish global linear convergence of HyDeFUse. Finally, we validate our theoretical findings and present fusion results on publicly available datasets to demonstrate the performance of HyDeFuse.

eess.IV

A Blue Start: A large-scale pairwise and higher-order social network dataset

Large-scale networks have been instrumental in shaping how we think about social systems, and have undergirded many foundational results in mathematical epidemiology, computational social science, and biology. However, many of the social systems through which diseases spread, information disseminates, and individuals interact are inherently mediated through groups, known as higher-order interactions. A gap exists between higher-order models of group formation and spreading processes and the data necessary to validate these mechanisms. Similarly, few datasets bridge the gap between pairwise and higher-order network data. The Bluesky social media platform is an ideal laboratory for observing social ties at scale through its open API. Not only does Bluesky contain pairwise following relationships, but it also contains higher-order social ties known as "starter packs" which are user-curated lists designed to promote social network growth. We introduce "A Blue Start", a large-scale network dataset comprising 39.7M user accounts, 2.4B pairwise following relationships, and 365.8K groups representing starter packs. This dataset will be an essential resource for the study of higher-order networks.

physics.soc-ph

Multi-strain spreading dynamics under arbitrary transmission kernels

Compartmental models of epidemic dynamics have long described the propagation of a single, immutable transmissible state through a population via pairwise contact, and multi-strain generalizations have extended this framework to incorporate mutation, competition, and cross-immunity. Here we study a minimal generalization with no sink states or feedback, in which transmission acts through an arbitrary column-stochastic kernel $Q$ on a finite set of strains, encoding mutation during transmission with no further structural assumptions. We derive the mean-field approximation for the well-mixed regime and show that it admits an exact closed-form solution for any $Q$, expressible as a single matrix exponential applied to the initial condition. A spectral decomposition of this solution reveals that the location of the long-time attractor and the rate of approach are governed by the eigenstructure of $Q$. We extend the analysis to structured populations via a pairwise mean-field approximation on regular contact networks, and validate both approximations against stochastic simulations. The framework provides an entry into the analysis of dynamical systems in which mutation and transmission occur on the same time scale, drawing parallels to the propagation of discrete signals through populations under noisy communication.

physics.soc-ph

What we should learn from pandemic publishing

Authors of COVID-19 papers produced during the pandemic were overwhelmingly not subject matter experts. Such a massive inflow of scholars from different expertise areas is both an asset and a potential problem. Domain-informed scientific collaboration is the key to preparing for future crises.

physics.soc-ph

How We Define Harm Impacts Data Annotations: Explaining How Annotators Distinguish Hateful, Offensive, and Toxic Comments

Computational social science research has made advances in machine learning and natural language processing that support content moderators in detecting harmful content. These advances often rely on training datasets annotated by crowdworkers for harmful content. In designing instructions for annotation tasks to generate training data for these algorithms, researchers often treat the harm concepts that we train algorithms to detect - 'hateful', 'offensive', 'toxic', 'racist', 'sexist', etc. - as interchangeable. In this work, we studied whether the way that researchers define 'harm' affects annotation outcomes. Using Venn diagrams, information gain comparisons, and content analyses, we reveal that annotators do not use the concepts 'hateful', 'offensive', and 'toxic' interchangeably. We identify that features of harm definitions and annotators' individual characteristics explain much of how annotators use these terms differently. Our results offer empirical evidence discouraging the common practice of using harm concepts interchangeably in content moderation research. Instead, researchers should make specific choices about which harm concepts to analyze based on their research goals. Recognizing that researchers are often resource constrained, we also encourage researchers to provide information to bound their findings when their concepts of interest differ from concepts that off-the-shelf harmful content detection algorithms identify. Finally, we encourage algorithm providers to ensure their instruments can adapt to contextually-specific content detection goals (e.g., soliciting instrument users' feedback).

cs.CL

Do LLMs Understand Social Knowledge? Evaluating the Sociability of Large Language Models with SocKET Benchmark

Large language models (LLMs) have been shown to perform well at a variety of syntactic, discourse, and reasoning tasks. While LLMs are increasingly deployed in many forms including conversational agents that interact with humans, we lack a grounded benchmark to measure how well LLMs understand \textit{social} language. Here, we introduce a new theory-driven benchmark, SocKET, that contains 58 NLP tasks testing social knowledge which we group into five categories: humor & sarcasm, offensiveness, sentiment & emotion, and trustworthiness. In tests on the benchmark, we demonstrate that current models attain only moderate performance but reveal significant potential for task transfer among different types and categories of tasks, which were predicted from theory. Through zero-shot evaluations, we show that pretrained models already possess some innate but limited capabilities of social language understanding and training on one category of tasks can improve zero-shot testing on others. Our benchmark provides a systematic way to analyze model performance on an important dimension of language and points to clear room for improvement to build more socially-aware LLMs. The associated resources are released at https://github.com/minjechoi/SOCKET.

cs.CL