arXiv ScienceSearch

arXiv subjects

Mingyue Zha

Publications and source records attributed to Mingyue Zha.

4 recordsLinked to original sources

Who Anchors AI Overviews in Health? Baidu, Google, and the Geography of Authority

Artificial intelligence is being rapidly incorporated into traditional search systems, yet scant work audits the information disparities across platforms, geography, and languages. We address this gap by comparing Google and Baidu's AI Overview systems for health queries, and measure informational anchors that emerge. Auditing 1,920 health queries across 12 countries and 4 languages, we find that Google and Baidu exhibit vertical integration, routing users toward their own company platforms in AI Overviews rather than a diverse set of primary sources. Smaller, lower-localization countries receive fewer domestically sourced references for health queries. Issuing the same query in a country's official language rather than English raises the share of locally sourced citations approximately 3.5- to 13.5-fold. Comparing queries across health topics of varying severity and controversy, including Traditional Chinese Medicine as an example, we also show that health disclaimers are multidimensional and vary across language and culture. We discuss how generative search influences access to health information, and the urgent need for culturally-aware oversight of these systems that influence critical health decisions.

cs.IR

Interpreting Multimodal Communication at Scale in Short-Form Video: Visual, Audio, and Textual Mental Health Discourse on TikTok

Short-form video platforms integrate text, visuals, and audio into complex communicative acts, yet existing research analyzes these modalities in isolation, lacking scalable frameworks to interpret their joint contributions. This study introduces a pipeline combining automated multimodal feature extraction with Shapley value-based interpretability to analyze how text, visuals, and audio jointly influence engagement. Applying this framework to 162,965 TikTok videos and 814,825 images about social anxiety disorder (SAD), we find that facial expressions outperform textual sentiment in predicting viewership, informational content drives more attention than emotional support, and cross-modal synergies exhibit threshold-dependent effects. These findings demonstrate how multimodal analysis reveals interaction patterns invisible to single-modality approaches. Methodologically, we contribute a reproducible framework for interpretable multimodal research applicable across domains; substantively, we advance understanding of mental health communication in algorithmically mediated environments.

cs.MM

Gender Inequalities in Content Collaborations: Asymmetric Creator Synergy and Symmetric Audience Biases

Content-creator collaborations are a widespread strategy for enhancing digital viewership and revenue. While existing research has explored the efficacy of collaborations, few have looked at inequities in collaborations, particularly from the perspective of the supply and demand of attention. Leveraging 42,376 videos and 6,117,441 comments from YouTube (across 150 channels and 3 games), this study examines gender inequality in collaborative environments. Utilizing Shapley value, a tool from cooperative game theory, results reveal dominant in-group collaborations based on in-game affordances. However, audience responses are aligned across games, reflecting symmetric biases across the gaming communities, with comments focusing more on peripherals than actual gameplay for women. We find supply-side asymmetries exist along with demand-side symmetries. Our results engage with the larger literature on digital and online biases, highlighting how genre and affordances moderate gendered collaboration, the direction of inequality, and contributing a general framework to quantify synergy across collaborations.

cs.CY

The Meme Is the Message: Generative Memesis and AI Visuals in the 2024 USA Presidential Elections

Visual content on social media has become increasingly influential in shaping political discourse and civic engagement, but it also limits participation due to the increased cost of multimedia production. In tandem, the growth of generative AI provides novel ways for citizens to participate in politics by lowering these costs. Drawing on a dataset of 239,526 Instagram images, we analyze the effects of synthetic images during the 2024 United States presidential election, using a multimodal workflow combining computer vision, large language models, and facial affect analysis. Results show that meme format is a stronger predictor of engagement than AI-generated content alone. However, AI-generated memes yield a significant interaction effect, suggesting synergistic increases in engagement when synthetic imagery is integrated with memes through human curation. We also characterize how users curate images. Partisans use AI in different ways: Democrat-leaning users tend to use it for in-group support, whereas Republican-leaning users more often employ it for out-group attacks. Users generally select happier synthetic faces compared to real photographs. We define generative memesis as a mode of communication in which memes are no longer shared person-to-person, but mediated by AI through customized visuals. We discuss how generative AI may empower civic participation, the bifurcation of content production and curation, and its implications for in the history of novel technologies and participatory culture.

cs.CY