arXiv Science⌕ Search

arXiv · 2609.34846

Conspiracy and Environment Communities on Reddit Sort Along Different Demographic Axes

Abstract

The impact of online platforms in political mobilization is widely documented, yet it remains unclear how it affects different demographic groups. To shed light on this phenomenon, we examine two contrasting cases: engagement with environmental causes on one side, and with conspiracy theories on the other. Both are well represented on Reddit, allowing us to study the different origins of these groups using a network-based method. By inferring age, gender, affluence, and partisanship of Reddit users, we construct stratified transition networks to compare the entry pathways of different demographic groups toward these communities. We find that demographic differentiation operates differently in these two domains. Entry into conspiracy communities is primarily structured by gender and partisanship: masculine users disproportionately arrive through overtly political and alt-right spaces and share right-wing media sources, whereas feminine users more often pass through esoteric and spiritual subreddits before converging on r/conspiracy. In contrast, environmental communities are differentiated mainly by age and affluence. Affluent users tend to approach environmental subreddits via discussions of individual energy management, technology, and finance, while less affluent users arrive through protest-oriented and climate movement spaces. By comparing these two issue domains, the study shows that sociodemographic sorting does not operate uniformly across polarized topics. Instead, distinct demographic groups build their own paths, shaping both the routes through which they engage with political content and the informational ecosystems that sustain their engagement. Our approach, based on the attention-flow graph, refines our understanding of how different groups become embedded in online political communities.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Miguel Á. Sánchez-Cortés, Corrado Monti, Gianmarco De Francisci Morales. 2026-09-28. Conspiracy and Environment Communities on Reddit Sort Along Different Demographic Axes. https://arxiv.org/abs/2609.34846

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

A Safety-First Gateway Architecture for Trusted Public Health Resource Navigation

Conversational AI can improve access to public health information, but public-facing healthcare applications require safeguards against inappropriate medical guidance and unsupported generation. We present a Safety-First Science Gateway for maternal and child health (MCH) resource navigation that combines large language models (LLMs) and retrieval-augmented generation (RAG) with a multi-layer safety architecture. The gateway integrates emergency handling, domain/scope screening, source attribution, anonymous session management, and operational audit logging while restricting retrieval to curated institutional resources. We describe the gateway architecture, prototype implementation, and functional verification of selected workflows. The current system provides resource provenance and safety-bounded navigation; it does not constitute a clinical decision-support system or automated claim-by-claim verification of generated health information. This work provides a reusable architectural framework for conversational navigation of curated public-health resources.

cs.CY↗

Greenpixie's AI Token Methodology: Assessing the Energy, Water and CO2-eq Impact of AI Tokens for Open and Closed Weight Models

We describe a methodology for estimating the per-token energy cost of cloud-hosted large language model (LLM) inference, separating between input (prefill) and output (decode) tokens. Graphics processing unit (GPU) energy usage is measured during inference benchmarking with open-weights models on a wide range of text-based tasks. The remaining server energy contribution from non-GPU hardware is estimated from the inference wall time. Bayesian linear regression is used to model the relationship between energy per token and LLM size, request traffic, and hardware deployment configuration. Proprietary frontier LLMs of unknown size and deployment are binned into size buckets based on naming conventions and performance priors, and the space of possible LLM configurations is sampled with Monte-Carlo methods to give a representative average energy per token and uncertainty. We also describe how these energy measurements can be used to estimate the carbon-dioxide equivalent ($\mathrm{CO_2\text{-}eq}$) emissions, both usage and embodied, and water consumed per token of AI inference. This methodology provides actionable data that enables reductions in cost, electricity usage, $\mathrm{CO_2\text{-}eq}$ emitted and water consumed in cloud and Software as a Service (SaaS).

cs.CY↗

Early Prediction of AI-Assisted Cheating Risk in Online Exams Through Learning Analytics

AI-assisted cheating has become an important threat to the security of online exams. This study examines whether the risk of AI-assisted cheating in the final exam can be predicted using students' digital traces in the learning management system (LMS) during the first eight weeks of the semester. The sample comprised 52 first-year undergraduates enrolled in a bachelor's program in Computer Education and Instructional Technology and taking an Introduction to Programming course at a public university in Turkiye. Students were labeled as low- or high-risk based on suspicious behaviors recorded in the final-exam logs, including copy, focus-loss, and right-click events. Of the 52 students, 23 (44.2%) were labeled as high-risk in a proctored, face-to-face exam. Group membership was then predicted using five features selected from 27 candidates extracted from students' digital traces. Logistic Regression, Naive Bayes, Random Forest, and Gradient Boosting algorithms were used to build the prediction models. Model performance was evaluated using leave-one-out cross-validation (LOOCV) with fold-specific preprocessing and feature selection. Logistic Regression achieved the best performance (Accuracy = 73.1%). The results indicate that LMS interaction data can provide an early signal of AI-assisted cheating risk. Course-module views, assignment submissions, and the number of days on which course videos were accessed were the most consistently selected features across the LOOCV folds. These predictions are intended to support timely academic guidance, not to establish misconduct or initiate disciplinary action.

cs.CY↗