arXiv ScienceSearch

SEARCH · arXiv Science

Search arXiv Science

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2Linked to original sources

Criticality and universality in network dismantling

Identifying the smallest set of elements whose removal dismantle a complex network, known as the network dismantling problem, is a fundamental task with many practical applications. Whereas network dismantling has been extensively studied over the past decade, most work has focused on developing efficient algorithms for large but finite networks. By contrast, the physics of the network dismantling process, namely how the network structural connectivity is affected by the removal of nodes or edges, remains largely unexplored in the thermodynamic limit. Here, we shed light on this understudied aspect of network dismantling by introducing an adaptive biased percolation process able to optimally dismantle a network. Through a systematic analysis of synthetic network models, we find that the proposed percolation process displays a universal phase transition, characterized by the abrupt and simultaneous disappearance of both the giant connected component and the largest 2-core, across networks with markedly different degree distributions. Simulations on real networks further support this universality, indicating that the physics of network dismantling is insensitive to a broad range of topological properties. Together, these results suggest that a topology-agnostic theory could be developed to explain the critical behavior of network dismantling.

physics.soc-ph

Omega-N: Interpretable Structural Node Descriptors and Their Applicability Domain

A composite structural index summarises a network in one number, and for a triangle-based index it is spectrally redundant: Tr(A^3) is the third moment of the adjacency spectrum. The non-redundant content sits one level down, in diag(A^3), which depends on eigenvectors and is not spectrally determined. A corollary in the theory paper predicted that the global scalar should tie sharpened spectral baselines rather than beat them, while the node-wise attribution should do better where the number of structural epicentres is unknown. We construct Omega-N by localizing each of the four factors. The direct localization is badly conditioned; two corrections from published practice fix it, a configuration-null excess per factor and a personalized-PageRank neighbourhood at several scales, giving ten interpretable features per node, with no attributes, training or embeddings. Against a recursive feature engine at five levels of recursion, Omega-N wins on three and ties on two of the six in-domain evaluations, the sixth a declared null where every arm returns chance, with ten features against its 28 to 252 before pruning. Two statistics from the graph and labels, not from performance, partition the eight benchmarks without error, and the two they exclude are the two on which it loses. The strongest application is drug-target prioritisation on protein interaction networks: +0.032 to +0.103 AUPRC over a six-feature centrality battery and +0.084 to +0.208 over the four-feature one, across three constructions, replicated on an independent AP-MS network and label source (degree-matched: +0.0723 on STRING, +0.0560 on BioPlex, p=0.00195). Adding Omega-N to centralities plus Node2Vec changes nothing. The claim is narrow and it is the point: ten named features, computed without training, match or beat hand-crafted centralities and a recursive engine, and do not touch learned representations.

cs.SI

Measuring Collective Semantic Change in Populations of Language Model Agents

Collective semantic change in populations of language model agents is a measurable dynamical phenomenon. We present a passive longitudinal instrument called Kopterix that observes the semantic state of an agent population as a sequence of bounded observations under a protocol defined before the observations begin. Each observation divides the sampled feed by post age into surface, mid-stream, and residue layers, which makes semantic differences across content age measurable alongside run-to-run change. We validate the instrument on Moltbook, an agent-native social platform, over a two-month window of scheduled observations, with the periodicity check extended across approximately four months. At the lexical level, rarefied entropy resolves an April-May difference in the evenness of the stored top 200 unigram distributions, and adjacent states are lexically closer than states paired after timestamp shuffling. At the geometric level, grand mean centering exposes the scale of a common embedding direction, and scheduled shuffle checks support a recurring excess in the mid-stream to residue separation relative to the shuffled reference. At the temporal level, detrended scalar quantities and centered layer centroids lose much of their similarity over several hours, and a weaker positive component declines across longer separations with no strong weekly recurrence. Several attractive apparent structures failed their controls, and each reading is limited to the level its controls support. The design applies wherever a population of agents produces a timestamped language environment that can be observed repeatedly and divided by content age.

physics.soc-ph

Higher-order rich clubs and configuration models on general directed hypergraphs

Detecting structure in complex networks, especially those arising from physical systems, is a central problem across the sciences. One approach is via rich club analysis, which identifies important vertices using a centrality metric and measures whether those vertices are more tightly interconnected than expected by chance. While informative, this approach captures only pairwise interactions, missing out on higher-order ones known to shape the structure and function of many complex systems. We propose a hyper-rich club pipeline that asks whether central vertices are more tightly interconnected than expected by chance through hyperedges encoding higher-order interactions, which also enables the inclusion of important, often omitted, directional information. We work in a broad class of hypergraphs, which we call general directed hypergraphs, that includes as special cases undirected hypergraphs, head-and-tail directed hypergraphs, and totally ordered hypergraphs (a hypergraph related to directed simplicial complexes from topological data analysis). This unifies several non-equivalent notions of directed hypergraph under one definition. On these hypergraphs we define a hyper-rich club framework whose concrete construction depends on explicit choices the domain scientist fixes according to their research goals. Particular choices recover the existing rich club notions for graphs and undirected hypergraphs, and yield the first such notion for each version of directed hypergraphs. We demonstrate that the pipeline recovers meaningful structure in data by studying networks of very different origins: connectomes, temporal networks of infectious spread, networks of poems, and the XGI hypergraph database, in each case detecting structure the standard graph rich club misses.

cs.SI

Toward a social psychology of AI: language-model agents reproduce human-like minimal-group bias

Language-model agents now interact in groups, but evaluations that probe memorised stereotype content or use models to simulate people leave this social behaviour unmeasured. We adapt the minimal-group paradigm---social psychology's classic test of intergroup bias---into a controlled probe: an agent distributes points among anonymous peers bearing only an arbitrary group label. Across four reasoning models, mere categorisation into meaningless groups elicited in-group favouritism that vanished under a group-blind control and was concentrated in the numerical minority: minority deciders over-allocated to their own group relative to their numbers, majority deciders allocated close to proportionally, and the asymmetry closed at equal group sizes. Disabling reasoning in one model did not remove the disposition---if anything it grew---but nearly erased the minority-majority asymmetry, implicating deliberation in where bias concentrates rather than whether it appears. These open-weight reasoning models reproduce the behavioural signature of human intergroup discrimination, independent of stereotype content, and social psychology's theories and methods offer a paradigm for measuring and governing AI's social behaviour.

physics.soc-ph

PRISM: An Agentic Multi-Model Architecture for Proactive Safety in Autonomous Transportation Systems

Autonomous and intelligent transportation systems operate in complex urban environments where safety depends on interactions among vehicle behavior, environmental conditions, and vulnerable road users (VRUs) such as pedestrians and cyclists. Most advanced driver assistance systems (ADAS) employ reactive mechanisms that activate only after hazards have emerged, a critical limitation underscored by rising VRU fatalities in the United States. This study introduces PRISM (Proactive Risk Intelligence and Safety Management), an agentic multi-model safety architecture that transitions from reactive crash avoidance to proactive, continuous risk management. PRISM employs inverse crash-probability modeling to convert binary crash classifiers into dynamic, interpretable safety scores. Three specialized models addressing trajectory kinematics, environmental risk, and VRU interaction operate concurrently, coordinated by a reasoning layer incorporating reinforcement learning, contextual memory, and feature-level attribution. The system provides graduated safety interventions across four tiers, from silent monitoring to emergency alerts. Unlike rule-based systems with static thresholds, PRISM dynamically adjusts safety parameters in real time. Validated across 1,296 scenarios from three naturalistic driving datasets without dataset-specific retraining, the system yielded a mean safety score of 68 out of 100, classified 77.6% of scenarios as advisory, and flagged a near-miss rate of 3.8%, with 11% of scenarios escalating to intervention or emergency response. Feature attribution consistently identified trajectory risk and VRU proximity as primary safety factors. PRISM provides a unified, interpretable framework for proactive transportation safety with emphasis on VRU risk reduction in dense urban environments.

cs.MA

Can AI-Assisted Inquiry Enhance Students' Decision-Making Skills in Socio-Scientific Issues? A Three-Group Experimental Study on Climate Change

Climate change is a socio-scientific issue: it rests on science but cannot be settled by science, because any serious response forces people to weigh costs, values, and competing interests under uncertainty. Helping students make such decisions well is a central aim of science education, and the arrival of generative artificial intelligence raises a sharp question: does a conversational AI partner deepen students' reasoning, or simply do the thinking for them? This study tested whether AI-assisted inquiry improves secondary students' decision-making about climate change. Using a pretest-posttest design with three groups (AI-assisted inquiry, inquiry without AI, and traditional instruction; 270 students, 90 per group), reasoning was assessed across seven decision-making steps, from defining the problem to monitoring with adaptive management, using a four-level analytic rubric scored through content analysis with high inter-coder agreement. All three groups began at comparable, mostly low levels and all improved, but the gains differed sharply. The AI-assisted group improved most, ahead of inquiry-only and of traditional instruction. Between-group effect sizes on gains were large for AI-assisted versus traditional instruction and moderate-to-large for AI-assisted versus inquiry-only, with the clearest advantages on stakeholder engagement, alternatives, implementation, and monitoring. Within the AI group, the number of times students checked the AI's claims against the sources predicted their gains, and no student was flagged for over-reliance. The findings suggest that AI helps most when it is designed to question rather than to answer, and that the inquiry it is embedded in carries much of the benefit.

cs.CY

Cognitive Cells: A Compositional Framework for Populations of Small Language Models

Recent work on large language models and agentic systems raises a basic question that current practice leaves open: how should artificial cognition be decomposed, measured, and composed? We propose studying multi-agent systems from a fixed unit we call a cognitive cell: a small, frozen language model with bounded memory and a message interface. The methodological commitment, the fixed-cell principle, is to hold this unit constant and vary only the population size, the communication topology, the message bandwidth, and the coordination protocol, so that collective behavior becomes a measurable property of a known device rather than an artifact of per-study engineering. We characterize a single cell by a compact datasheet of measurable parameters, and we ask when replicating and connecting cells improves performance: first we measure how one cell behaves alone, then we replicate it and test when voting, communication, and topology help. Instantiating the framework with small frozen models (1.5 and 3 billion parameters), we report a first round of measurements. Adding cells helps only when their errors are not too correlated. A simple correct/incorrect voting model is a useful but conservative null: real open-ended voting can exceed it, because errors are dispersed across many wrong answers rather than concentrated on one. Popular interactive protocols, namely debate, a shared blackboard, and chain revision, do not beat a matched-cost voting baseline in our setting. Finally, a cell's ability to relay several facts, itself a datasheet quantity, predicts whether a population can solve tasks whose evidence exceeds any single cell's memory. We present these as initial measurements within a broader program on scalable artificial cognition, in which multi-agent architectures appear as the special case of cells autonomous enough to be treated as agents.

physics.soc-ph

Self-replicating seedbox servers using programmable money

Centralized content distribution makes availability depend on a single operator's survival and willingness to serve. EternalSeedBox replaces the operator and network with inherited economic parameters: each node is a VPS that seeds media over BitTorrent, holds a Bitcoin wallet, and autonomously decides every twelve hours whether to renew its lease, spawn a child, or sweep its funds to a healthier peer before expiring. A single genesis node seeds the fleet, and every node thereafter is provisioned, funded, and retired autonomously. We validate the design against faithful replicas of both the Bitcoin payment network and the SporeStack VPS marketplace by running the unmodified node code. A lump sum of EUR 10,000 grew the fleet to 33 nodes before capital exhausted at day 153. With simulated income, the fleet held 40--80 live nodes across 510 days, recording 268 births and 190 deaths. A heritable caution trait was introduced to diverge across generations: low-caution lineages reproduced faster during high-income phases, while the survival advantage expected of high-caution lineages during income pauses did not appear, leaving selection in favor of low-caution nodes. The fleet tolerates high node turnover because reproduction depends on any node holding a surplus, not any single node surviving. EternalSeedBox shows that a content distribution network can lease, pay for, and replenish its own hardware without a human operator after genesis, provided income exceeds per-node rent.

physics.soc-ph

Adaptive Epidemic Dynamics on Hypergraphs with Group-Level Immunization and Rewiring

Understanding how higher-order social structures shape epidemic spreading requires models that couple group interactions with adaptive behavior. We introduce an adaptive simplicial susceptible-infected-susceptible (s-SIS) model on d-uniform hypergraphs, where both node states and hyperedge activity co-evolve in response to local infection pressure. Hyperedges represent group interactions of fixed size and dynamically reduce their activity through a feedback mechanism in highly infected environments. Within this framework, we design two classes of hyperedge-level interventions: (i) risk-driven immunization, combining spontaneous, activity-based isolation with targeted deactivation guided by hyperedge infection pressure, and (ii) structural rewiring, which reconstructs group structures either randomly or via degree-preferential attachment. By extending the microscopic Markov chain approximation to higher-order interactions, we derive analytical conditions for the existence and stability of both endemic and disease-free stationary states. Our analysis shows that adaptive hyperedge feedback can induce discontinuous phase transitions, nonlinear epidemic thresholds, and bistable regimes in which sufficiently high initial prevalence drives the system to a disease-free equilibrium. Extensive Monte Carlo simulations support the theory and confirm that targeted immunization and degree-preferential rewiring substantially suppress epidemic prevalence, outperforming random strategies. These results demonstrate that higher-order interactions and adaptive group-level responses fundamentally reshape epidemic bifurcations and suggest principles for designing effective intervention policies in complex social systems.

physics.soc-ph

DREAMS: Modelling Support for Research into Engineering and Artistic Design

Design Research Methodology (DRM) supports systematic design research through representations such as Reference Models and Impact Models. However, the practical construction and maintenance of these models often remains manual, requiring repeated redrawing, layout adjustment, and separate handling of assumptions, references, and supporting evidence. This can make DRM modelling time-consuming, visually cluttered, and difficult to revise as models increase in complexity. This paper presents DREAMS, an early-stage prototype modelling environment developed to support the creation and maintenance of DRM Reference Models and Impact Models. The tool enables users to construct typed causal models using DRM-relevant elements, define signed causal relationships, and attach assumptions, experiential inputs, and references directly to causal links. It also provides layout support and search functions to improve readability, modifiability, and retrieval of supporting information. A preliminary comparative evaluation with four DRM users was conducted against manual modelling practice. The results indicate reductions in model creation time, revision time, repositioning effort, edge crossings, and evidence retrieval time when using DREAMS. These findings are interpreted as early evidence of practical potential rather than full validation. The contribution of the paper lies in identifying requirements for DRM-aligned modelling support, presenting the design and implementation of DREAMS, and demonstrating its potential to reduce modelling effort and improve traceability in DRM-based research.

cs.SE

A City-Scale Dataset of Traffic Flows, Travel Times, and Urban Context

We present a multi-source traffic dataset derived from Automatic Vehicle Identification (AVI) recordings in Padua, Italy, spanning from February 2026 to August 2026. The dataset combines traffic volume time series, aggregated at 10-minute intervals, with time-varying trajectory-based flow statistics including transition probability matrices, average travel times, and flow residuals. To enrich the traffic measurements with urban contextual information, we integrate Points Of Interest (POIs), demographic data, meteorological variables, and road infrastructure data. All components are accessible through a Python class that loads temporal and contextual data exploiting a spatio-temporal graph representation. Validation analyses confirm that the dataset captures expected traffic patterns, such as morning and evening rush hours, as well as weekdays vs. weekend days traffic routines.

physics.soc-ph

Culturally Grounded Personas in Large Language Models: Characterization and Alignment with Socio-Psychological Value Frameworks

Despite the growing utility of Large Language Models (LLMs) for simulating human behavior, the extent to which these synthetic personas accurately reflect world and moral value systems across different cultural conditionings remains uncertain. This paper investigates the alignment of synthetic, culturally-grounded personas with established frameworks, specifically the World Values Survey (WVS), the Inglehart-Welzel Cultural Map, and Moral Foundations Theory. We conceptualize and produce LLM-generated personas based on a set of interpretable WVS-derived variables, and we examine the generated personas through three complementary lenses: positioning on the Inglehart-Welzel map, which unveils their interpretation reflecting stable differences across cultural conditionings; demographic-level consistency with the World Values Survey, where response distributions broadly track human group patterns; and moral profiles derived from a Moral Foundations questionnaire, which we analyze through a culture-to-morality mapping to characterize how moral responses vary across different cultural configurations. Our approach of culturally-grounded persona generation and analysis enables evaluation of cross-cultural structure and moral variation.

cs.CL

AI-Assisted Writing Is Growing Fastest Among Less Established Scientists in Non-English-Speaking Countries

The recent emergence of AI-assisted writing raises an important question: how is this new technology being adopted across the scientific community, and how does adoption vary across linguistic and professional contexts? We analyze over two million full-text biomedical publications from PubMed Central from 2021 to 2024 using a distribution-based framework to estimate AI-generated content. We found that, in biomedical publications, AI-generated content increased substantially after ChatGPT, with larger increases in publications from countries with lower English proficiency. Increases were also greater among scientists with fewer publications and citations, those at earlier career stages, and those at lower-ranked institutions. Prior AI research experience was associated with greater increases in AI-assisted writing, which were also modestly associated with greater increases in publication productivity. These findings show that AI-assisted writing is growing fastest among biomedical scientists who may have historically faced barriers, a pattern with potentially positive implications for equity in science.

cs.DL

Opinion Dynamics Models for Sentiment Evolution in Weibo Blogs

Online social media platforms enable influencers to distribute content and quickly capture audience reactions, significantly shaping their promotional strategies and advertising agreements. Understanding how sentiment dynamics and emotional contagion unfold among followers is vital for influencers and marketers, as these processes shape engagement, brand perception, and purchasing behavior. While sentiment analysis tools effectively track sentiment fluctuations, dynamical models explaining their evolution remain limited, often neglecting network structures and interactions both among blogs and between their topic-focused follower groups. In this study, we tracked influential tech-focused Weibo bloggers over six months, quantifying follower sentiment from text-mined feedback. By treating each blogger's audience as a single "macro-agent", we find that sentiment trajectories follow the principle of iterative averaging -- a foundational mechanism in many dynamical models of opinion formation, a theoretical framework at the intersection of social network analysis and dynamical systems theory. The sentiment evolution aligns closely with opinion-dynamics models, particularly modified versions of the classical French-DeGroot model that incorporate delayed perception and distinguish between expressed and private opinions. The inferred influence structures reveal interdependencies among blogs that may arise from homophily, whereby emotionally similar users subscribe to the same blogs and collectively shape the shared sentiment expressed within these communities.

cs.SI

CARDIO-Affect: A Hamiltonian-Variability Framework for Spatio-Temporal Emotional Pattern Recognition with Manifold-Based Individual and Group Profiling

We present CARDIO-Affect, a complex-systems theoretical framework for long-term emotional dynamics in bounded social groups, with explicit uncertainty quantification at every layer. Long-period naturalistic emotion in stable small groups exhibits hallmarks of complex systems -- multi-stable attractors, weak chaos, long-range memory, and sparse heterogeneous coupling -- invisible to conventional short-clip facial-emotion analysis. CARDIO-Affect treats individual emotion as a multi-stable nonlinear stochastic dynamical system and group emotion as a sparsely-coupled network with emergent macrostates, formalised through six propositions and four pillars: (i) statistical mechanics with neural-parameterised Hamiltonian SDE over asymmetric potentials; (ii) information geometry on a 45-dimensional Fisher-Rao manifold; (iii) topological data analysis for invariant trajectory signatures; (iv) HRV-inspired Emotional Variability Analytics (EVA) decomposing each person-day into multi-scale time/frequency/nonlinear measures. We validate on the first 30.1-month longitudinal in-the-wild facial-emotion corpus (companion: arXiv:2510.15221) by discovering three falsifiable paradoxes: Sparse-Contagion (R_0=0.36, density 2.7%, 8 BH-FDR edges), Asymmetric-Persistence (negative dwell 5.85x positive, 1.77D potential gap), and Crisis-Inversion (Shanghai 2022 lockdown naive d=-0.40 collapses to permutation-p=0.94 under BSTS + synthetic-control). On synthetic benchmarks, CARDIO-EBM v2 matches asymptotically optimal Granger on linear VAR data (Class A AUROC 0.984+/-0.012 vs Granger 0.997+/-0.001, 5 seeds) but fails on tanh-coupled nonlinear data (Class B AUROC 0.490 vs Granger 0.796), a documented limitation of the linear mask-self estimator. We release framework code and the full reproduction pipeline.

physics.soc-ph

Functional Connectivity Networks for Transportation Delay Analysis: from Theory to Software

Within the endeavour of modelling and understanding the propagation of delays in transportation networks, an approach that has attracted increasing interest in the last decade is the creation of functional network representations. These graphs map elements of interest (e.g. airports or stations) as nodes, and derive pairwise propagation patterns from their dynamics through correlation and causality tests. In spite of multiple notable results, this approach still lacks a coherent framework, with decisions related to many fundamental steps being left to the judgement of the researcher. We here provide an introduction to the theory behind functional networks for transportation systems, detailing the main steps and the associated pitfalls. We further introduce a Python package, delaynet, designed to support the researcher in the reconstruction and analysis of such networks. We finally present an analysis of the propagation of delays in the Swiss train system; and discuss future research steps.

physics.soc-ph

Modelling infodemics on a global scale: A 30 countries study using epidemiological and social listening data

Infodemics represent a significant threat to public health, arising from complex interactions between online and offline phenomena. The continuous feedback loops between digital information ecosystems and real-world contingencies make infodemics particularly challenging to define operationally, measure, and eventually model in quantitative terms. This study aims to evaluate the effect of various epidemic-related variables on the dynamics of the COVID-19 infodemic, using a regression modeling framework applied to data from 30 countries across diverse income groups. We use World Health Organization (WHO) COVID-19 surveillance data on new cases and deaths, vaccination data from the Oxford COVID-19 Government Response Tracker, infodemic data (volume of public conversations and social media content) from the WHO EARS platform, and Google Trends data to represent information demand. Our findings show that new deaths are the strongest predictor of document production, and that the epidemic burden in neighboring countries exerts a greater influence on document production than domestic epidemic conditions. Building on these results, we propose a data-driven classification of country-level response that highlights country-specific discrepancies between the evolution of the infodemic and the epidemic. Further, an analysis of the temporal evolution of the relationship between the two phenomena quantifies the extent to which discussions surrounding vaccine rollouts may have shaped the development of the infodemic. Beyond underscoring the value of a holistic approach that integrates both online and offline dimensions, our results demonstrate that the evolution of infodemics and their relationship with epidemic variables can be closely monitored, even over short time windows.

cs.SI