arXiv ScienceSearch

arXiv subjects

Michael Small

Publications and source records attributed to Michael Small.

At least 19 recordsLinked to original sources

Pricing the Unpriced Asset: A Standards-Based Method for Valuing Enterprise Data under IAS 38 and IAS 2

The recognition and measurement of data assets under current accounting standards presents significant challenges. While International Accounting Standard 38 (IAS 38) provides a framework for intangible asset recognition, data assets frequently fail to meet capitalisation criteria due to difficulties in demonstrating separability, establishing reliable cost measurement, and proving probable future economic benefits. The widespread failure to easily and reliably value data causes mispricing and allocative distortions across data and artificial intelligence markets. This paper introduces a two-layer valuation progression for authenticated data assets, that is, datasets that have met IAS 38 recognition criteria through established legal provenance and contractual boundaries. The first layer, D-Val, is the auditable cost-basis valuation consistent with IAS 38. D-Val is defined as D-Val = Cp * Avt, where Cp is the reliably measurable production cost and Avt is the appreciation or depreciation factor applied over time. Under prevailing interpretations of IAS 38, Av is constrained to values less than or equal to one absent an active market revaluation, rendering D-Val a strictly cost-less-amortisation figure. The second layer, A-Val, is a theoretically grounded commercial valuation that incorporates scarcity, rivalry, completeness, accuracy, and explicit premia for provenance authentication and independent audit. A-Val is not auditable as fair value under current practice but serves as a defensible commercial valuation during the period before active markets for authenticated data assets mature. As authenticated data markets mature parameter assumptions improve providing a foundation for iterative refinement of the model.

cs.CE

Dynamics, Complexity and Time Series Analysis

The aim of this text is to provide a linguistically accessible, but comprehensive introduction into a variety of topics in dynamical systems and its applications. Whilst preliminary knowledge of dynamical systems is useful, it is not essential and readers are only assumed to have familiarity with foundational undergraduate mathematics topics of calculus, linear algebra and rudimentary statistics. A variety of extended topics on recent publications and research activities in the field have been included in the last four chapters, which the interested reader may use as an introduction into further reading. A collection of exercises and questions both theoretical and computational are also included in this text.

math.DS

Do triangles matter? Replicating hypergraph disease dynamics with lower-order interactions

Disease spreading models such as the ubiquitous SIS compartmental model and its numerous variants are widely used to understand and predict the behaviour of a given epidemic or information diffusion process. A common approach to imbue more realism to the spreading process is to constrain simulations to a network structure, where connected nodes update their disease state based on pairwise interactions along the edges of their local neighbourhood. Simplicial contagion models (SCM) extend this to hypergraphs such that groups of three nodes are able to interact and propagate the disease along higher-order hyperedges (triangles). Though more flexible, it is not clear the extent to which the inclusion of these higher-order interactions result in dynamics that are characteristically different to those attained from simpler pairwise interactions. Here, we propose an agent-based model that unifies the classical SIS/SIR compartmental model and SCM, and extends it to allow for interactions along hyperedges of arbitrary order. Using this model, we demonstrate how the steady-state dynamics of pairwise interactions can be made to replicate those of simulations that include higher-order topologies by linearly scaling disease parameters based on a proposed measure of network activity. By allowing disease parameters to dynamically vary over time, lower-order pairwise interactions can be made to closely replicate both the transient and steady-state dynamics of higher-order simulations. We demonstrate that this relationship is robust to misspecification in the assumed higher-order interaction model, and applies to non-clique complex hypergraphs with non-trivial heterogeneous topology. For the latter case, it is found that heterogeneities in hypergraph topology result in weakened approximations of higher-order dynamics by pairwise interactions.

math.DS

Economy and Geography Shape the Collective Attention of Cities

Complex networks are commonly used to explore human behavior. However, previous studies largely overlooked the geographical and economic factors embedded in collective attention. To address this, we construct attention networks from time-series data for the United States and China, each a key economic power in the West and the East, respectively. We reveal a strong macroscale correlation between urban attention and Gross Domestic Product (GDP). At the mesoscale, community detection of attention networks shows that high-GDP cities consistently act as core nodes within their communities and occupy strategic geographic positions. At the microscale, structural hole theory identifies these cities as key connectors between communities, with influence proportional to economic output. Overlapping community detection further reveals tightly connected urban clusters, prompting us to introduce geographic and topic-based metrics, which show that closely linked cities are spatially proximate and topically coherent. Of course, not all patterns were consistent across regions. A notable distinction emerged in the relationship between population size and urban attention, which was evident in the United States but absent in China. Building on these insights, we integrate key variables reflecting GDP, geography, and scenic resources into regression model to cross-verify the influence of economic and geographic factors on collective user attention, and unexpectedly discover that a composite index of population, access, and scenery fails to account for cross-city variations in attention. Our study bridges the gap between economic prosperity and geographic centrality in shaping urban attention landscapes.

physics.soc-ph

Dynamic Object Geographic Coordinate Recognition: An Attitude-Free and Reference-Free Framework via Intrinsic Linear Algebraic Structures

The Earth, a temporal complex system, is witnessing a shift in research on its coordinate system, moving away from conventional static positioning toward embracing dynamic modeling. Early positioning concentrates on static natural geographic features, with the emergence of geographic information systems introducing a growing demand for spatial data, the focus turns to capturing dynamic objects. However, previous methods typically rely on expensive devices or external calibration objects for attitude measurement. We propose an applied mathematical model that utilizes time series, the nature of dynamic object, to determine relative attitudes without absolute attitude measurements, then employs SVD-based methods for 3D coordinate recognition. The model is validated with negligible error in a numerical simulation, which is inherent in computer numerical approximations. What in follows, to assess our model in the engineering scenario, we propose a framework featuring the integration of applied mathematics with AI, utilizing only three cameras to capture an UAV. We enhance the YOLOv8 model by leveraging time series for the accurate 2D coordinate acquisitions, which is then used as input for 2D-to-3D conversion via our mathematics model. As a result, the framework demonstrates high precision, as evidenced by low error metrics including root mean square error, mean absolute error, maximum error, and a strong R-squared value. It is important to note that the mathematical method itself is inherently error-free; any observed inaccuracies are due solely to external hardware or the AI-based 2D coordinate acquisition process, which represents an improved version of the current state-of-the-art. Our framework enriches geodetic theory by providing a streamlined model for the 3D positioning of non-cooperative targets, minimizing input attitude parameters, leveraging applied mathematics and AI.

physics.geo-ph

Triadic Closure-Heterogeneity-Harmony GCN for Link Prediction

Link prediction aims to estimate the likelihood of connections between pairs of nodes in complex networks, which is beneficial to many applications from friend recommendation to metabolic network reconstruction. Traditional heuristic-based methodologies in the field of complex networks typically depend on predefined assumptions about node connectivity, limiting their generalizability across diverse networks. While recent graph neural network (GNN) approaches capture global structural features effectively, they often neglect node attributes and intrinsic structural relationships between node pairs. To address this, we propose TriHetGCN, an extension of traditional Graph Convolutional Networks (GCNs) that incorporates explicit topological indicators -- triadic closure and degree heterogeneity. TriHetGCN consists of three modules: topology feature construction, graph structural representation, and connection probability prediction. The topology feature module constructs node features using shortest path distances to anchor nodes, enhancing global structure perception. The graph structural module integrates topological indicators into the GCN framework to model triadic closure and heterogeneity. The connection probability module uses deep learning to predict links. Evaluated on nine real-world datasets, from traditional networks without node attributes to large-scale networks with rich features, TriHetGCN achieves state-of-the-art performance, outperforming mainstream methods. This highlights its strong generalization across diverse network types, offering a promising framework that bridges statistical physics and graph deep learning.

cs.SI

Machine Learning Informed by Micro and Mesoscopic Statistical Physics Methods for Community Detection

Community detection plays a crucial role in understanding the structural organization of complex networks. Previous methods, particularly those from statistical physics, primarily focus on the analysis of mesoscopic network structures and often struggle to integrate fine-grained node similarities. To address this limitation, we propose a low-complexity framework that integrates machine learning to embed micro-level node-pair similarities into mesoscopic community structures. By leveraging ensemble learning models, our approach enhances both structural coherence and detection accuracy. Experimental evaluations on artificial and real-world networks demonstrate that our framework consistently outperforms conventional methods, achieving higher modularity and improved accuracy in NMI and ARI. Notably, when ground-truth labels are available, our approach yields the most accurate detection results, effectively recovering real-world community structures while minimizing misclassifications. To further explain our framework's performance, we analyze the correlation between node-pair similarity and evaluation metrics. The results reveal a strong and statistically significant correlation, underscoring the critical role of node-pair similarity in enhancing detection accuracy. Overall, our findings highlight the synergy between machine learning and statistical physics, demonstrating how machine learning techniques can enhance network analysis and uncover complex structural patterns.

cs.SI

A simple model of global cascades on random hypergraphs

This study introduces a comprehensive framework that situates information cascades within the domain of higher-order interactions, utilizing a double-threshold hypergraph model. We propose that individuals (nodes) gain awareness of information through each communication channel (hyperedge) once the number of information adopters surpasses a threshold $\phi_m$. However, actual adoption of the information only occurs when the cumulative influence across all communication channels exceeds a second threshold, $\phi_k$. We analytically derive the cascade condition for both the case of a single seed node using percolation methods and the case of any seed size employing mean-field approximation. Our findings underscore that when considering the fractional seed size, $r_0 \in (0,1]$, the connectivity pattern of the random hypergraph, characterized by the hyperdegree, $k$, and cardinality, $m$, distributions, exerts an asymmetric impact on the global cascade boundary. This asymmetry manifests in the observed differences in the boundaries of the global cascade within the $(\phi_m, \langle m \rangle)$ and $(\phi_k, \langle k \rangle)$ planes. However, as $r_0 \to 0$, this asymmetric effect gradually diminishes. Overall, by elucidating the mechanisms driving information cascades within a broader context of higher-order interactions, our research contributes to theoretical advancements in complex systems theory.

physics.soc-ph

Model Calibration and Validation From A Statistical Inference Perspective

Despite the general consensus in transport research community that model calibration and validation are necessary to enhance model predictive performance, there exist significant inconsistencies in the literature. This is primarily due to a lack of consistent definitions, and a unified and statistically sound framework. In this paper, we provide a general and rigorous formulation of the model calibration and validation problem, and highlight its relation to statistical inference. We also conduct a comprehensive review of the steps and challenges involved, as well as point out inconsistencies, before providing suggestions on improving the current practices. This paper is intended to help the practitioners better understand the nature of model calibration and validation, and to promote statistically rigorous and correct practices. Although the examples are drawn from a transport research background - and that is our target audience - the content in this paper is equally applicable to other modelling contexts.

stat.ME

Ordinal Poincar\'e Sections: Reconstructing the First Return Map from an Ordinal Segmentation of Time Series

We propose a robust and computationally efficient algorithm to generically construct first return maps of dynamical systems from time series without the need for embedding. Typically, a first return map is constructed using a heuristic convenience (maxima or zero-crossings of the time series, for example) or a computationally delicate geometric approach (explicitly constructing a Poincar\'e section from a hyper-surface normal to the flow and then interpolating to determine intersections with trajectories). Our approach relies on ordinal partitions of the time series and builds the first return map from successive intersections with particular ordinal sequences. Generically, we can obtain distinct first return maps for each ordinal sequence. We define entropy-based measures to guide our selection of the ordinal sequence for a ``good'' first return map and show that this method can robustly be applied to time series from classical chaotic systems to extract the underlying first return map dynamics. The results are shown on several well-known dynamical systems (Lorenz, R{\"o}ssler and Mackey-Glass in chaotic regimes).

math.DS

Selecting embedding delays: An overview of embedding techniques and a new method using persistent homology

Delay embedding methods are a staple tool in the field of time series analysis and prediction. However, the selection of embedding parameters can have a big impact on the resulting analysis. This has led to the creation of a large number of methods to optimise the selection of parameters such as embedding lag. This paper aims to provide a comprehensive overview of the fundamentals of embedding theory for readers who are new to the subject. We outline a collection of existing methods for selecting embedding lag in both uniform and non-uniform delay embedding cases. Highlighting the poor dynamical explainability of existing methods of selecting non-uniform lags, we provide an alternative method of selecting embedding lags that includes a mixture of both dynamical and topological arguments. The proposed method, {\em Significant Times on Persistent Strands} (SToPS), uses persistent homology to construct a characteristic time spectrum that quantifies the relative dynamical significance of each time lag. We test our method on periodic, chaotic and fast-slow time series and find that our method performs similar to existing automated non-uniform embedding methods. Additionally, $n$-step predictors trained on embeddings constructed with SToPS was found to outperform other embedding methods when predicting fast-slow time series.

math.DS

Backpropagation on Dynamical Networks

Dynamical networks are versatile models that can describe a variety of behaviours such as synchronisation and feedback. However, applying these models in real world contexts is difficult as prior information pertaining to the connectivity structure or local dynamics is often unknown and must be inferred from time series observations of network states. Additionally, the influence of coupling interactions between nodes further complicates the isolation of local node dynamics. Given the architectural similarities between dynamical networks and recurrent neural networks (RNN), we propose a network inference method based on the backpropagation through time (BPTT) algorithm commonly used to train recurrent neural networks. This method aims to simultaneously infer both the connectivity structure and local node dynamics purely from observation of node states. An approximation of local node dynamics is first constructed using a neural network. This is alternated with an adapted BPTT algorithm to regress corresponding network weights by minimising prediction errors of the dynamical network based on the previously constructed local models until convergence is achieved. This method was found to be succesful in identifying the connectivity structure for coupled networks of Lorenz, Chua and FitzHugh-Nagumo oscillators. Freerun prediction performance with the resulting local models and weights was found to be comparable to the true system with noisy initial conditions. The method is also extended to non-conventional network couplings such as asymmetric negative coupling.

math.DS

Characterisation of neonatal cardiac dynamics using ordinal partition network

The maturation of the autonomic nervous system (ANS) starts in the gestation period and it is completed after birth in a variable time, reaching its peak in adulthood. However, the development of ANS maturation is not entirely understood in newborns. Clinically, the ANS condition is evaluated with monitoring of gestational age, Apgar score, heart rate, and by quantification of heart rate variability using linear methods. Few researchers have addressed this problem from the perspective nonlinear data analysis. This paper proposes a new data-driven methodology using nonlinear time series analysis, based on complex networks, to classify ANS conditions in newborns. We map $74$ time series given by RR intervals from premature and full-term newborns to ordinal partition networks and use complexity quantifiers to discriminate the dynamical process present in both conditions. We obtain three complexity quantifiers (permutation, conditional and global node entropies) using network mappings from forward and reverse directions, and considering different time lags and embedding dimensions. The results indicate that time asymmetry is present in the data of both groups and the complexity quantifiers can differentiate the groups analysed. We show that the conditional and global node entropies are sensitive for detecting subtle differences between the neonates, particularly for small embedding dimensions ($m < 7$). This study reinforces the assessment of nonlinear techniques for RR intervals time series analysis.

math.DS

Analysis of Activity Dependent Development of Topographic Maps in Neural Field Theory with Short Time Scale Dependent Plasticity

Topographic maps are a brain structure connecting pre-synpatic and post-synaptic brain regions. Topographic development is dependent on Hebbian-based plasticity mechanisms working in conjunction with spontaneous patterns of neural activity generated in the pre-synaptic regions. Studies performed in mouse have shown that these spontaneous patterns can exhibit complex spatial-temporal structures which existing models cannot incorporate. Neural field theories are appropriate modelling paradigms for topographic systems due to the dense nature of the connections between regions and can be augmented with a plasticity rule general enough to capture complex time-varying structures. We propose a theoretical framework for studying the development of topography in the context of complex spatial-temporal activity fed-forward from the pre-synaptic to post-synaptic regions. Analysis of the model leads to an analytic solution corroborating the conclusion that activity can drive the refinement of topographic projections. The analysis also suggests that biological noise is used in the development of topography to stabilise the dynamics. MCMC simulations are used to analyse and understand the differences in topographic refinement between wild-type and the $\beta2$ knock-out mutant in mice. The time scale of the synaptic plasticity window is estimated as $0.56$ seconds in this context with a model fit of $R^2 = 0.81$.

q-bio.NC

Consistency capacity of reservoir computers

We study the propagation and distribution of information-carrying signals injected in dynamical systems serving as a reservoir computers. A multivariate correlation analysis in tailored replica tests reveals consistency spectra and capacities of a reservoir. These measures provide a high-dimensional portrait of the nonlinear functional dependence on the inputs. For multiple inputs a hierarchy of capacity measures characterizes the interference of signals from each source. For each input the time-resolved capacity forms a nonlinear fading memory profile. We illustrate the methodology with various types of echo state networks.

cond-mat.dis-nn

Navigating differential structures in complex networks

Structural changes in a network representation of a system (e.g.,different experimental conditions, time evolution), can provide insight on its organization, function and on how it responds to external perturbations. The deeper understanding of how gene networks cope with diseases and treatments is maybe the most incisive demonstration of the gains obtained through this differential network analysis point-of-view, which lead to an explosion of new numeric techniques in the last decade. However, {\it where} to focus ones attention, or how to navigate through the differential structures can be overwhelming even for few experimental conditions. In this paper, we propose a theory and a methodological implementation for the characterization of shared "structural roles" of nodes simultaneously within and between networks, whose outcome is a highly {\em interpretable} map. The main features and accuracy are investigated with numerical benchmarks generated by a stochastic block model. Results show that it can provide nuanced and interpretable information in scenarios with very different (i) community sizes and (ii) total number of communities, and (iii) even for a large number of 100 networks been compared (e.g., for 100 different experimental conditions). Then, we show evidence that the strength of the method is its "story-telling"-like characterization of the information encoded in a set of networks, which can be used to pinpoint unexpected differential structures, leading to further investigations and providing new insights. We provide an illustrative, exploratory analysis of four gene co-expression networks from two cell types $\times$ two treatments (interferon-$\beta$ stimulated or control). The method proposed here allowed us to elaborate and test a set of very specific hypotheses related to {\em unique} and {\em subtle} nuances of the structural differences between these networks.

physics.data-an

West Australian Pandemic Response: The Black Swan of Black Swans

The COVID-19 Pandemic has been described as the global challenge of our time, an enormous human tragedy with dramatic economic impacts. This paper describes the response and expected recovery process for Western Australia, where a rapid and effective response was implemented. This has enabled an early transition into an expected recovery both in health and economic terms. The positive lessons learned from this experience are documented as they emerge in order to support other states and nations as they address this issue globally in the near-term and consider enduring improvements for the longer term. While the authors have personal experience in the WA context, wider observations across Australia and selected international benchmarks are also included. Key lessons include the importance of good health advice in Australia's interest; timely, synchronized and aligned action at all levels of government; a program of well communicated, aligned health and economic measures which support all in society allowing a very high level of appropriate community behaviour, ensuring the health system was not overloaded; innovation in telehealth, testing, pandemic modelling, and integrated operations which also allowed essential industries to continue; and strong border and travel controls with highly effective isolation preventing community spread, ultimately enabling rapid elimination of the disease from the hospital system. In combination, these demonstrate that in the case of Western Australia the result of first eliminating the disease from the community, and then reopening the economy progressively at a strong pace, has enabled a world leading outcome in both in health and economic terms. The lessons from this experience are widely applicable, shareable both as supporting service to other regions and through knowledge transfer.

q-bio.PE

Modelling remote epidemic transmission in Western Australia and implications for pandemic response

We develop an agent-based model of disease transmission in remote communities in Western Australia. Despite extreme isolation, we show that the movement of people amongst a large number of small but isolated communities has the effect of causing transmission to spread quickly. Significant movement between remote communities, and regional and urban centres allows for infection to quickly spread to and then among these remote communities. Our conclusions are based on two characteristic features of remote communities in Western Australia: (1) high mobility of people amongst these communities, and (2) relatively high proportion of travellers from very small communities to major population centres. In models of infection initiated in the state capital, Perth, these remote communities are collectively and uniquely vulnerable. Our model and analysis does not account for possibly heightened impact due to preexisting conditions, such additional assumptions would only make the projections of this model more dire. We advocate stringent monitoring and control of movement to prevent significant impact on the indigenous population of Western Australia.

q-bio.PE