arXiv ScienceSearch

arXiv subjects

Nigel Bean

Publications and source records attributed to Nigel Bean.

9 recordsLinked to original sources

Weak convergence of quasi-birth-and-death processes with rational arrival process components to fluid queues

In this paper we construct a new approximation to a fluid queue as a quasi-birth-and-death process with rational arrival process components (QBD-RAP) and prove its convergence. Fluid queues are stochastic processes that move linearly at a rate governed by the state of a continuous-time Markov chain (CTMC), and are widely used to model telecommunications, power, risk, and storage systems. A key motivating application is to fluid-fluid queues, whose analysis proceeds via operator-analytic expressions involving the generator of the underlying fluid queue; these expressions are differential operators that are not, in general, readily computable, so approximation is needed. Existing approximations with a probabilistic interpretation guarantee valid probabilities but require a fine discretisation to be accurate, while methods such as the Discontinuous Galerkin approach are more accurate at a given discretisation level but can produce negative mass or probabilities exceeding one. Our (QBD-RAP) approximation addresses this. Because the QBD-RAP is itself a stochastic process, the approximation it produces automatically retains the defining properties of a probability, while promising improved numerical accuracy over existing probabilistic schemes for a given discretisation level. We prove that the generator of the QBD-RAP converges to the generator of the fluid queue, which is, to our knowledge, the first generator-theoretic convergence result for a process with rational arrival process components. The proof introduces a new technique for the analysis of RAP-modulated processes, analysing the generator via bases of conditional residual time distributions rather than the orbit process used in prior RAP analyses, and along the way establishes that the phase process of the QBD-RAP and of the fluid queue share the same distribution.

math.PR

Are we always in strife? A longitudinal study of the echo chamber effect in the Australian Twittersphere

Contrary to expectations that the increased connectivity offered by the internet and particularly Online Social Networks (OSNs) would result in broad consensus on contentious issues, we instead frequently observe the formation of polarised echo chambers, in which only one side of an argument is entertained. These can progress to filter bubbles, actively filtering contrasting opinions, resulting in vulnerability to misinformation and increased polarisation on social and political issues. These have real-world effects when they spread offline, such as vaccine hesitation and violence. This work seeks to develop a better understanding of how echo chambers manifest in different discussions dealing with different issues over an extended period of time. We explore the activities of two groups of polarised accounts across three Twitter discussions in the Australian context. We found Australian Twitter accounts arguing against marriage equality in 2017 were more likely to support the notion that arsonists were the primary cause of the 2019/2020 Australian bushfires, and those supporting marriage equality argued against that arson narrative. We also found strong evidence that the stance people took on marriage equality in 2017 did not predict their political stance in discussions around the Australian federal election two years later. Although mostly isolated from each other, we observe that in certain situations the polarised groups may interact with the broader community, which offers hope that the echo chambers may be reduced with concerted outreach to members.

cs.SI

Bayesian estimation of trend components within Markovian regime-switching models for wholesale electricity prices: an application to the South Australian wholesale electricity market

We discuss and extend methods for estimating Markovian-Regime-Switching (MRS) and trend models for wholesale electricity prices. We argue the existing methods of trend estimation used in the electricity price modelling literature either require an ambiguous definition of an extreme price, or lead to issues when implementing model selection [23]. The first main contribution of this paper is to design and infer a model which has a model-based definition of extreme prices and permits the use of model selection criteria. Due to the complexity of the MRS models inference is not straightforward. In the existing literature an approximate EM algorithm is used [26]. Another contribution of this paper is to implement exact inference in a Bayesian setting. This also allows the use of posterior predictive checks to assess model fit. We demonstrate the methodologies with South Australian electricity market.

stat.ME

Estimation of Markovian-regime-switching models with independent regimes

Markovian-regime-switching (MRS) models are commonly used for modelling economic time series, including electricity prices where independent regime models are used, since they can more accurately and succinctly capture electricity price dynamics than dependent regime MRS models can. We can think of these independent regime MRS models for electricity prices as a collection of independent AR(1) processes, of which only one process is observed at each time; which is observed is determined by a (hidden) Markov chain. Here we develop novel, computationally feasible methods for MRS models with independent regimes including forward, backward and EM algorithms. The key idea is to augment the hidden process with a counter which records the time since the hidden Markov chain last visited each state that corresponding to an AR(1) process.

stat.ME

A framework for streamlined statistical prediction using topic models

In the Humanities and Social Sciences, there is increasing interest in approaches to information extraction, prediction, intelligent linkage, and dimension reduction applicable to large text corpora. With approaches in these fields being grounded in traditional statistical techniques, the need arises for frameworks whereby advanced NLP techniques such as topic modelling may be incorporated within classical methodologies. This paper provides a classical, supervised, statistical learning framework for prediction from text, using topic models as a data reduction method and the topics themselves as predictors, alongside typical statistical tools for predictive modelling. We apply this framework in a Social Sciences context (applied animal behaviour) as well as a Humanities context (narrative analysis) as examples of this framework. The results show that topic regression models perform comparably to their much less efficient equivalents that use individual words as predictors.

stat.AP

Semi-supervised graph labelling reveals increasing partisanship in the United States Congress

Graph labelling is a key activity of network science, with broad practical applications, and close relations to other network science tasks, such as community detection and clustering. While a large body of work exists on both unsupervised and supervised labelling algorithms, the class of random walk-based supervised algorithms requires further exploration, particularly given their relevance to social and political networks. This work refines and expands upon a new semi-supervised graph labelling method, the GLaSS method, that exactly calculates absorption probabilities for random walks on connected graphs. The method models graphs exactly as discrete-time Markov chains, treating labelled nodes as absorbing states. The method is applied to roll call voting data for 42 meetings of the United States House of Representatives and Senate, from 1935 to 2019. Analysis of the 84 resultant political networks demonstrates strong and consistent performance of GLaSS when estimating labels for unlabelled nodes in graphs, and reveals a significant trend of increasing partisanship within the United States Congress.

cs.SI

A discontinuous Galerkin method for approximating the stationary distribution of stochastic fluid-fluid processes

Introduced by Bean and O'Reilly (2014), a stochastic fluid-fluid process is a Markov processes $\{X_t, Y_t, \varphi_t\}_{t \geq 0}$, where the first fluid $X_t$ is driven by the Markov chain $\varphi_t$, and the second fluid $Y_t$ is driven by $\varphi_t$ as well as by $X_t$. That paper derived a closed-form expression for the joint stationary distribution, given in terms of operators acting on measures, which does not lend itself easily to numerical computations. Here, we construct a discontinuous Galerkin method for approximating this stationary distribution, and illustrate the methodology using an on-off bandwidth sharing system, which is a special case of a stochastic fluid-fluid process.

math.PR

Pachinko Prediction: A Bayesian method for event prediction from social media data

The combination of large open data sources with machine learning approaches presents a potentially powerful way to predict events such as protest or social unrest. However, accounting for uncertainty in such models, particularly when using diverse, unstructured datasets such as social media, is essential to guarantee the appropriate use of such methods. Here we develop a Bayesian method for predicting social unrest events in Australia using social media data. This method uses machine learning methods to classify individual postings to social media as being relevant, and an empirical Bayesian approach to calculate posterior event probabilities. We use the method to predict events in Australian cities over a period in 2017/18.

cs.CY

Doubling Algorithms for Stationary Distributions of Fluid Queues: A Probabilistic Interpretation

Fluid queues are mathematical models frequently used in stochastic modelling. Their stationary distributions involve a key matrix recording the conditional probabilities of returning to an initial level from above, often known in the literature as the matrix $\Psi$. Here, we present a probabilistic interpretation of the family of algorithms known as \emph{doubling}, which are currently the most effective algorithms for computing the return probability matrix $\Psi$. To this end, we first revisit the links described in \cite{ram99, soares02} between fluid queues and Quasi-Birth-Death processes; in particular, we give new probabilistic interpretations for these connections. We generalize this framework to give a probabilistic meaning for the initial step of doubling algorithms, and include also an interpretation for the iterative step of these algorithms. Our work is the first probabilistic interpretation available for doubling algorithms.

math.PR