arXiv Science⌕ Search

arXiv · 1405.5623

Stochastic variational inference for large-scale discrete choice models using adaptive batch sizes

Abstract

Discrete choice models describe the choices made by decision makers among alternatives and play an important role in transportation planning, marketing research and other applications. The mixed multinomial logit (MMNL) model is a popular discrete choice model that captures heterogeneity in the preferences of decision makers through random coefficients. While Markov chain Monte Carlo methods provide the Bayesian analogue to classical procedures for estimating MMNL models, computations can be prohibitively expensive for large datasets. Approximate inference can be obtained using variational methods at a lower computational cost with competitive accuracy. In this paper, we develop variational methods for estimating MMNL models that allow random coefficients to be correlated in the posterior and can be extended easily to large-scale datasets. We explore three alternatives: (1) Laplace variational inference, (2) nonconjugate variational message passing and (3) stochastic linear regression. Their performances are compared using real and simulated data. To accelerate convergence for large datasets, we develop stochastic variational inference for MMNL models using each of the above alternatives. Stochastic variational inference allows data to be processed in minibatches by optimizing global variational parameters using stochastic gradient approximation. A novel strategy for increasing minibatch sizes adaptively within stochastic variational inference is proposed.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Linda S. L. Tan. 2015-10-08. Stochastic variational inference for large-scale discrete choice models using adaptive batch sizes. https://doi.org/10.1007/s11222-015-9618-x

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Jagged AI in Scientific Peer Review: Evidence from POMP Data Analysis

Despite their growing use in academic writing and statistical analysis, the performance of artificial intelligence (AI) tools in scientific peer review remains a largely unexplored area. A key challenge is jagged AI, a phenomenon where AI exhibits strong ability spikes in some domains while remaining deficient in others. To study this jaggedness in a practical data science context, we considered the task of reviewing partially observed Markov process (POMP) data analyses. POMP models, a generalization of state-space models or hidden Markov models, are used to fit mechanistic dynamic models to time series data in diverse applications including disease transmission, ecological dynamics, and financial risk assessment. High-quality peer review in this area entails assessment of scientific context, identification of errors in implementing complex algorithms, and decisions concerning methodological best practices. We studied 72 POMP projects from four semesters of a University of Michigan graduate time series course for which the project reports, the source code, and student peer reviews are anonymized and open access. We compared the human reviews with four AI review agents, using Claude Code with differing instructions implemented as skill files. We found that the AI review agents exhibited a jagged capability profile, identifying plausible technical errors and instances of invalid inference methodology overlooked by humans, while showing lower capability at finding issues involving interpretive errors, narrative coherence, and domain-informed model critique. The jaggedness was similar for all agents, consistent with other research finding that the specific instructions can have limited effect on the capability of AI models for some tasks. Skill file configuration shifted which weaknesses agents emphasized, without removing the jaggedness.

stat.AP↗

Hybrid Models for Short-Term Sea-Level Forecasting

Accurate tide forecasts are essential for coastal management, navigation, flood-risk reduction, and infrastructure protection. Observed sea level can be decomposed into astronomical and non-astronomical components, the latter mainly driven by meteorological effects. This study investigates a hybrid framework for hourly sea-level forecasting that combines harmonic analysis (HA) for the astronomical component with data-driven models for the non-astronomical contribution. The approach is evaluated at six tide-gauge stations with different tidal regimes: Venice, Trieste, Saint-Malo, Vardø, Nikiski, and Nagasaki. Four data-driven model classes are considered: (i) linear parametric models, represented by autoregressive models with exogenous variables; (ii) functional parametric models, based on functional autoregressive models with exogenous variables; (iii) semiparametric and nonlinear models, including generalized additive models and autoregressive neural networks; and (iv) a semi-functional non-standard k-nearest-neighbours approach combining similarity in recent non-astronomical trajectories and meteorological conditions. Results reveal that hybrid models reduce forecast errors by 52.9-54.9% on average relative to HA. The generalized additive model is the most competitive across locations, while k-nearest neighbours performs best at Saint-Malo and the autoregressive model with exogenous variables is favoured in Nagasaki. For Venice, an economic decision-making case study assesses the operational use of sea-level forecasts in managing the MoSE flood-barrier system.

stat.AP↗

Assessing the impact of climate change and rising temperatures on life insurance portfolios

Climate change may materially affect long-term life insurance liabilities by altering both the level and seasonal pattern of mortality. This article develops a multi-population mortality framework that combines a Hermite spline model with a distributed lag non-linear model to capture age-, region-, and temperature-specific mortality effects. We apply the framework to mortality and temperature data from 15 Spanish NUTS-2 regions and project future mortality under three shared socioeconomic pathway (SSP) scenarios. We then assess the implications for a hypothetical whole life insurance portfolio through expected death-benefit payments and portfolio profit and loss. The results reveal an important seasonal offset: warmer conditions reduce expected payoffs during winter periods but increase them during summer periods, with these effects becoming more pronounced under more severe climate scenarios and for policies issued in later years. Over longer horizons, adverse summer mortality effects become increasingly important. The portfolio analysis further shows that climate-related mortality risk can materially increase the dispersion and downside risk of portfolio outcomes, particularly under SSP5-8.5. These findings highlight the importance of incorporating temperature-related mortality effects into long-term life insurance liability projections and risk assessment.

stat.AP↗