arXiv ScienceSearch

arXiv subjects

Scott Johnson

Publications and source records attributed to Scott Johnson.

2 recordsLinked to original sources

Utility-Based Reinforcement Learning: Unifying Single-objective and Multi-objective Reinforcement Learning

Research in multi-objective reinforcement learning (MORL) has introduced the utility-based paradigm, which makes use of both environmental rewards and a function that defines the utility derived by the user from those rewards. In this paper we extend this paradigm to the context of single-objective reinforcement learning (RL), and outline multiple potential benefits including the ability to perform multi-policy learning across tasks relating to uncertain objectives, risk-aware RL, discounting, and safe RL. We also examine the algorithmic implications of adopting a utility-based approach.

cs.LG

Observation of Yamaji magic angles in bismuth surfaces

Bismuth consist of bismuth bilayers that are two-dimensional topological insulators and correspondingly, the surface is an array of edge states. Moreover, topological models, including second order topologic order, predict an interlayer electrical coupling mediated by hinge states. Here we report that angle dependent magnetoresistance measurements of small diameter single-crystal bismuth nanowires exhibit the sequence of magnetoresistance (MR) peaks at Yamaji magic angles and a peak for B//bilayer, indicating coherent transport between layers, and showing that the Fermi surface of surface electrons is a warped cylinder. The MR peaks are associated with magnetic field induced flat bands that are reminiscent of the well-known flat bands in bilayer graphene. Coherent transport across layers is interpreted in term of transport by topological hinge states.

cond-mat.mes-hall