arXiv ScienceSearch

arXiv subjects

Andrew Tu

Publications and source records attributed to Andrew Tu.

2 recordsLinked to original sources

Measuring Intelligence Beyond Human Scale

How can we measure intelligence beyond human capability? Human-authored benchmarks saturate, and above human capability, examiners may not know which tasks are both hard and verifiable. We argue that this difficulty is inherent to absolute-scale evaluation and propose a new paradigm based on relative measurement in which models generate public challenges that separate other systems. Aggregating these outcomes yields an adversarial psychometric rating system that can scale with the systems being measured. We describe practical protocols that reduce incentives for private-information attacks, support judge-free adjudication, and naturally scale with agent capabilities. We instantiate the framework across verifiable and open-ended, non-verifiable domains, illustrating how model-generated evaluation can continue to measure systems beyond the human frontier.

cs.AI

Two Games on Arithmetic Functions: SALIQUANT and NONTOTIENT

We investigate the Sprague-Grundy sequences for two normal-play impartial games based on arithmetic functions, first described by Iannucci and Larsson in \cite{sum}. In each game, the set of positions is N (natural numbers). In saliquant, the options are to subtract a non-divisor. Here we obtain several nice number theoretic lemmas, a fundamental theorem, and two conjectures about the eventual density of Sprague-Grundy values. In nontotient, the only option is to subtract the number of relatively prime residues. Here are able to calculate certain Sprague-Grundy values, and start to understand an appropriate class function.

math.NT