arXiv · 2610.10240
Benchmarking Adult Addressee Classification Across Child- and Adult-Directed Speech Datasets
Abstract
In this work, we present a comprehensive analysis of classification performance for distinguishing child-directed speech (CDS) from adult-directed speech (ADS) using speech data from corpora containing natural in-lab and in-the-wild CDS and ADS. We establish classification benchmarks for these datasets using self-supervised learning (SSL) representations, along with a range of time-pooled representations that go beyond first- and second-order statistics by incorporating cross-channel covariances in high-dimensional embeddings. In addition, we probe these representations to examine how different pooling methods capture prosodic information using linear probes. Overall, SSL-based representations prove particularly effective, achieving the best performance, while different pooling methods offer complementary advantages for the task.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Sofoklis Kakouros, Daniil Kocharov, Okko Räsänen. 2026-10-07. Benchmarking Adult Addressee Classification Across Child- and Adult-Directed Speech Datasets. https://arxiv.org/abs/2610.10240
Cite the original work for its findings. Save a collection to share your selection of sources.