arXiv · 2609.37343
Classification as Search Infrastructure: How Category Creation, Addition and Cleanup Shape Knowledge Retrieval
Abstract
Classification systems shape how searchers find relevant objects. We study how revisions to classification architecture affect retrieval by changing the routes through which objects enter consideration. We theorize that creating a cross-cutting category can reduce classificatory distance and translation costs, whereas adding routes to an established system can increase route-selection and updating costs; category cleanup can restore contrast. Empirically, we exploit the 2013 creation and 2020 revision of Y02/Y04 green-technology codes in the Cooperative Patent Classification. We match treated patents to unchanged green patents within technology-cohort cells and estimate examiner- and applicant-specific effects on first-citation timing and citation volume. Category creation is associated with a 25% higher examiner first-citation hazard. In the mature system, additions are associated with 13-15% lower examiner hazards, while cleanup is associated with a 14% higher hazard and 12% higher citation volume. Effects are generally stronger for examiners, whose search task is more tightly coupled to patent classification. Timing and volume need not move together: creation and green-domain entry shift timing without detectable changes in volume, while cleanup increases both speed and recorded reach. Classification infrastructure may therefore change when relevant objects are found even when overall recorded attention is unchanged, and greater classificatory detail need not improve search efficiency.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Kerstin Hötte, Nicolò Barbieri, Su Jung Jee. 2026-09-29. Classification as Search Infrastructure: How Category Creation, Addition and Cleanup Shape Knowledge Retrieval. https://arxiv.org/abs/2609.37343
Cite the original work for its findings. Save a collection to share your selection of sources.