arXiv · 2609.16000
Using Codebooks to Detect Cybercrime Topics in Text Narratives
Abstract
In the United States, management of cybercrime-related consumer complaints increasingly falls on state and city governments given de-staffing of federal agencies. AI, and in particular, large language models (LLMs), shows promise for detecting cybercrime in text complaints, but often via specialized models that local governments are not resourced to develop and maintain. We present an LLM prompting method that uses codebooks from qualitative cybercrime research to detect cybercrime topics in consumer narratives. For two cybercrime topics, impostor scams and identity theft, we demonstrate the method achieves high precision and recall across multiple runs of 5 models in the Gemini and GPT model families. This strategy suggests a path for resource-constrained organizations, like many local governments, to leverage frontier models to support community safety.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Shufan Chai, Liangliang Sun, Jessica Staddon. 2026-07-23. Using Codebooks to Detect Cybercrime Topics in Text Narratives. https://arxiv.org/abs/2609.16000
Cite the original work for its findings. Save a collection to share your selection of sources.