arXiv · 2602.01618
SEA-Guard: Culturally Grounded Multilingual Safeguard for Southeast Asia
Abstract
Culturally aware safeguards are crucial for AI alignment in real-world settings, where safety extends beyond common sense and encompasses diverse local values, norms, and region-specific regulations. However, building large-scale, culturally grounded datasets is challenging due to limited resources and a scarcity of native annotators. Consequently, many safeguard models rely on machine translation of English datasets, often missing regional and cultural nuances. We present a novel agentic data-generation framework to scalably create authentic, region-specific safety datasets for Southeast Asia (SEA). On this foundation, we introduce the SEA-Guard family, the first multilingual safeguard models grounded in SEA cultural contexts. Evaluated across multiple benchmarks and cultural variants, SEA-Guard consistently outperforms existing safeguards at detecting regionally sensitive or harmful content while maintaining strong general safety performance.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Panuthep Tasawong, Jian Gang Ngui, Alham Fikri Aji, Trevor Cohn, Peerat Limkonchotiwat. 2026-02-02. SEA-Guard: Culturally Grounded Multilingual Safeguard for Southeast Asia. https://arxiv.org/abs/2602.01618
Cite the original work for its findings. Save a collection to share your selection of sources.