arXiv Science⌕ Search

arXiv subjects

Ethan Z. Rong

Publications and source records attributed to Ethan Z. Rong.

3 recordsLinked to original sources

You Really Didn't Get That? Benchmarking Social Pragmatic Inference for Indirect and Playful Chinese Online Comments

Chinese online comments often convey social meaning through indirect and playful language that is hard to interpret without context. Existing evaluations largely organize items around predefined phenomena or controlled pragmatic categories, leaving open whether models can distinguish plausible readings of what a naturally occurring comment is doing in a particular exchange. We introduce a benchmark for evaluating whether LLMs can recover such situated pragmatic meanings. From more than 200,000 public Chinese social media interaction records, we construct 4,735 human-validated diagnostic items, each pairing a target comment with reconstructed preceding context and plausible misreadings. We evaluate eight LLMs as both question writers and solvers in a cross-writer setting. The task is challenging: the strongest model achieves 81.42% leave-writer-out accuracy. Across all eight models, the mean leave-writer-out accuracy is 68.70% while human accuracy was 90.8%. Case analysis shows that models often recognize broad irony or playfulness while misidentifying the mechanism or interactional move.

cs.CL↗

Multi-Agent Comedy Club: Investigating Community Discussion Effects on LLM Humor Generation

Prior work has explored multi-turn interaction and feedback for LLM writing, but evaluations still largely center on prompts and localized feedback, leaving persistent public reception in online communities underexamined. We test whether broadcast community discussion improves stand-up comedy writing in a controlled multi-agent sandbox: in the discussion condition, critic and audience threads are recorded, filtered, stored as social memory, and later retrieved to condition subsequent generations, whereas the baseline omits discussion. Across 50 rounds (250 paired monologues) judged by five expert annotators using A/B preference and a 15-item rubric, discussion wins 75.6% of instances and improves Craft/Clarity (Δ = 0.440) and Social Response (Δ = 0.422), with occasional increases in aggressive humor.

cs.CL↗

"It Feels Like Being Locked in A Cage": Understanding Blind or Low Vision Streamers' Perceptions of Content Curation Algorithms

Blind or low vision (BLV) people were recently reported to be live streamers on the online platforms that employed content curation algorithms. Recent research uncovered algorithm biases suppressing the content created by marginalized populations. However, little is known about the effects of the algorithms adopted by live streaming platforms on BLV streamers and how they, as a marginalized population, perceive the effects of the algorithms. We interviewed BLV streamers (N=19) of Douyin -- a popular live stream platform in China -- to understand their perceptions of algorithms, perceived challenges, and mitigation strategies. Our findings show the perceived factors contributing to disadvantages under algorithmic evaluation of BLV streamers' content (e.g., issues with filming and timely interaction with viewers) and perceived algorithmic suppression (e.g., content not amplified to sighted users but suppressed within the BLV community). Their mitigation strategies (e.g., not watching other BLV streamers' shows) tended to be passive. We discuss design considerations to design a more inclusive and fair live streaming platform.

cs.HC↗