Evaluating Online Moderation Via LLM-Powered Counterfactual Simulations
通过LLM赋能的反事实模拟评估在线 moderation
AI总结 本文提出一种基于LLM的反事实模拟方法,用于评估在线社交网络的 moderation 策略,揭示了社交传染现象及个性化策略的有效性。
Comments Accepted for publication at AAAI Conference on Artificial Intelligence 2026
Journal ref Proceedings of the AAAI Conference on Artificial Intelligence 2026