2024
LLM-based Rewriting of Inappropriate Argumentation using Reinforcement Learning from Machine Feedback
ACL 2024long
Ensuring that online discussions are civil and productive is a major challenge for social media platforms. Such platforms usually rely both on users and on automated detection tools to flag inappropriate arguments of other users, which moderators then review. However, this kind of post-hoc moderatio…