Self-Explaining Hate Speech Detection with Moral Rationales

Published in Findings of the Association for Computational Linguistics: ACL 2026, 2026

Recommended citation: Vargas, F., Trager, J., Alves, D., Guida, M., et al. (2026). "Self-Explaining Hate Speech Detection with Moral Rationales." Findings of the Association for Computational Linguistics: ACL 2026. https://arxiv.org/abs/2601.03481