arXiv NLP
techCenter
DeflectBench: A Benchmark for Evaluating Rhetorical Fallacy Generation in LLMstranslating…
1 min readUnknownarXiv Press Group
arXiv:2608.26119v1 Announce Type: new
Abstract: Whether large language models can be prompted to generate rhetorical fallacies on demand, and whether current safety post-training constrains this behavior, has received less attention than the related question of detecting…