Polite but Misaligned: Evaluating LLM Politeness Judgments Against Human Pragmatic Norms
arXiv:2609.29001v1 Announce Type: new Abstract: Despite strong performance on standard benchmarks, it remains unclear whether large language models (LLMs) evaluate social pragmatics in ways that align with human judgments. We evaluate LLM politeness judgments using two English-language datasets wit…
Read the full story at arXiv cs.CL ↗
Timeline · 1 report
- 2026-09-25 04:00 · arXiv cs.CL
Polite but Misaligned: Evaluating LLM Politeness Judgments Against Human Pragmatic Norms