Are You Sure You're Sure? On the Impact of Instruction Tuning on Confidence and Lexical Diversity Paper • 2608.13430 • Published 8 days ago • 12
Alaya-EVOKE: From Linear-Scaling Supervision to Endless World Paper • 2608.13546 • Published 8 days ago • 130
HatePrototypes: Interpretable and Transferable Representations for Implicit and Explicit Hate Speech Detection Paper • 2511.06391 • Published Apr 5 • 1
Fair-GPTQ: Bias-Aware Quantization for Large Language Models Paper • 2509.15206 • Published Sep 18, 2025 • 1
Histoires Morales: A French Dataset for Assessing Moral Alignment Paper • 2501.17117 • Published Jan 28, 2025 • 5