Post-Training Language Models for Gold-Medal Performance in Coding Competitions Paper • 2609.02849 • Published 3 days ago • 9
ChildVox: A Speech, Audio, and Large Audio-Language Model Benchmark in Understanding and Characterizing Sound across Childhood Paper • 2605.29257 • Published May 28 • 12
RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Paper • 2605.21195 • Published May 20 • 20
EvalVerse: Pipeline-Aware and Expert-Calibrated Benchmarking for Professional Cinematic Video Generation Paper • 2605.23271 • Published May 22 • 83