Submitted by Jesse Cresswell 5 Unifying Conformal Language Tasks with In-Context Ensembles Layer 6 AI 2 1
Submitted by Jesse Cresswell 9 A Gradient Perspective on RLVR Stability and Winner Advantage Policy Optimization Layer 6 AI 1 2
Submitted by Joseph Tang 17 RankJudge: A Multi-Turn LLM-as-a-Judge Synthetic Benchmark Generator Layer 6 AI 9 3
Submitted by Noel Vouitsis 11 Inconsistencies In Consistency Models: Better ODE Solving Does Not Imply Better Samples Layer 6 AI 4 2