math.ST 2504.01781

Proper scoring rules for estimation and forecast evaluation

This paper reviews the mathematical foundations of proper scoring rules, emphasizing their role in probabilistic estimation and forecast evaluation, with theoretical and practical insights.

Kartik Waghmare, Johanna Ziegel

2025-04-02 54
cs.CL 2504.00050

JudgeLRM: Large Reasoning Models as a Judge

JudgeLRM employs reinforcement learning with outcome-driven rewards to surpass SFT models, with 7B/8B models exceeding GPT-4 in F1 scores.

Nuo Chen, Zhiyuan Hu, Qingyun Zou et al.

2025-03-31 36