cs.LG 2506.00700

Central Path Proximal Policy Optimization

C3PO introduces a central path-inspired modification to PPO, improving constraint satisfaction and reward performance in constrained RL.

Nikola Milosevic, Johannes Müller, Nico Scherf

2025-06-01 27
cs.CL 2506.00400

Scaling Textual Gradients via Sampling-Based Momentum

Introduces TSGD-M, a sampling momentum method that scales textual gradients effectively within limited context windows, improving prompt optimization performance.

Zixin Ding, Junyuan Hong, Zhan Shi et al.

2025-05-31 50
cs.LG 2505.24492

Object Centric Concept Bottlenecks

Object-Centric Concept Bottlenecks (OCB) integrates pretrained object detection and concept discovery to enhance performance and interpretability in complex visual tasks, achieving 68.84% accuracy on COCOLogic.

David Steinmann, Wolfgang Stammer, Antonia Wüst et al.

2025-05-30 12 citations 53