cs.LG 2510.13259

Hypernetworks for Perspectivist Adaptation

Utilizing hypernetwork and adapters architecture, this study enhances perspective adaptation in hate speech detection with fewer parameters.

Daniil Ignatev, Denis Paperno, Massimo Poesio

2025-10-15 1
cs.LG 2510.02259

Transformers Discover Molecular Structure Without Graph Priors

This study demonstrates that a standard Transformer, trained directly on Cartesian coordinates without graph priors, can achieve energy and force prediction accuracy comparable to state-of-the-art equivariant GNNs on OMol25, with faster inference.

Tobias Kreiman, Yutong Bai, Fadi Atieh et al.

2025-10-03 40
cs.LG 2510.01132

A Practitioner's Guide to Multi-turn Agentic Reinforcement Learning

This study introduces a systematic multi-turn RL framework for large language models, emphasizing environment, reward, and policy pillars, validated across TextWorld, ALFWorld, and SWE-Gym with key improvements of up to 88%.

Ruiyi Wang, Prithviraj Ammanabrolu

2025-10-02 47