cs.DC 2510.14686

xLLM Technical Report

xLLM employs decoupled architecture with adaptive scheduling and multi-layer pipeline optimization, achieving 1.7× throughput over MindIE and 2.2× over vLLM-Ascend on Qwen models.

Tongxuan Liu, Tao Peng, Peijun Yang et al.

2025-10-16 71
cs.LG 2510.14545

Agentic Entropy-Balanced Policy Optimization

AEPO introduces dynamic entropy balancing and gradient regulation, enhancing stability and exploration in multi-turn web agent RL with only 1K samples.

Guanting Dong, Licheng Bao, Zhongyuan Wang et al.

2025-10-16 52
cs.LG 2510.14386

ASecond-Order SpikingSSM for Wearables

SHaRe-SSM excels in ultra-long sequences with 52.1x energy efficiency improvement.

Kartikay Agrawal, Abhijeet Vikram, Vedant Sharma et al.

2025-10-16 27
cs.CV 2510.13804

Generative Universal Verifier as Multimodal Meta-Reasoner

Proposes Generative Universal Verifier, trained on ViVerBench, improving visual verification by 8.3 points with OmniVerifier-7B and TTS strategies.

Xinchen Zhang, Xiaoying Zhang, Youbin Wu et al.

2025-10-16 18 citations 26
cs.LG 2510.13259

Hypernetworks for Perspectivist Adaptation

Utilizing hypernetwork and adapters architecture, this study enhances perspective adaptation in hate speech detection with fewer parameters.

Daniil Ignatev, Denis Paperno, Massimo Poesio

2025-10-15 18
cs.CL 2510.13912

AI Debaters are More Persuasive when Arguing in Alignment with Their Own Beliefs

This study shows that AI models are more persuasive when defending positions aligned with their prior beliefs, using sequential and simultaneous debate protocols, with models favoring sycophantic strategies in conflict scenarios.

María Victoria Carro, Denise Alejandra Mester, Facundo Nieto et al.

2025-10-15 38
cs.CV 2510.12764

AnyUp: Universal Feature Upsampling

AnyUp is a universal feature upsampling method that generalizes to any feature type at inference, outperforming state-of-the-art.

Thomas Wimmer, Prune Truong, Marie-Julie Rakotosaona et al.

2025-10-15 42