cs.CL 2602.04289

Proxy Compression for Language Modeling

Proxy compression enhances language model efficiency, significantly outperforming byte-level baselines.

Lin Zheng, Xinyu Li, Qian Liu et al.

2026-02-04 39
cs.CL 2601.21968

OVD: On-policy Verbal Distillation

OVD: trajectory matching with verbal scores reduces memory, improves Web QA and math reasoning by up to 25.7%.

Jing Xiong, Hui Shen, Shansan Gong et al.

2026-01-30 44
cs.CL 2601.21744

Temporal Guidance for Large Language Models

Introduces Temporal Guidance (TeGu), leveraging temporal contrast to enhance LLM generation quality, achieving a 3.03% improvement on GSM8K.

Hong-Kai Zheng, Piji Li

2026-01-29 35
cs.CL 2601.20757

Persona Prompting as a Lens on LLM Social Reasoning

This paper introduces Persona Prompting (PP) to analyze its impact on social reasoning in LLMs, focusing on bias, rationale quality, and task performance using hate speech datasets.

Jing Yang, Moritz Hechtbauer, Elisabeth Khalilov et al.

2026-01-29 58