GptGet
Features PaperForge Apps Papers Blog Contact AI Chat 中文
Sort: Latest Popular Citations
All Artificial Intelligence Computation and Language Computer Vision Information Retrieval Machine Learning Machine Learning (Stats) Neural and Evolutionary Computing Robotics
cs.RO 2604.03181

SpatialVAM:Spatial-Aware Multi-View Video Diffusion as a Data-Efficient Robot Policy

SpatialVAM achieves data-efficient robot policy learning via 3D video diffusion, improving Meta-World success rates by 22%.

Peiyan Li, Yixiang Chen, Yuan Xu et al.

2026-04-04 43
cs.RO 2604.03037

ARM: Advantage Reward Modeling for Long-Horizon Manipulation

Advantage Reward Modeling (ARM) uses tri-state labels and a MIMO Transformer to estimate relative advantage, achieving 99.4% success in long-horizon towel-folding tasks.

Yiming Mao, Zixi Yu, Weixin Mao et al.

2026-04-03 44
cs.AI 2604.02971

InfoSeeker: A Scalable Hierarchical Parallel Agent Framework for Web Information Seeking

InfoSeeker uses Host–Manager–Worker parallelism, reaching 8.38% WideSearch success and 3–5× faster execution.

Ka Yiu Lee, Yuxuan Huang, Zhiyuan He et al.

2026-04-03 34
cs.CL 2604.02795

Rubrics to Tokens: Bridging Response-level Rubrics and Token-level Rewards in Instruction Following Tasks

RTT framework maps response-level scores to token rewards via a Token Relevance Discriminator, improving instruction-following accuracy by 2.5% over baselines.

Tianze Xu, Yanzhao Zheng, Pengrui Lu et al.

2026-04-03 46
q-fin.CP 2604.14199

PolyBench: Benchmarking LLM Forecasting and Trading Capabilities on Live Prediction Market Data

PolyBench combines 38,666 binary prediction markets, order book states, and news streams to evaluate 7 LLMs, with only MiMo-V2-Flash and Gemini-3-Flash achieving positive returns.

Pu Cheng, Juncheng Liu, Yunshen Long

2026-04-03 8 citations 44
cs.CL 2604.02699

Trivial Vocabulary Bans Improve LLM Reasoning More Than Deep Linguistic Constraints

Trivial vocabulary bans (e.g., 'very', 'just') outperform deep linguistic constraints like E-Prime in improving LLM reasoning, due to output regularization effects.

Rodney Jehu-Appiah

2026-04-03 57
cs.CL 2604.02637

Train Yourself as an LLM: Exploring Effects of AI Literacy on Persuasion via Role-playing LLM Training

LLMimic uses role-playing LLM training to enhance AI literacy, reduce persuasion success, and improve social responsibility.

Qihui Fan, Min Ge, Chenyan Jia et al.

2026-04-03 41
cs.CV 2604.02467

VERTIGO: Visual Preference Optimization for Cinematic Camera Trajectory Generation

VERTIGO optimizes visual preference, reducing off-screen rate to 0% and enhancing shot quality.

Mengtian Li, Yuwei Lu, Feifei Li et al.

2026-04-03 24
cs.IR 2604.02211

Multi-Agent Video Recommenders: Evolution, Patterns, and Open Challenges

Multi-agent video recommenders with LLMs enhance precision and explainability.

Srivaths Ranganathan, Abhishek Dharmaratnakar, Anushree Sinha et al.

2026-04-03 26
cs.CL 2605.20191

Shiny Stories, Hidden Struggles: Investigating the Representation of Disability Through the Lens of LLMs

Using social media simulation, analysis reveals LLMs over-idealize disability, reinforcing biases and stereotypes.

Marco Bombieri, Simone Paolo Ponzetto, Marco Rospocher

2026-04-02 55
cs.CV 2604.01764

Hidden Meanings in Plain Sight: RebusBench for Evaluating Cognitive Visual Reasoning

RebusBench evaluates LVLMs' cognitive visual reasoning, performance below 10% exact match.

Seyed Amir Kasaei, Arash Marioriyad, Mahbod Khaleti et al.

2026-04-02 18
cs.CV 2604.01715

SteerFlow: Steering Rectified Flows for Faithful Inversion-Based Image Editing

SteerFlow introduces fixed-point and trajectory interpolation techniques to improve faithful inversion-based image editing, outperforming existing methods in source preservation.

Thinh Dao, Zhen Wang, Kien T. Pham et al.

2026-04-02 42
cs.CL 2604.01707

Memory in the LLM Era: Modular Architectures and Strategies in a Unified Framework

Unified modular framework for agent memory; fusion model outperforms SOTA in long-term tasks with 8-15% improvements.

Yanchen Wu, Tenghui Lin, Yingli Zhou et al.

2026-04-02 31
cs.CV 2604.01561

ReFlow: Self-correction Motion Learning for Dynamic Scene Reconstruction

ReFlow employs self-correction flow matching for monocular 4D scene reconstruction, surpassing existing methods with no external motion guidance, achieving PSNR of 27.65dB.

Yanzhe Liang, Ruijie Zhu, Hanzhi Chang et al.

2026-04-02 54
cs.DB 2604.06231

Automating Database-Native Function Code Synthesis with LLMs

DBCooker leverages LLMs to automate database native function synthesis, achieving 34.55% higher accuracy than state-of-the-art methods.

Wei Zhou, Xuanhe Zhou, Qikang He et al.

2026-04-02 3 citations 40
cs.CL 2604.01504

Magic, Madness, Heaven, Sin: LLM Output Diversity is Everything, Everywhere, All at Once

Proposes the Magic-Madness-Heaven-Sin framework, categorizing LLM output diversity by task goals, revealing trade-offs across contexts.

Harnoor Dhingra

2026-04-02 58
cs.CR 2604.01444

Cooking Up Risks: Benchmarking and Reducing Food Safety Risks in Large Language Models

Introduced FoodGuardBench to evaluate LLMs' food safety, revealing three major vulnerabilities.

Weidi Luo, Xiaofei Wen, Tenghao Huang et al.

2026-04-02 22
cs.LG 2604.01430

Improving Latent Generalization Using Test-time Compute

Enhancing latent generalization using test-time compute with reinforcement learning for long chain-of-thoughts.

Arslan Chaudhry, Sridhar Thiagarajan, Andrew Lampinen

2026-04-02 5
cs.CR 2604.01346

Safety, Security, and Cognitive Risks in World Models

Study reveals GRU-based RSSM's safety risks under adversarial attacks, with a 59.5% reward reduction.

Manoj Parmar

2026-04-02 24
cs.AI 2604.01221

HippoCamp: Benchmarking Contextual Agents on Personal Computers

HippoCamp benchmarks multimodal file management agents, revealing limitations in user environments with top accuracy only 48.3%.

Zhe Yang, Shulin Tian, Kairui Hu et al.

2026-04-02 295
Prev 1 ... 171 172 173 174 175 176 177 ... 573 Next

© 2026 GptGet.net - Paper Insights Platform

Paper List Submit Paper Help GptGet Home