GptGet
Features PaperForge Apps Papers Blog Contact AI Chat 中文
Sort: Latest Popular Citations
All Artificial Intelligence Computation and Language Computer Vision Information Retrieval Machine Learning Machine Learning (Stats) Neural and Evolutionary Computing Robotics
cs.AI 2601.21754

Language-based Trial and Error Falls Behind in the Era of Experience

SCOUT framework decouples exploration and exploitation, enabling Qwen2.5-3B-Instruct to score 0.86 on unseen tasks, saving 60% GPU hours.

Haoyu Wang, Guozheng Ma, Shugang Cui et al.

2026-01-29 29
cs.AI 2601.21526

KAPSO: A Knowledge-grounded framework for Autonomous Program Synthesis and Optimization

KAPSO employs a knowledge-grounded long-horizon optimization framework integrating git isolation, structured knowledge graphs, and episodic memory to enhance autonomous program synthesis.

Alireza Nadafian, Alireza Mohammadshahi, Majid Yazdani

2026-01-29 8 citations 48
cs.AI 2601.21181

MAD: Modality-Adaptive Decoding for Mitigating Cross-Modal Hallucinations in Multimodal Large Language Models

MAD method reduces cross-modal hallucinations by adaptive decoding, achieving 7.8% and 2.0% improvements on CMM and AVHBench.

Sangyun Chung, Se Yeon Kim, Youngchae Chee et al.

2026-01-29 3
cs.AI 2601.20352

AMA: Adaptive Memory via Multi-Agent Collaboration

AMA framework achieves adaptive memory via multi-agent collaboration, reducing token consumption by 80%.

Weiquan Huang, Zixuan Wang, Hehai Lin et al.

2026-01-28 0
cs.AI 2601.19402

PROTEUS: SLA-Aware Routing via Lagrangian RL for Multi-LLM Serving Systems

PROTEUS uses Lagrangian RL for SLA-aware multi-LLM routing, achieving 94.0% accuracy and 89.8% cost savings.

Amit Singh Bhatti, Vishal Vaddina, Dagnachew Birru

2026-01-27 19
cs.AI 2601.18226

Yunjue Agent Tech Report: A Fully Reproducible, Zero-Start In-Situ Self-Evolving Agent System for Open-Ended Tasks

Yunjue Agent achieves zero-start adaptability for open tasks via parallel batch evolution and tool optimization, outperforming baselines.

Haotian Li, Shijun Yang, Weizhen Qi et al.

2026-01-26 31
cs.AI 2601.16344

DSGym: A Holistic Framework for Evaluating and Training Data Science Agents

Proposes DSGym, a modular framework for end-to-end evaluation and training of data science agents, covering diverse scientific and predictive tasks.

Fan Nie, Junlin Wang, Harper Hua et al.

2026-01-23 26
cs.AI 2601.15876

EvoCUA: Evolving Computer Use Agents via Learning from Scalable Synthetic Experience

EvoCUA employs a self-evolving cycle combining verifiable synthetic task generation, large-scale asynchronous interaction, and iterative policy optimization, achieving 56.7% success on OSWorld.

Taofeng Xue, Chong Peng, Mianqiu Huang et al.

2026-01-22 35 citations 45
cs.AI 2601.15737

PhysProver: Advancing Automatic Theorem Proving for Physics

PhysProver combines RLVR and PhysLeanData, enhancing physics theorem proving by 2.4%.

Hanning Zhang, Ruida Wang, Rui Pan et al.

2026-01-22 0
cs.AI 2601.11100

ReCreate: Reasoning and Creating Domain Agents Driven by Experience

ReCreate employs experience-driven, interaction-based scaffold updates to automate domain agent creation, outperforming black-box methods with 5%+ gains.

Zhezheng Hao, Hong Wang, Jian Luo et al.

2026-01-16 39
cs.AI 2601.10413

LADFA: A Framework of Using Large Language Models and Retrieval-Augmented Generation for Personal Data Flow Analysis in Privacy Policies

LADFA combines LLMs and RAG to analyze personal data flows in privacy policies.

Haiyue Yuan, Nikolay Matyunin, Ali Raza et al.

2026-01-15 37
cs.AI 2601.10306

Evidence-Augmented Policy Optimization with Reward Co-Evolution for Long-Context Reasoning

EAPO enhances long-context reasoning with reward co-evolution, achieving significant performance gains.

Xin Guan, Zijian Li, Shen Huang et al.

2026-01-15 1
cs.AI 2601.08679

PersonaDual: Balancing Personalization and Objectivity via Adaptive Reasoning

PersonaDual balances personalization and objectivity via adaptive reasoning, improving accuracy by 3%.

Xiaoyou Liu, Xinyi Mou, Shengbin Yue et al.

2026-01-14 1
cs.AI 2601.08444

Beyond Linearization: Attributed Table Graphs for Table Reasoning

Proposes TabGR using Attributed Table Graph and QG-PPR to improve table reasoning accuracy by 9.7%.

Yuxiang Wang, Junhao Gan, Shengxiang Gao et al.

2026-01-13 39
cs.AI 2601.07477

JudgeFlow: Agentic Workflow Optimization via Block Judge

JudgeFlow uses block responsibility scores to optimize LLM workflows, improving efficiency and interpretability.

Zihan Ma, Zhikai Zhao, Chuanbo Hua et al.

2026-01-12 37
cs.AI 2601.07226

Lost in the Noise: How Reasoning Models Fail with Contextual Distractors

Introduced NoisyBench benchmark; state-of-the-art models drop up to 80% in noisy environments; RARE improves robustness.

Seongyun Lee, Yongrae Jo, Minju Seo et al.

2026-01-12 46
cs.AI 2601.07055

Dr. Zero: Self-Evolving Search Agents without Training Data

Dr. Zero achieves search capabilities comparable to supervised learning through self-evolution and HRPO without training data.

Zhenrui Yue, Kartikeya Upasani, Xianjun Yang et al.

2026-01-12 30
cs.AI 2601.06338

Circuit Mechanisms for Spatial Relation Generation in Diffusion Transformers

Mechanistic analysis of DiT reveals two distinct circuits for spatial relation generation: stepwise attention heads with random embeddings and integrated fusion with T5.

Binxu Wang, Jingxuan Fan, Xu Pan

2026-01-10 28
cs.AI 2601.04888

SmartSearch: Process Reward-Guided Query Refinement for Search Agents

SmartSearch enhances search agent query quality via process rewards and query refinement, significantly boosting retrieval efficiency.

Tongyu Wen, Guanting Dong, Zhicheng Dou

2026-01-08 27
cs.AI 2601.04426

XGrammar-2: Dynamic and Efficient Structured Generation Engine for Agentic LLMs

XGrammar-2 introduces TagDispatch and cross-grammar cache, achieving 6× faster structured generation for dynamic agent workloads.

Linzhang Li, Yixin Dong, Guanjie Wang et al.

2026-01-08 5 citations 80
Prev 1 ... 24 25 26 27 28 29 30 ... 44 Next

© 2026 GptGet.net - Paper Insights Platform

Paper List Submit Paper Help GptGet Home