GptGet
Features PaperForge Apps Papers Blog Contact AI Chat 中文
Sort: Latest Popular Citations
All Artificial Intelligence Computation and Language Computer Vision Information Retrieval Machine Learning Machine Learning (Stats) Neural and Evolutionary Computing Robotics
cs.HC 2506.10762

Integrating Large Language Models into Text Animation: An Intelligent Editing System with Inline and Chat Interaction

Proposed an LLM-based text animation editing system with inline suggestions and chat interaction, reducing non-professional users' editing time by 30%.

Bao Zhang, Zihan Li, Zhenglei Liu et al.

2025-06-12 39
cs.CL 2506.10728

Beyond True or False: Retrieval-Augmented Hierarchical Analysis of Nuanced Claims

ClaimSpect uses retrieval-augmented hierarchical analysis to decompose nuanced claims, identifying multi-angle evidence and perspectives.

Priyanka Kargupta, Runchu Tian, Jiawei Han

2025-06-12 31
cs.CL 2506.10641

Spelling-out is not Straightforward: LLMs' Capability of Tokenization from Token to Characters

Study reveals complexities in LLMs' character-level spelling tasks, showing reliance on mid-to-high Transformer layers for reconstructing character information.

Tatsuya Hiraoka, Kentaro Inui

2025-06-12 7
cs.RO 2506.12095

DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion

DoublyAware decomposes planning and policy uncertainties, improving sample efficiency and robustness in humanoid locomotion via conformal prediction and group-relative policy constraints.

Khang Nguyen, An T. Le, Jan Peters et al.

2025-06-12 38
cs.LO 2506.10558

StepProof: Step-by-step verification of natural language mathematical proofs

StepProof enhances verification success rates and efficiency through step-by-step validation of natural language mathematical proofs.

Xiaolin Hu, Qinghua Zhou, Bogdan Grechuk et al.

2025-06-12 2
cs.AI 2507.00008

DiMo-GUI: Advancing Test-time Scaling in GUI Grounding via Modality-Aware Visual Reasoning

DiMo-GUI combines modality decoupling and dynamic zooming to improve GUI grounding accuracy without extra training, boosting performance over baseline models.

Hang Wu, Hongkai Chen, Yujun Cai et al.

2025-06-12 52
cs.CV 2506.09989

Hearing Hands: Generating Sounds from Physical Interactions in 3D Scenes

Rectified flow turns 3D hand trajectories into interaction sounds that are often indistinguishable from real audio.

Yiming Dou, Wonseok Oh, Yuqing Luo et al.

2025-06-12 33
cs.LG 2506.10054

Uni-DPO: A Unified Paradigm for Dynamic Preference Optimization of LLMs

Uni-DPO enhances LLM performance by dynamically reweighting samples, outperforming Claude 3 Opus by 6.7 points on Arena-Hard.

Shangpin Peng, Weinong Wang, Zhuotao Tian et al.

2025-06-12 21
cs.CV 2506.09980

Efficient Part-level 3D Object Generation via Dual Volume Packing

Proposes dual volume packing for efficient part-level 3D generation from a single image, avoiding segmentation priors.

Jiaxiang Tang, Ruijie Lu, Zhaoshuo Li et al.

2025-06-12 39
cs.CL 2506.09902

PersonaLens: A Benchmark for Personalization Evaluation in Conversational AI Assistants

PersonaLens uses multi-task, multi-domain user profiles and two LLM agents to evaluate personalization in task-oriented dialogue systems.

Zheng Zhao, Clara Vania, Subhradeep Kayal et al.

2025-06-12 40
cs.RO 2506.09800

Reinforced Refinement with Self-Aware Expansion for End-to-End Autonomous Driving

Proposes R2SE framework combining hard-case identification and reinforcement fine-tuning to enhance end-to-end autonomous driving.

Haochen Liu, Tianyu Li, Haohan Yang et al.

2025-06-11 36
cs.LG 2506.09477

On a few pitfalls in KL divergence gradient estimation for RL

Identifies three common pitfalls in KL gradient estimation for RL, emphasizing the correct full-sequence approach.

Yunhao Tang, Rémi Munos

2025-06-11 26 citations 31
math.NA 2506.09394

Subspace-constrained randomized coordinate descent for linear systems with good low-rank matrix approximations

Introduces Subspace-Constrained Randomized Coordinate Descent (SC-RCD) leveraging Nyström low-rank approximation, robust against spectral outliers, for large-scale positive semidefinite linear systems.

Jackie Lok, Elizaveta Rebrova

2025-06-11 4 citations 46
cs.LG 2506.09373

LPO: Towards Accurate GUI Agent Interaction via Location Preference Optimization

LPO combines information entropy and dynamic distance rewards to enhance GUI interaction accuracy, achieving state-of-the-art results.

Jiaqi Tang, Yu Xia, Yi-Feng Wu et al.

2025-06-11 63
cs.LG 2506.09026

e3: Learning to Explore Enables Extrapolation of Test-Time Compute for LLMs

e3 trains models with chain-based exploration, enabling extrapolation to twice the training inference length, significantly improving reasoning performance beyond training limits.

Amrith Setlur, Matthew Y. R. Yang, Charlie Snell et al.

2025-06-11 36
cs.CV 2506.09022

Do Multiple Instance Learning Models Transfer?

MIL models show superior transfer across organs, boosting performance by 9.8%.

Daniel Shao, Richard J. Chen, Andrew H. Song et al.

2025-06-11 7
eess.SP 2506.08807

Confidence Boosts Trust-Based Resilience in Cooperative Multi-Robot Systems

Proposes a trust-based resilient multi-robot protocol with dynamic λt to ensure robustness against malicious robots, validated through theoretical proofs and experiments.

Luca Ballotta, Áron Vékássy, Stephanie Gil et al.

2025-06-10 58
cs.SD 2506.08570

Auto-Regressive vs Flow-Matching: a Comparative Study of Modeling Paradigms for Text-to-Music Generation

Comparative study of auto-regressive decoding and conditional flow matching for text-to-music generation.

Or Tal, Felix Kreuk, Yossi Adi

2025-06-10 2
cs.CV 2506.08015

4DGT: Learning a 4D Gaussian Transformer Using Real-World Monocular Videos

Proposes 4D Gaussian Transformer (4DGT) for real-time dynamic scene reconstruction from monocular videos, significantly improving speed and accuracy.

Zhen Xu, Zhengqin Li, Zhao Dong et al.

2025-06-10 43 citations 41
cs.CV 2506.08009

Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion

Self Forcing introduces autoregressive training with distribution matching, enabling real-time video generation with high quality and low latency, outperforming traditional slow diffusion models.

Xun Huang, Zhengqi Li, Guande He et al.

2025-06-10 533 citations 42
Prev 1 ... 274 275 276 277 278 279 280 ... 573 Next

© 2026 GptGet.net - Paper Insights Platform

Paper List Submit Paper Help GptGet Home