GptGet
Features PaperForge Apps Papers Blog Contact AI Chat 中文
Sort: Latest Popular Citations
All Artificial Intelligence Computation and Language Computer Vision Information Retrieval Machine Learning Machine Learning (Stats) Neural and Evolutionary Computing Robotics
cs.LG 2510.08748

Conformal Risk Training: End-to-End Optimization of Conformal Risk Control

Proposes end-to-end conformal risk training extending CRC to OCE risks, improving model performance and guarantees.

Christopher Yeh, Nicolas Christianson, Adam Wierman et al.

2025-10-10 39
cs.CV 2510.08527

FlexTraj: Image-to-Video Generation with Flexible Point Trajectory Control

FlexTraj introduces a point-based trajectory control framework enabling multi-granularity, alignment-agnostic image-to-video synthesis with improved efficiency.

Zhiyuan Zhang, Can Wang, Dongdong Chen et al.

2025-10-10 48
cs.AI 2510.08511

AutoMLGen: Navigating Fine-Grained Optimization for Coding Agents

AutoMLGen integrates knowledge base and MCGS to optimize end-to-end ML pipelines, achieving 36.4% medal rate within 12 hours, outperforming baselines.

Shangheng Du, Xiangchao Yan, Dengyang Jiang et al.

2025-10-10 57
cs.CL 2510.08457

ARES: Multimodal Adaptive Reasoning via Difficulty-Aware Token-Level Entropy Shaping

ARES uses difficulty-aware window entropy shaping for adaptive multimodal reasoning, achieving superior performance across benchmarks.

Shuang Chen, Yue Guo, Yimeng Ye et al.

2025-10-10 14
cs.CV 2510.08363

Hyperspectral data augmentation with transformer-based diffusion models

Proposed a guided diffusion model with Transformer for hyperspectral data augmentation, boosting forest classification accuracy by 4.8%.

Mattia Ferrari, Lorenzo Bruzzone

2025-10-09 52
cs.LG 2510.08294

Counterfactual Identifiability via Dynamic Optimal Transport

Proposes a dynamic optimal transport-based framework for multivariate counterfactual identification, ensuring uniqueness and consistency.

Fabio De Sousa Ribeiro, Ainkaran Santhirasekaram, Ben Glocker

2025-10-09 50
cs.HC 2510.08242

Simulating Teams with LLM Agents: Interactive 2D Environments for Studying Human-AI Dynamics

VirT-Lab simulates customizable LLM teams in 2D worlds, validated through rescue alignment, ablations, a 12-person user study, and scaling tests.

Mohammed Almutairi, Charles Chiang, Haoze Guo et al.

2025-10-09 31
cs.SD 2510.08176

Leveraging Whisper Embeddings for Audio-based Lyrics Matching

Leveraging Whisper decoder embeddings for audio lyrics matching, achieving performance comparable to state-of-the-art methods.

Eleonora Mancini, Joan Serrà, Paolo Torroni et al.

2025-10-09 43
cs.SD 2510.08078

Detecting and Mitigating Insertion Hallucination in Video-to-Audio Generation

HALCON method reduces insertion hallucination in video-to-audio generation by over 50%.

Liyang Chen, Hongkai Chen, Yujun Cai et al.

2025-10-09 27
cs.LG 2510.07919

GRADE: Personalized Multi-Task Fusion via Group-relative Reinforcement Learning with Adaptive Dirichlet Exploration

GRADE framework uses group-relative policy optimization and Dirichlet exploration for personalized multi-task fusion, improving CTR by 0.595% and CVR by 1.193%.

Tingfeng Hong, Pingye Ren, Xinlong Xiao et al.

2025-10-09 43
cs.CL 2510.07877

Ready to Translate, Not to Represent? Bias and Performance Gaps in Multilingual LLMs Across Language Families and Domains

Proposes Translation Tangles framework, integrating multi-metric, multi-domain, bias detection for 24 language pairs, analyzing translation quality and biases.

Md. Faiyaz Abdullah Sayeedi, Md. Mahbub Alam, Subhey Sadi Rahman et al.

2025-10-09 55
cs.CV 2510.07721

RePainter: Empowering E-commerce Object Removal via Spatial-matting Reinforcement Learning

Repainter uses spatial-matting reinforcement learning to significantly enhance e-commerce image removal.

Zipeng Guo, Lichen Ma, Xiaolong Fu et al.

2025-10-09 26
cs.IR 2510.07644

ISMIE: A Framework to Characterize Information Seeking in Modern Information Environments

Proposed ISMIE framework models information seeking via components, variables, activities; validated in misinformation and AI content trust scenarios.

Shuoqi Sun, Danula Hettiachchi, Damiano Spina

2025-10-09 2 citations 66
cs.CL 2510.07248

Don't Adapt Small Language Models for Tools; Adapt Tool Schemas to the Models

Proposes PA-Tool, using peakedness to align tool schemas with pretrained models, boosting accuracy by 17% without retraining.

Jonggeun Lee, Woojung Song, Jongwook Han et al.

2025-10-09 55
cs.CL 2510.07230

Customer-R1: Personalized Simulation of Human Behaviors via RL-based LLM Agent in Online Shopping

Customer-R1 employs RL with explicit user personas, boosting next-action prediction accuracy from 7.32% to 39.58%, outperforming baselines.

Ziyi Wang, Yuxuan Lu, Yimeng Zhang et al.

2025-10-09 42
cs.CV 2510.07190

MV-Performer: Taming Video Diffusion Model for Faithful and Synchronized Multi-view Performer Synthesis

MV-Performer employs depth-guided video diffusion to synthesize 360° synchronized multi-view human videos from monocular input.

Yihao Zhi, Chenghong Li, Hongjie Liao et al.

2025-10-09 47
cs.LG 2510.07358

Encode, Think, Decode: Scaling test-time reasoning with recursive latent thoughts

Proposes ETD, training models to iterate over key layers during mid-training, boosting reasoning accuracy by up to 36% on benchmarks.

Yeskendir Koishekenov, Aldo Lipani, Nicola Cancedda

2025-10-08 37
cs.RO 2510.07134

TrackVLA++: Unleashing Reasoning and Memory Capabilities in VLA Models for Embodied Visual Tracking

TrackVLA++ enhances visual tracking with spatial reasoning and memory modules, achieving 5.1% and 12% improvements.

Jiahang Liu, Yunpeng Qi, Jiazhao Zhang et al.

2025-10-08 39
cs.AI 2510.07073

VRPAgent: LLM-Driven Discovery of Heuristic Operators for Vehicle Routing Problems

VRPAgent uses LLMs to generate heuristic operators, optimized via genetic algorithms, outperforming handcrafted methods on multiple VRP variants.

André Hottung, Federico Berto, Chuanbo Hua et al.

2025-10-08 48
cs.CL 2510.06727

Scaling LLM Multi-turn RL with End-to-end Summarization-based Context Management

SUPO algorithm scales LLM multi-turn RL via summarization-based context management, enhancing success rate and reducing context length.

Miao Lu, Weiwei Sun, Weihua Du et al.

2025-10-08 51
Prev 1 ... 244 245 246 247 248 249 250 ... 573 Next

© 2026 GptGet.net - Paper Insights Platform

Paper List Submit Paper Help GptGet Home