GptGet
Features PaperForge Apps Papers Blog Contact AI Chat 中文
Sort: Latest Popular Citations
All Artificial Intelligence Computation and Language Computer Vision Information Retrieval Machine Learning Machine Learning (Stats) Neural and Evolutionary Computing Robotics
cs.CV 2510.00818

PhraseStereo: The First Open-Vocabulary Stereo Image Segmentation Dataset

Introduces PhraseStereo, a stereo dataset for open-vocabulary phrase segmentation leveraging depth cues, with SSIM=0.601 and LPIPS=0.352 at optimal scale.

Thomas Campagnolo, Ezio Malis, Philippe Martinet et al.

2025-10-01 37
cs.AI 2510.00732

EvolProver: Advancing Automated Theorem Proving by Evolving Formalized Problems via Symmetry and Difficulty

EvolProver enhances automated theorem proving by evolving problems via symmetry and difficulty, achieving 53.8% pass@32.

Yuchen Tian, Ruiyuan Huang, Xuanwu Wang et al.

2025-10-01 22
cs.CV 2510.00634

LAKAN: Landmark-assisted Adaptive Kolmogorov-Arnold Network for Face Forgery Detection

LAKAN integrates facial landmarks to dynamically modulate Kolmogorov-Arnold networks, boosting deepfake detection accuracy.

Jiayao Jiang, Bin Liu, Qi Chu et al.

2025-10-01 50
cs.CL 2510.00496

Agent-ScanKit: Unraveling Memory and Reasoning of Multimodal Agents via Sensitivity Perturbations

Agent-ScanKit reveals multimodal agents' memory and reasoning via sensitivity perturbations, showing memory often outweighs reasoning.

Pengzhou Cheng, Lingzhong Dong, Zeng Wu et al.

2025-10-01 7
cs.AI 2510.00492

Rethinking Reward Models for Multi-Domain Test-Time Scaling

This paper compares four reward models (dORM, dPRM, gORM, gPRM) across 14 multi-domain tasks, finding gORM most robust and effective.

Dong Bok Lee, Seanie Lee, Sangwoo Park et al.

2025-10-01 36
cs.LG 2510.00399

How Can Mamba Learn In Context with Outliers and Generalize Provably?

This paper provides the first theoretical analysis of one-layer Mamba's training dynamics and robustness to outliers in in-context learning.

Hongkang Li, Songtao Lu, Xiaodong Cui et al.

2025-10-01 80
cs.LG 2510.00386

Train on Validation (ToV): Fast data selection with applications to fine-tuning

ToV method quickly selects data by reversing train-validation roles, significantly reducing test loss.

Ayush Jain, Andrea Montanari, Eren Sasoglu

2025-10-01 7
cs.SE 2510.02387

CWM: An Open-Weights LLM for Research on Code Generation with World Models

CWM: An open-weight large language model with world modeling, achieving 65.8% pass@1 on SWE-bench Verified.

FAIR CodeGen team, Jade Copet, Quentin Carbonneaux et al.

2025-10-01 39
cs.MM 2510.01284

Ovi: Twin Backbone Cross-Modal Fusion for Audio-Video Generation

Ovi employs twin DiT backbones with blockwise bidirectional cross-attention and scaled-RoPE to unify audio-video generation, trained on hundreds of thousands of hours of raw data.

Chetwin Low, Weimin Wang, Calder Katyal

2025-10-01 60
cs.RO 2510.00272

BC-MPPI: A Probabilistic Constraint Layer for Safe Model-Predictive Path-Integral Control

BC-MPPI enhances safety in path planning by integrating a Bayesian constraint layer, ensuring constraint adherence in complex environments.

Odichimnma Ezeji, Michael Ziegltrum, Giulio Turrisi et al.

2025-10-01 33
cs.AI 2510.00229

AgentFlux: Decoupled Fine-Tuning & Inference for On-Device Agentic Systems

AgentFlux uses decoupled fine-tuning with LoRA adapters, boosting local tool-calling accuracy by 46% on MCP-Bench.

Rohan Kadekodi, Zhan Jin, Keisuke Kamahori et al.

2025-10-01 55
cs.RO 2510.00225

TGPO: Temporal Grounded Policy Optimization for Signal Temporal Logic Tasks

TGPO method improves task success rate by 31.6% in Signal Temporal Logic tasks.

Yue Meng, Fei Chen, Chuchu Fan

2025-10-01 25
cs.LG 2509.26578

Linking Process to Outcome: Conditional Reward Modeling for LLM Reasoning

CRM links step rewards to outcomes via conditional hazards, reaching 43.3% on AIME24 in verifier-free RL, 16.7 points above PURE.

Zheng Zhang, Ziwei Shan, Kaitao Song et al.

2025-10-01 37
cs.LG 2509.26468

fev-bench: A Realistic Benchmark for Time Series Forecasting

fev-bench introduces a benchmark with 100 tasks across 7 domains, emphasizing covariates, using bootstrap confidence intervals for robust evaluation.

Oleksandr Shchur, Abdul Fatir Ansari, Caner Turkmen et al.

2025-10-01 29
cs.CL 2509.26048

RE-Searcher: Robust Agentic Search with Goal-oriented Planning and Self-reflection

RE-Searcher combines goal-oriented planning and self-reflection, achieving state-of-the-art robustness and accuracy in complex search environments.

Daocheng Fu, Jianbiao Mei, Licheng Wen et al.

2025-09-30 59
cs.CL 2509.25911

Mem-α: Learning Memory Construction via Reinforcement Learning

Mem-α optimizes memory construction via reinforcement learning, enhancing long-sequence processing.

Yu Wang, Ryuichi Takanobu, Zhiqi Liang et al.

2025-09-30 9
cs.AI 2509.25885

SafeMind: Benchmarking and Mitigating Safety Risks in Embodied LLM Agents

Proposes SafeMind with a three-layer safety architecture, boosting safety rate by over 30% in embodied LLM agents evaluated on 5558 multimodal samples.

Ruolin Chen, Yinqian Sun, Jihang Wang et al.

2025-09-30 41
cs.LG 2509.25810

Learning to Reason as Action Abstractions with Scalable Mid-Training RL

Proposes RA3, a mid-training algorithm that learns action abstractions via temporal ELBO, improving code generation by 8 points on average.

Shenao Zhang, Donghan Yu, Yihao Feng et al.

2025-09-30 54
cs.RO 2509.25756

SAC Flow: Sample-Efficient Reinforcement Learning of Flow-Based Policies via Velocity-Reparameterized Sequential Modeling

Proposes SAC Flow, reparameterizing velocity with modern sequence models to stabilize flow policies, achieving state-of-the-art results.

Yixian Zhang, Shu'ang Yu, Tonghe Zhang et al.

2025-09-30 48
stat.ML 2509.25741

Test time training enhances in-context learning of nonlinear functions

Combining test-time training (TTT) with in-context learning (ICL), this work achieves low prediction risk for nonlinear single-index models, with error approaching noise levels as data grows.

Kento Kuwataka, Taiji Suzuki

2025-09-30 70
Prev 1 ... 247 248 249 250 251 252 253 ... 573 Next

© 2026 GptGet.net - Paper Insights Platform

Paper List Submit Paper Help GptGet Home