GptGet
Features PaperForge Apps Papers Blog Contact AI Chat 中文
Sort: Latest Popular Citations
All Artificial Intelligence Computation and Language Computer Vision Information Retrieval Machine Learning Machine Learning (Stats) Neural and Evolutionary Computing Robotics
cs.LG 2505.17708

The Third Pillar of Causal Analysis? A Measurement Perspective on Causal Representations

Proposes T-MEX score within measurement model framework to evaluate causal representations in CRL.

Dingling Yao, Shimeng Huang, Riccardo Cadei et al.

2025-05-23 34
cs.CV 2505.17685

FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving

FSDrive employs visual spatio-temporal Chain-of-Thought (CoT) to unify future scene prediction and trajectory planning, improving accuracy and safety in autonomous driving.

Shuang Zeng, Xinyuan Chang, Mengwei Xie et al.

2025-05-23 39
cs.CV 2505.17677

Towards Dynamic 3D Reconstruction of Hand-Instrument Interaction in Ophthalmic Surgery

OphNet-3D dataset enables dynamic 3D hand-instrument reconstruction in ophthalmic surgery, reducing MPJPE to 2.3mm and improving interaction metrics by 23%.

Ming Hu, Zhengdi Yu, Feilong Tang et al.

2025-05-23 64
cs.CL 2505.17667

QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning

QwenLong-L1 employs progressive context scaling and RL algorithms (GRPO, DAPO) to enhance long-text reasoning, outperforming existing models.

Fanqi Wan, Weizhou Shen, Shengyi Liao et al.

2025-05-23 46 citations 68
cs.CL 2505.17612

Distilling LLM Agent into Small Models with Retrieval and Code Tools

Proposes Agent Distillation, transferring LLM agent behaviors into small models with retrieval and code tools, achieving performance comparable to larger models.

Minki Kang, Jongwon Jeong, Seanie Lee et al.

2025-05-23 48
stat.ME 2505.17468

Efficient Adaptive Experimentation with Noncompliance

Proposes AMRIV, an adaptive IV estimator achieving semiparametric efficiency under noncompliance.

Miruna Oprescu, Brian M Cho, Nathan Kallus

2025-05-23 28
cs.CV 2505.17412

Direct3D-S2: Gigascale 3D Generation Made Easy with Spatial Sparse Attention

Direct3D-S2 uses Spatial Sparse Attention for efficient 3D generation, achieving significant speedups.

Shuang Wu, Youtian Lin, Feihu Zhang et al.

2025-05-23 39
cs.LG 2505.17373

Value-Guided Search for Efficient Chain-of-Thought Reasoning

Introduced Value-Guided Search (VGS) using a 1.5B value model trained on 2.5M reasoning traces for efficient long-context reasoning.

Kaiwen Wang, Jin Peng Zhou, Jonathan Chang et al.

2025-05-23 11
cs.CL 2505.17322

From Compression to Expression: A Layerwise Analysis of In-Context Learning

Layerwise Compression-Expression phenomenon reveals task info capture and generation in ICL.

Jiachen Jiang, Yuxin Dong, Jinxin Zhou et al.

2025-05-23 11
cs.CV 2505.17020

CrossLMM: Decoupling Long Video Sequences from LMMs via Dual Cross-Attention Mechanisms

CrossLMM employs dual cross-attention to compress long video sequences, maintaining performance with fewer tokens.

Shilin Yan, Jiaming Han, Joey Tsai et al.

2025-05-23 37
cs.LG 2505.17016

Interactive Post-Training for Vision-Language-Action Models

RIPT-VLA fine-tunes pretrained VLA models via reinforcement learning, boosting success rate to 97.5% with only one demonstration.

Shuhan Tan, Kairan Dou, Yue Zhao et al.

2025-05-23 121 citations 35
cs.AI 2505.16854

Think or Not? Selective Reasoning via Reinforcement Learning for Vision-Language Models

Proposes TON framework with Thought Dropout and GRPO for selective reasoning, reducing 90% inference length while maintaining accuracy.

Jiaqi Wang, Kevin Qinghong Lin, James Cheng et al.

2025-05-23 49
cs.LG 2505.16829

Contextual Learning for Stochastic Optimization

Capped Squared Loss learns contextual value distributions with O(dξ²cmax⁴/ε⁸δ²) samples.

Anna Heuser, Thomas Kesselheim

2025-05-23 31
cs.IR 2505.16810

DeepRec: Towards a Deep Dive Into the Item Space with Large Language Model Based Recommendation

DeepRec enhances recommendation by multi-turn interactions between LLMs and TRMs, significantly improving performance.

Bowen Zheng, Xiaolei Wang, Enze Liu et al.

2025-05-22 35
cs.CV 2505.16707

KRIS-Bench: Benchmarking Next-Level Intelligent Image Editing Models

KRIS-Bench employs a cognitive-inspired taxonomy to evaluate models' reasoning, covering 22 tasks across 7 dimensions with a focus on knowledge plausibility.

Yongliang Wu, Zonghui Li, Xinting Hu et al.

2025-05-22 41
cs.CV 2505.16687

One-Step Diffusion-Based Image Compression with Semantic Distillation

Proposes OneDC, a one-step diffusion image codec with semantic distillation, reducing bitrate by 39% and decoding time by 20×.

Naifu Xue, Zhaoyang Jia, Jiahao Li et al.

2025-05-22 32
cs.CL 2505.16297

ToDi: Token-wise Distillation via Fine-Grained Divergence Control

ToDi adaptively combines FKL and RKL per token via probability ratio, significantly improving distillation accuracy.

Seongryong Jung, Suwan Yoon, DongGeon Kim et al.

2025-05-22 69
cs.LG 2505.16265

Think-RM: Enabling Long-Horizon Reasoning in Generative Reward Models

Think-RM models internal reasoning to enable long-horizon inference, outperforming BT RM and scaled GenRM by 8% on RM-Bench.

Ilgee Hong, Changlong Yu, Liang Qiu et al.

2025-05-22 31
cs.CL 2505.16258

IRONIC: Coherence-Aware Reasoning Chains for Multi-Modal Sarcasm Detection

IRONIC framework achieves state-of-the-art zero-shot sarcasm detection using multi-modal coherence relations.

Aashish Anantha Ramakrishnan, Aadarsh Anantha Ramakrishnan, Dongwon Lee

2025-05-22 56
cs.CV 2505.16174

Erased or Dormant? Rethinking Concept Erasure Through Reversibility

Using lightweight fine-tuning to probe concept erasure reversibility, revealing existing methods only achieve superficial suppression.

Ping Liu, Chi Zhang

2025-05-22 37
Prev 1 ... 282 283 284 285 286 287 288 ... 573 Next

© 2026 GptGet.net - Paper Insights Platform

Paper List Submit Paper Help GptGet Home