GptGet
Features PaperForge Apps Papers Blog Contact AI Chat 中文
Sort: Latest Popular Citations
All Artificial Intelligence Computation and Language Computer Vision Information Retrieval Machine Learning Machine Learning (Stats) Neural and Evolutionary Computing Robotics
cs.SE 2607.02703

LLMoxie: Exploring Agentic AI for Scientific Software Development

LLMoxie platform enhances AI agent performance in scientific software development through a three-tier architecture and RSE-Plugins ecosystem.

Landung Setiawan, Anant Mittal, Cordero Core et al.

2026-07-03 10
cs.RO 2607.02646

EVA-Client: A Unified Data Collection, Inference, and Deployment Framework for Embodied Policies on Real Robots

EVA-Client provides a unified deployment, debugging, and data collection framework for embodied policies across multiple robot platforms, integrating various real-time inference strategies.

Heqing Yang, Yang Yi, Liyao Wang et al.

2026-07-03 47
cs.AI 2607.02440

EvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive Environments

EvoPolicyGym benchmarks autonomous policy evolution; GPT-5.5 achieves top performance across 16 RL environments with limited interactions.

Zhilin Wang, Han Song, Runzhe Zhan et al.

2026-07-03 59
cs.LG 2607.02637

Post-Generation Curation of Synthetic Images via Homogeneous-Heterogeneous Splitting

HO-HE post-generation curation balances fidelity and diversity, reaching 95% on CIFAR-10 with 300K synthetic images.

Disheng Liu, Tuo Liang, Chaoda Song et al.

2026-07-02 27
cs.RO 2607.02322

The Moving Eye: Enhancing VLA Spatial Generalization via Hybrid Dynamic Data Collection

Proposed a hybrid dynamic data collection method, significantly enhancing VLA spatial generalization with a 40% success rate increase.

Jincheng Tang, Yilong Zhu, Zhengyuan Xie et al.

2026-07-02 15
cs.LG 2607.02291

Optimizing Visual Generative Models via Distribution-wise Rewards

Proposes distribution-wise reward with subset-replace strategy for RL fine-tuning, reducing FID from 8.30 to 5.77.

Ruihang Li, Mengde Xu, Shuyang Gu et al.

2026-07-02 29
cs.RO 2607.02195

Bridge-WA: Predicting Where and How the World Changes for Robotic Action

Bridge-WA enhances robotic action success by predicting scene changes, achieving a 9.7% improvement on VLABench.

Yongjie Bai, Hanting Wang, Mingtong Dai et al.

2026-07-02 10
cs.CV 2607.02096

LongEgoRefer: A Benchmark for Long-Form Egocentric Video Referring Expression Comprehension

LongEgoRefer benchmark challenges existing models with long-form egocentric video referring expression comprehension.

Shunya Kato, Taiki Miyanishi, Shuhei Kurita et al.

2026-07-02 19
cs.AI 2607.02073

Evidence-State Rewards for Long-Context Reasoning

MAVEN framework enhances long-context reasoning by rewarding dynamic evidence state transitions, improving performance by 3.5% on LongBench v2.

Ya Gao, Pekka Marttinen

2026-07-02 17
cs.CV 2607.02045

PWM-ArtGen: Part World Model for Articulated Object Generation

PWM-ArtGen uses diffusion-based joint modeling of visual dynamics and kinematic parameters, outperforming baselines with strong zero-shot generalization.

Wentao Zheng, Ancong Wu

2026-07-02 43
cs.AI 2607.01978

Multimodal Knowledge Edit-Scoped Generalization for Online Recursive MLLM Editing

ScopeEdit employs a dual-branch scope-aware mechanism with orthogonal low-rank geometry to control knowledge propagation in online multimodal model editing.

Siyuan Li, Youyuan Zhang, Ruitong Liu et al.

2026-07-02 22
cs.RO 2607.01938

PhysMani: Physics-principled 3D World Model for Dynamic Object Manipulation

PhysMani combines a physics-principled 3D Gaussian world model with a future-aware action policy model to enhance dynamic object manipulation success rates.

Peng Yun, Shouwang Huang, Hao Li et al.

2026-07-02 15
cs.LG 2607.01918

Zeus: Towards Tuning-Free Foundation Model for Time Series Analysis

Zeus is a tuning-free foundation model for time series, using multi-scale Transformer and multi-objective masking to excel across tasks.

Yisong Fu, Zezhi Shao, Chengqing Yu et al.

2026-07-02 40
cs.AI 2607.01874

SkillCoach: Self-Evolving Rubrics for Evaluating and Enhancing Agentic Skill-Use

SkillCoach improves agentic skill-use evaluation with self-evolving rubrics, significantly enhancing assessment quality.

Jiayin Zhu, Kelong Mao, Yudong Guo et al.

2026-07-02 16
eess.SY 2607.01819

Koopman operator theory: fundamentals, control, and applications

Koopman operator theory enables global linearization of complex dynamical systems using methods like EDMD.

Igor Mezić, Jorge Cortés, Karl Worthmann et al.

2026-07-02 16
cs.LG 2607.01763

Denser $\neq$ Better: Limits of On-Policy Self-Distillation for Continual Post-Training

Study finds SDPO accelerates in-domain learning under specific conditions but struggles in cross-domain scenarios.

Meng Wang, Haohan Zhao, Wenzhuo Liu et al.

2026-07-02 14
cs.CV 2607.01737

ReQuest: Rethinking-based Question-Aware Frame Selection for Long-Form Video QA

ReQuest improves long-video QA accuracy through uncertainty-driven keyframe selection.

Minkuk Kim, Suyong Yun, Young Tae Kim et al.

2026-07-02 17
cs.CV 2607.01707

LASER: A Corrective Lens for LVLMs via Visual Attention Preservation and Sink Suppression

LASER regulates visual attention via grounding and sink suppression rewards, effectively mitigating visual forgetting in LVLMs.

Bowen Yuan, Zijian Wang, Yadan Luo et al.

2026-07-02 43
cs.CV 2607.01677

ICDepth: Taming Video Diffusion Models for Video Depth Estimation via In-Context Conditioning

ICDepth uses In-Context Conditioning and SAND-Attention for video depth estimation, achieving 6-13x data efficiency improvement.

Xuanhua He, Jiaxin Xie, Mingzhe Zheng et al.

2026-07-02 35
cs.CV 2607.01663

Unified Panoramic-Gaussian Representation for Monocular 4D Scene Synthesis

Proposes PanoGaussian, combining panoramic trajectories and Gaussian distillation for consistent 4D scene synthesis from monocular videos, even in unseen regions.

Yuankun Yang, Yi Wei, Wenyang Zhou et al.

2026-07-02 47
Prev 1 ... 78 79 80 81 82 83 84 ... 552 Next

© 2026 GptGet.net - Paper Insights Platform

Paper List Submit Paper Help GptGet Home