GptGet
Features PaperForge Apps Papers Blog Contact AI Chat 中文
Sort: Latest Popular Citations
All Artificial Intelligence Computation and Language Computer Vision Information Retrieval Machine Learning Machine Learning (Stats) Neural and Evolutionary Computing Robotics
cs.CV 2606.13460

VISA: VLM-Guided Instance Semantic Auditing for 3D Occupancy World Models

VISA uses offline VLM auditing to improve 3D semantic occupancy mIoU, significantly enhancing rare-class performance.

Ruiqi Xian, Yuehan Xian, Jing Liang et al.

2026-06-11 40
cs.IR 2606.13438

CQC-RAG: Robust Retrieval-Augmented Generation via Cross-Query Consistency

CQC-RAG introduces cross-query consistency to enhance robustness in retrieval-augmented generation, outperforming baselines by +4.76 EM on TriviaQA and +9.12 EM on MuSiQue.

Yanjia Sun, Sifan Liu, Jie Shao

2026-06-11 883
stat.ME 2606.13433

Smoothed-KL Reweighting: A Principled Account and Matching Rule for SNR-Based Diffusion Training

Introduced Smoothed-KL weighting, validated on CIFAR-10 and CelebA-64 with 0.45 FID improvement on average.

Lei Li

2026-06-11 26
cs.AR 2606.13354

SupraSNN: Exploiting Synapse-Level Parallelism in Spiking Neural Network Accelerators through Co-Optimized Mapping and Scheduling

SupraSNN employs a superscalar-inspired architecture with synapse-level parallelism, achieving 149μs latency and 0.025mJ/image on FPGA for MNIST, outperforming prior accelerators.

Seyed Sadra Ghavami, Mohammad Hossein Nikkhah, Mohammad Rasoul Roshanshah et al.

2026-06-11 163
cs.CL 2606.13317

SkillCAT: Contrastive, Assessment-Augmented and Topology-AwareSkill Self-Evolution for LLM Agents

SkillCAT framework boosts LLM skill self-evolution by 49.69% through Contrastive Causal Extraction, Assessment-Augmented Evolution, and Topology-Aware Task Execution.

Kunfeng Chen, Qihuang Zhong, Juhua Liu et al.

2026-06-11 4
cs.LG 2606.13300

Quantizing Time-Series Models As Dynamical Systems: Trajectory-Based Quantization Sensitivity Score

Introduces Trajectory Sensitivity Score (TQS) based on dynamical systems stability to evaluate quantization impact on time-series models.

Mariya Pavlova, Harrison Bo Hua Zhu, Lidia Vitanova et al.

2026-06-11 47
stat.ML 2606.13277

ProtoX-AD: Self-Explainable Time Series Anomaly Detection and Characterization

ProtoX-AD is a prototype-based self-explainable time series anomaly detection framework that achieves comparable performance to black-box models by leveraging transformation-aware latent representations.

Aitor Sánchez-Ferrera, Elisabeth Wetzer, Kristoffer Wickstrøm et al.

2026-06-11 267
cs.CL 2606.13171

NTS-CoT: Mitigating Hallucinations in LLM-based News Timeline Summarization with Chain-of-Thought Reasoning

NTS-CoT reduces hallucinations in news timeline summarization using Chain-of-Thought, improving AR-1 by 23.4%.

Feng Lyu, Huiqin Yan, Sijing Duan et al.

2026-06-11 18
cs.RO 2606.13102

FTP-1: A Generalist Foundation Tactile Policy Across Tactile Sensors for Contact-Rich Manipulation

FTP-1 is a generalist tactile policy improving contact-rich manipulation success by 31%.

Chengbo Yuan, Zicheng Zhang, Mingjie Zhou et al.

2026-06-11 17
cs.PL 2606.13097

Functional Cache Grafting: Robust and Rapid Code-Policy Synthesis for Embodied Agents

FCGRAFT uses function-level KV-cache grafting to improve embodied-policy success by 18.31% and synthesis speed by 2.3×.

Saehun Chun, Wonje Choi, Sera Choi et al.

2026-06-11 44
cs.AI 2606.13038

Nous: An Attempt to Extract and Inject the Cognition Behind Prediction-Market Behavior

Nous extracts behavioral parameters from trading data and attempts prompt-based injection to induce cognitive diversity, but results show limited effectiveness.

Haowei Qian

2026-06-11 55
cs.CV 2606.14792

Efficient Reinforcement for Visual-Textual Thinking with Discrete Diffusion Model

Introduces LocFac-RL using discrete diffusion models to enhance visual-textual reasoning efficiency, reducing computation by 26.9%.

Yoonjeon Kim, Yuhta Takida, Chieh-Hsin Lai et al.

2026-06-11 36
cs.LG 2606.12994

DeepJEB++: Foundation Model-Driven Large-Scale 3D Engineering Dataset via 2D Latent Space Augmentation

DeepJEB++ leverages 2D latent space interpolation and foundation models to expand a small seed set into 15,360 labeled 3D jet engine brackets, with minimal resources.

Soyoung Yoo, Leekyo Jeong, Jinsu Ra et al.

2026-06-11 47
cs.AI 2606.13720

Refusal Beyond a Single Direction: A Preliminary Comparison of Diff-in-Means and INLP

This study compares Diff-in-Means (DiM) and INLP for extracting linear directions controlling model refusal, finding INLP's counterfactual flipping highly effective.

Elisabetta Rocchetti, Alfio Ferrara

2026-06-11 43
cs.AI 2606.12969

Multi-Modal Agents for Power Distribution Defect Detection: An Evaluation of Foundation Models

This study systematically evaluates multimodal foundation models' perception, reasoning, and tool use in power defect detection, with detailed experimental data.

Quan Quan

2026-06-11 47
cs.AI 2607.24772

RSMeM: Knowledge-Enhanced Memory Evolution for Remote Sensing Agents with Systematic Evaluation

RSMeM enhances remote sensing agents' tool usage accuracy by 6% through knowledge-enhanced memory evolution.

Bingxian Wu, Yu Zhang, Zonghao Guo et al.

2026-06-11 39
cs.RO 2606.12759

Sparse2Act: Learning Action-Aligned Sparse 3D Representations for Cross-Domain Robot Manipulation

Sparse2Act achieves cross-domain robot manipulation using action-aligned sparse 3D representations, with 86.9% success on LIBERO-10.

Yu Guo, Chang Yu, Siyu Ma et al.

2026-06-11 21
cs.CV 2606.12706

VLADriveBench: Evaluating CoT-Action Relationship in VLA for Autonomous Driving

VLADriveBench combines observational metrics and causal intervention to evaluate CoT–action causality in VLA autonomous driving models.

Thach Nguyen, Danhua Guo, Tom Lampo et al.

2026-06-11 44
cs.CR 2606.12666

CAPED: Context-Aware Privacy Exposure Defense for Mobile GUI Agents

CAPED employs task-driven selective exposure, reducing incidental visual privacy leaks by over 70% while maintaining high task utility.

Siyu Shen, Fenghao Xu, Wenrui Diao et al.

2026-06-11 68
cs.RO 2606.12402

DIRECT: When and Where Should You Allocate Test-Time Compute in Embodied Planners?

This paper introduces DIRECT, a multimodal scene-aware routing framework that dynamically allocates test-time compute among embodied planners, reducing latency by up to 65% while maintaining or surpassing top performance.

Jadelynn Dao, Milan Ganai, Yasmina Abukhadra et al.

2026-06-11 217
Prev 1 ... 102 103 104 105 106 107 108 ... 559 Next

© 2026 GptGet.net - Paper Insights Platform

Paper List Submit Paper Help GptGet Home