GptGet
Features PaperForge Apps Papers Blog Contact AI Chat 中文
Sort: Latest Popular Citations
All Artificial Intelligence Computation and Language Computer Vision Information Retrieval Machine Learning Machine Learning (Stats) Neural and Evolutionary Computing Robotics
cs.LG 2609.10490

Learning with Covariance Matrices: Principal Component Analysis Meets Learning with Graphs

Introduces covariance neural networks (VNNs), combining PCA and graph learning to improve stability and transferability on multiscale datasets.

Saurabh Sihag, Andrea Cavallo, Elvin Isufi et al.

2026-09-10 88
cs.LG 2609.10464

Semigroup-JEPA: Latent Dynamics Consistency for Zero-Shot Physics Generalization

Introduced Semigroup-JEPA for zero-shot physics generalization, achieving 34% error reduction and 23.3% control success improvement.

Andy Zeyi Liu, Haoran Sun, Lucas Baker et al.

2026-09-10 94
cs.AI 2609.10451

JarvisGUI: Towards Cross-Device GUI Agents with Dynamic Task Composition

JarvisGUI evaluates cross-device GUI agents with dynamic task composition, revealing capability gaps in real-world workflows.

Zixiang Chen, Yuheng Lu, Zihao Cheng et al.

2026-09-10 95
cs.AI 2609.10441

ConvMem: Convolutional Memory for Long-Context Reasoning

ConvMem reformulates long-context reasoning as hierarchical convolution, outperforming training-free baselines on RULER-HotpotQA.

Hongming Zhang, Zhaozhen Gu, Fengshuo Bai et al.

2026-09-10 96
cs.LG 2609.10439

Forgetting Only What Matters: Layer-Selective Unlearning toward Robust LLMs

FOM-UL achieves efficient forgetting by selectively updating Transformer layers, maintaining utility and privacy under 8/4-bit quantization.

Ravi Ranjan, Olivera Kotevska, Agoritsa Polyzou

2026-09-10 79
cs.AI 2609.10413

Fortunate Recall: Ontology-Driven Memory Lifecycle Management for Persistent Coherence in LLMs

Fortunate Recall uses a 10+1 behavioral ontology and lifecycle policies to enhance LLM memory management, achieving 76.9% on LifecycleBench.

Ansuman Mullick, Eray Tüzün

2026-09-10 81
cs.RO 2609.10405

Frequency-Conditioned Flow Matching for Vision-Language-Action Models

FreqFM improves VLA models by frequency conditioning, achieving a 9.3-point gain on LIBERO-Plus.

Haochen Niu, Shengye Dong, Hao Liu et al.

2026-09-10 85
cs.HC 2609.10385

MOONWALK: Mediating Operations with Intent-Evidence-Action Alignment Across Junior-Supervisor Review Workflows in Animation/VFX Pre-Production

MOONWALK optimizes animation/VFX pre-production reviews with an intent-evidence-action alignment framework, improving intent clarity and task executability.

Shih-Yu Lai, Wen-Fan Wang, Sai Ling et al.

2026-09-10 88
cs.RO 2609.10377

Data-Driven Risk Fields for Safer End-to-End Autonomous Driving

DRiF framework uses data-driven risk fields for safer end-to-end driving, achieving 88.78 driving score on Bench2Drive.

Yuanxin Tian, Zhiyuan Liu, Jinhao Li et al.

2026-09-10 94
cs.CV 2609.10372

PACE: Perceived-Latency-Aware Cascading Service Routing and Filler Control for QoE-Efficient Retrieval-Augmented Dialogue Serving

PACE reduces P95 PTFR from 0.53s to 0.29s via perceived-latency-aware dialogue routing.

Lin Huang, Yujuan Tan, Weisheng Li et al.

2026-09-10 85
cs.CV 2609.10355

Why Is Video Still So Expensive? A Survey of Inference-Efficiency Mechanisms in Video and Audiovisual LLMs

Survey of inference-efficiency mechanisms in VideoLLMs, analyzing frame sampling, modality encoding, and token compression techniques.

Killian Steunou, Yannis Tevissen, Mounîm A. El Yacoubi

2026-09-09 87
cs.RO 2609.10339

A Confidence-Aware Multimodal Fusion Framework for Industrial Human-Robot Collaboration

Proposed CAMF achieves 91.86% intention recognition accuracy for industrial human-robot collaboration.

Xinyu Liu, Qiqi Dong, Boya Jia et al.

2026-09-09 83
cs.CV 2609.10317

Decoupled Self-Forcing Distillation for Streaming Talking Head Generation

Motar achieves 15.4 FPS high-fidelity streaming talking head generation by fusing conditions in a low-dimensional motion space.

Yanru An, Ruiyan Wang, Wenwu Wei et al.

2026-09-09 7
cs.CV 2609.10156

ScopeMamba-YOLO: Widening the Perceptual Scope Inward and Outward for Small Object Detection in Remote Sensing Imagery

ScopeMamba-YOLO enhances small object detection in remote sensing imagery by widening perceptual scope, achieving a 10.8 pp mAP50 improvement.

Junjie Fan, Yijun Mai, Linduo Wei et al.

2026-09-09 4
cs.CV 2609.09924

Multimodal Emotion Recognition in Conversations via Class-Wise Adaptive Modality Fusion and Affective Geometry

An SDT extension combining geometric vision, class-wise fusion, and affective priors reaches 75.93% and 74.11% weighted F1 on MELD and IEMOCAP.

Oriol Marín, Roger Marí, Gloria Haro et al.

2026-09-09 44
eess.IV 2609.09801

Morphological Decoupling-Based Skeletal Classification for Clinical Assessment of Malocclusion

TeethGNN fuses CBCT, ANB/MP-FH morphology, and graph calibration, reaching 77.08% accuracy and 89.61% AUC.

Zhichun Jin, Zhicheng He, Hao Xu et al.

2026-09-09 36
cs.CL 2609.09764

SocialRL: Refining LLMs' Social Intelligence through Multi-turn Reinforcement Learning and Reward Design

SocialRL combines multi-turn PPO with dynamic process rewards, raising goal achievement by an average of 9.2 percentage points.

Jianing Wang, Xintao Wang, Aili Chen et al.

2026-09-09 42
cs.CV 2609.09528

MotionBlind: Probing the Illusion of Motion Understanding in Video-LLMs

MotionBlind reveals Video-LLMs' inability to accurately understand motion in videos; only Gemini3.1 Pro passes.

Dhairya Bhatia, Bishoy Galoaa, Oliver Fritsche et al.

2026-09-09 6
cs.CV 2609.09396

VANTAGE-Bench: Evaluating the Infrastructure AI Gap in Vision-Language Models

VANTAGE-Bench evaluates the Infrastructure AI gap in VLMs, revealing deficiencies in event verification and temporal localization tasks.

Zaid Pervaiz Bhat, Nimra Nayyar, Arihant Jain et al.

2026-09-09 5
cs.CV 2609.08230

ActionSplice: In-Flight Action Editing for Interactive World Models

ActionSplice uses Counterfactual State Transport to enable in-flight action editing, reducing LPIPS by up to 75.9% and speeding up sampling 2.73×.

Pardis Taghavi, Tingyu Guo, Jonas Lossner et al.

2026-09-08 21
Prev 1 2 3 4 5 6 7 8 ... 511 Next

© 2026 GptGet.net - Paper Insights Platform

Paper List Submit Paper Help GptGet Home