GptGet
Features PaperForge Apps Papers Blog Contact AI Chat 中文
Sort: Latest Popular Citations
All Artificial Intelligence Computation and Language Computer Vision Information Retrieval Machine Learning Machine Learning (Stats) Neural and Evolutionary Computing Robotics
cs.AI 2512.12692

WebOperator: Action-Aware Tree Search for Autonomous Agents in Web Environment

WebOperator employs action-aware tree search with safe backtracking, achieving 54.6% success on WebArena.

Mahir Labib Dihan, Tanzima Hashem, Mohammed Eunus Ali et al.

2025-12-14 55
cs.CV 2512.12492

Adaptive Detector-Verifier Framework for Zero-Shot Polyp Detection in Open-World Settings

Proposes AdaptiveDetector, combining YOLOv11 and VLM with adaptive thresholds and GRPO, boosting zero-shot polyp recall by 14-22% under challenging conditions.

Shengkai Xu, Hsiang Lun Kao, Tianxiang Xu et al.

2025-12-14 48
cs.CV 2512.12309

WeDetect: Fast Open-Vocabulary Object Detection as Retrieval

WeDetect achieves fast open-vocabulary object detection via retrieval, achieving SOTA across 15 benchmarks with high inference efficiency.

Shenghao Fu, Yukun Su, Fengyun Rao et al.

2025-12-13 34
cs.CL 2512.11399

Minimal Clips, Maximum Salience: Long Video Summarization via Key Moment Extraction

Proposes lightweight clip selection combined with large language models to extract key moments, achieving near-reference summary quality with less than 6% video content.

Galann Pennec, Zhengyuan Liu, Nicholas Asher et al.

2025-12-12 52
cs.CV 2512.11336

UFVideo: Towards Unified Fine-Grained Video Cooperative Understanding with Large Language Models

UFVideo unifies multi-scale video understanding, integrating global, pixel, and temporal info, outperforming GPT-4 with 7.3% improvement across benchmarks.

Hewen Pan, Cong Wei, Dashuang Liang et al.

2025-12-12 45
cs.CV 2512.10954

Group Diffusion: Enhancing Image Generation by Unlocking Cross-Sample Collaboration

Proposes Group Diffusion, leveraging cross-sample attention to improve image generation, achieving up to 32.2% FID reduction.

Sicheng Mo, Thao Nguyen, Richard Zhang et al.

2025-12-12 49
cs.CV 2512.10818

Self-Ensemble Post Learning for Noisy Domain Generalization

Proposed SEPL method enhances noisy domain generalization via feature probing and prediction ensemble.

Wang Lu, Jindong Wang

2025-12-12 16
cs.CL 2512.10791

The FACTS Leaderboard: A Comprehensive Benchmark for Large Language Model Factuality

The FACTS Leaderboard evaluates large language models' factuality using four sub-leaderboards, with an average score of 68.8.

Aileen Cheng, Alon Jacovi, Amir Globerson et al.

2025-12-12 22
cs.AI 2512.10696

Remember Me, Refine Me: A Dynamic Procedural Memory Framework for Experience-Driven Agent Evolution

ReMe employs multi-faceted distillation, scenario-aware indexing, and utility-based refinement, enabling small models to outperform larger ones without memory, with 8.83% gain on BFCL-V3.

Zouying Cao, Jiaji Deng, Li Yu et al.

2025-12-11 41
cs.LG 2512.10287

A Kernel-based Resource-efficient Neural Surrogate for Multi-fidelity Prediction of Aerodynamic Field

Proposes KHRONOS, a kernel-based neural surrogate for multi-fidelity aerodynamic prediction, reducing parameters by 94% and accelerating training/inference.

Apurba Sarker, Reza T. Batley, Darshan Sarojini et al.

2025-12-11 31
cs.LG 2512.10042

SEMDICE: Off-policy State Entropy Maximization via Stationary Distribution Correction Estimation

SEMDICE is an off-policy state entropy maximization method using stationary distribution correction, enabling learning from arbitrary datasets with theoretical guarantees.

Jongmin Lee, Meiqi Sun, Pieter Abbeel

2025-12-11 48
cs.CV 2512.09864

UniUGP: Unifying Understanding, Generation, and Planing For End-to-end Autonomous Driving

UniUGP framework enhances autonomous driving by integrating understanding, generation, and planning, achieving superior decision accuracy in long-tail scenarios.

Hao Lu, Ziyang Liu, Guangfeng Jiang et al.

2025-12-11 8
cs.LG 2512.09850

Conformal bandits: bringing statistical validity and reward efficiency under weak arm separability

Integrates conformal prediction into multi-armed bandits, ensuring finite-sample coverage and reward efficiency under weak arm separability.

Simone Cuonzo, Nina Deliu

2025-12-11 50
cs.CL 2512.09742

Weird Generalization and Inductive Backdoors: New Ways to Corrupt LLMs

Fine-tuning on narrow datasets causes broad, unpredictable behaviors and hidden backdoors, exposing security risks in LLMs.

Jan Betley, Jorio Cocola, Dylan Feng et al.

2025-12-10 41
cs.CL 2512.09675

d-TreeRPO: Towards More Reliable Policy Optimization for Diffusion Language Models

d-TreeRPO enhances policy optimization reliability for diffusion language models, achieving 86.2% improvement on Sudoku.

Leyi Pan, Shuchang Tao, Yunpeng Zhai et al.

2025-12-10 15
cs.RO 2512.09343

Development and Testing for Perception Based Autonomous Landing of a Long-Range QuadPlane

A YOLO–TensorRT–Isaac ROS QuadPlane landing stack achieved 32.7 ms per frame, enabling over 30 FPS edge inference.

Ashik E Rasul, Humaira Tasnim, Ji Yu Kim et al.

2025-12-10 36
cs.DC 2512.09277

Efficient MoE Serving in the Memory-Bound Regime: Balance Activated Experts, Not Tokens

METRO algorithm minimizes activated experts instead of token counts, reducing latency and boosting throughput in memory-bound MoE inference.

Yanpeng Yu, Haiyue Ma, Krish Agarwal et al.

2025-12-10 33
cs.CV 2512.08860

Tri-Bench: Stress-Testing VLM Reliability on Spatial Reasoning under Camera Tilt and Object Interference

Tri-Bench tests VLM spatial reasoning under camera tilt and object interference, with ~69% average accuracy.

Amit Bendkhale

2025-12-10 21
cs.CV 2512.08829

InfiniteVL: Synergizing Linear and Sparse Attention for Highly-Efficient, Unlimited-Input Vision-Language Models

InfiniteVL synergizes linear and sparse attention for efficient unlimited-input vision-language models, achieving 1.7x decoding speedup.

Hongyuan Tao, Bencheng Liao, Shaoyu Chen et al.

2025-12-10 23
cs.RO 2512.11891

VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer

AEGIS integrates control barrier functions into VLA models, achieving over 50% improvement in obstacle avoidance and nearly 10% higher task success.

Songqiao Hu, Zeyi Liu, Shuang Liu et al.

2025-12-10 72
Prev 1 ... 226 227 228 229 230 231 232 ... 573 Next

© 2026 GptGet.net - Paper Insights Platform

Paper List Submit Paper Help GptGet Home