GptGet
Features PaperForge Apps Papers Blog Contact AI Chat 中文
Sort: Latest Popular Citations
All Artificial Intelligence Computation and Language Computer Vision Information Retrieval Machine Learning Machine Learning (Stats) Neural and Evolutionary Computing Robotics
cs.LG 2601.11259

Latent Dynamics Graph Convolutional Networks for model order reduction of parameterized time-dependent PDEs

Proposes LD-GCN, an encoder-free GNN framework for low-dimensional modeling of parameterized time-dependent PDEs, improving interpretability and efficiency.

Lorenzo Tomada, Federico Pichi, Gianluigi Rozza

2026-01-16 44
cs.AI 2601.11100

ReCreate: Reasoning and Creating Domain Agents Driven by Experience

ReCreate employs experience-driven, interaction-based scaffold updates to automate domain agent creation, outperforming black-box methods with 5%+ gains.

Zhezheng Hao, Hong Wang, Jian Luo et al.

2026-01-16 44
cs.CV 2601.11035

Your One-Stop Solution for AI-Generated Video Detection

AIGVDBench: benchmark with 31 models, 440k videos, evaluating 33 detectors, offering comprehensive analysis.

Long Ma, Zihao Xue, Yan Wang et al.

2026-01-16 56
cs.CV 2601.14037

Human detectors are surprisingly powerful reward models

HuDA model uses human detection and temporal alignment to enhance video generation, achieving a 73% win rate.

Kumar Ashutosh, XuDong Wang, Xi Yin et al.

2026-01-16 27
cs.CV 2601.10611

Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding

Molmo2 is an open-source vision-language model with pixel-level grounding, outperforming existing open and proprietary models on video understanding tasks.

Christopher Clark, Jieyu Zhang, Zixian Ma et al.

2026-01-16 27
cs.CV 2601.10553

Inference-time Physics Alignment of Video Generative Models with Latent World Models

Improved video generation physics plausibility using WMReward and VJEPA-2, achieving 62.64% in ICCV 2025 challenge.

Jianhao Yuan, Xiaofeng Zhang, Felix Friedrich et al.

2026-01-16 37
cs.SD 2601.10547

HeartMuLa: A Family of Open Sourced Music Foundation Models

HeartMuLa integrates four modules to achieve controllable music generation with 7B parameters, excelling in fidelity and structure.

Dongchao Yang, Yuxin Xie, Yuguo Yin et al.

2026-01-16 31
cs.AI 2601.10413

LADFA: A Framework of Using Large Language Models and Retrieval-Augmented Generation for Personal Data Flow Analysis in Privacy Policies

LADFA combines LLMs and RAG to analyze personal data flows in privacy policies.

Haiyue Yuan, Nikolay Matyunin, Ali Raza et al.

2026-01-15 39
cs.CR 2601.10338

Agent Skills in the Wild: An Empirical Study of Security Vulnerabilities at Scale

SkillScan framework reveals 26.1% vulnerabilities in AI skill markets, urging enhanced security vetting.

Yi Liu, Weizhe Wang, Ruitao Feng et al.

2026-01-15 28
cs.AI 2601.10306

Evidence-Augmented Policy Optimization with Reward Co-Evolution for Long-Context Reasoning

EAPO enhances long-context reasoning with reward co-evolution, achieving significant performance gains.

Xin Guan, Zijian Li, Shen Huang et al.

2026-01-15 25
cs.CV 2601.10214

Beyond Inpainting: Unleash 3D Understanding for Precise Camera-Controlled Video Generation

DepthDirector uses depth video guidance for precise camera control and consistent content generation.

Dong-Yu Chen, Yixin Guo, Shuojin Yang et al.

2026-01-15 38
cs.CV 2601.10168

RAG-3DSG: Enhancing 3D Scene Graphs with Re-Shot Guided Retrieval-Augmented Generation

RAG-3DSG enhances 3D scene graphs by re-shot guided uncertainty estimation and retrieval-augmented generation, achieving state-of-the-art results in semantic consistency and precision.

Yue Chang, Rufeng Chen, Zhaofan Zhang et al.

2026-01-15 44
cs.CV 2601.09708

Fast-ThinkAct: Efficient Vision-Language-Action Reasoning via Verbalizable Latent Planning

Fast-ThinkAct compresses chain-of-thought reasoning into verbalizable latent representations, reducing inference latency by 89.3%.

Chi-Pin Huang, Yunze Man, Zhiding Yu et al.

2026-01-15 41
cs.CL 2601.09609

DPWriter: Reinforcement Learning with Diverse Planning Branching for Creative Writing

DPWriter combines semi-structured long Chain-of-Thought with diversity-guided RL, achieving over 10% improvement in output diversity metrics on creative writing benchmarks.

Qian Cao, Yahui Liu, Wei Bi et al.

2026-01-15 62
cs.AI 2601.09770

GUI-Eyes: Tool-Augmented Perception for Visual Grounding in GUI Agents

GUI-Eyes achieves 44.8% grounding accuracy with tool-augmented perception using only 3k samples.

Chen Chen, Jiawei Shao, Dakuan Lu et al.

2026-01-14 17
cs.CL 2601.09065

Beyond Consensus: Perspectivist Modeling and Evaluation of Annotator Disagreement in NLP

Utilizing perspectivist modeling to analyze annotator disagreement in NLP, enhancing model fairness.

Yinuo Xu, David Jurgens

2026-01-14 21
cs.IR 2601.08816

MemRec: Collaborative Memory-Augmented Agentic Recommender System

MemRec enhances recommender systems with collaborative memory, achieving a 28.98% H@1 improvement on Goodreads.

Weixin Chen, Yuhan Zhao, Jingyuan Huang et al.

2026-01-14 27
cs.AI 2601.08679

PersonaDual: Balancing Personalization and Objectivity via Adaptive Reasoning

PersonaDual balances personalization and objectivity via adaptive reasoning, improving accuracy by 3%.

Xiaoyou Liu, Xinyi Mou, Shengbin Yue et al.

2026-01-14 18
cs.CL 2601.08892

Evaluating Role-Consistency in LLMs for Counselor Training

Introduces adversarial attack-based evaluation of role consistency in Vicuna LLMs for virtual counseling.

Eric Rudolph, Natalie Engert, Jens Albrecht

2026-01-13 35
cs.AI 2601.08444

Beyond Linearization: Attributed Table Graphs for Table Reasoning

Proposes TabGR using Attributed Table Graph and QG-PPR to improve table reasoning accuracy by 9.7%.

Yuxiang Wang, Junhao Gan, Shengxiang Gao et al.

2026-01-13 47
Prev 1 ... 217 218 219 220 221 222 223 ... 573 Next

© 2026 GptGet.net - Paper Insights Platform

Paper List Submit Paper Help GptGet Home