GptGet
Features PaperForge Apps Papers Blog Contact AI Chat 中文
Sort: Latest Popular Citations
All Artificial Intelligence Computation and Language Computer Vision Information Retrieval Machine Learning Machine Learning (Stats) Neural and Evolutionary Computing Robotics
cs.CL 2606.11220

LifeSentence: Language models can encode human life course trajectories from longitudinal panel data

LifeSentence leverages pretrained language models with structured event data to predict and analyze human life trajectories, outperforming baselines with 35.3% joint accuracy.

Samuel Liu, Muchen Xi, William Yeoh et al.

2026-05-15 42
cs.CL 2605.15081

ML-Embed: Inclusive and Efficient Embeddings for a Multilingual World

ML-Embed integrates 3D-ML framework with MEL, MLL, MRL for efficient, inclusive multilingual embeddings; achieves SOTA on 9/17 MTEB benchmarks.

Ziyin Zhang, Zihan Liao, Hang Yu et al.

2026-05-15 45
cs.CL 2605.14600

SciPaths: Forecasting Pathways to Scientific Discovery

Introduces SciPaths, a dataset with 262 expert-annotated pathways, using ML models like GPT-4 to predict scientific dependencies with F1=0.189.

Eric Chamoun, Yizhou Chi, Yulong Chen et al.

2026-05-14 32
cs.CL 2605.14473

Does RAG Know When Retrieval Is Wrong? Diagnosing Context Compliance under Knowledge Conflict

CDD method enhances RAG model accuracy under knowledge conflict, especially in entity swap and logical contradiction scenarios.

Yihang Chen, Pin Qian, Su Wang et al.

2026-05-14 3
cs.CL 2605.14401

Agentic Recommender System with Hierarchical Belief-State Memory

MARS employs a three-tier memory architecture with adaptive lifecycle management, achieving 26.4% HR@1 improvement in recommendation tasks.

Xiang Shen, Yuhang Zhou, Yifan Wu et al.

2026-05-14 14
cs.CL 2605.12493

LongMemEval-V2: Evaluating Long-Term Agent Memory Toward Experienced Colleagues

LongMemEval-V2 achieves 72.5% accuracy with AgentRunbook-C, evaluating long-term memory in agents.

Di Wu, Zixiang Ji, Asmi Kawatkar et al.

2026-05-13 1806
cs.CL 2605.12487

Task-Adaptive Embedding Refinement via Test-time LLM Guidance

Task-Adaptive Embedding Refinement via Test-time LLM Guidance improves zero-shot search and classification by up to 25%.

Ariel Gera, Shir Ashury-Tahan, Gal Bloch et al.

2026-05-13 219
cs.CL 2605.12452

The Algorithmic Caricature: Auditing LLM-Generated Political Discourse Across Crisis Events

Using a Computational Social Science framework, audit LLM-generated political discourse across nine crisis events, finding it more negative and structurally consistent.

Gunjan, Sidahmed Benabderrahmane, Talal Rahwan

2026-05-13 158
cs.CL 2605.12398

Question Difficulty Estimation for Large Language Models via Answer Plausibility Scoring

Q-DAPS estimates question difficulty by computing the entropy of plausibility scores, excelling on four QA datasets.

Jamshid Mozafari, Bhawna Piryani, Adam Jatowt

2026-05-13 219
cs.CL 2605.12361

MedHopQA: A Disease-Centered Multi-Hop Reasoning Benchmark and Evaluation Framework for LLM-Based Biomedical Question Answering

MedHopQA evaluates biomedical QA via multi-hop reasoning with 1,000 expert-curated question-answer pairs.

Rezarta Islamaj, Robert Leaman, Joey Chan et al.

2026-05-13 197
cs.CL 2605.12185

Mitigating Context-Memory Conflicts in LLMs through Dynamic Cognitive Reconciliation Decoding

Dynamic Cognitive Reconciliation Decoding (DCRD) predicts and mitigates context-memory conflicts, achieving state-of-the-art performance across six QA datasets.

Yigeng Zhou, Wu Li, Yifan Lu et al.

2026-05-12 3
cs.CL 2605.11317

SOMA: Efficient Multi-turn LLM Serving via Small Language Model

SOMA enables efficient multi-turn LLM serving via small models, reducing costs and latency while maintaining high response quality.

Xueqi Cheng, Qiong Wu, Zhengyi Zhou et al.

2026-05-12 35
cs.CL 2605.10415

Aligning LLM Uncertainty with Human Disagreement in Subjectivity Analysis

Proposes DPUA framework combining disagreement perception and uncertainty calibration, improving reliability in subjective analysis.

Junyu Lu, Deyi Ji, Xuanyi Liu et al.

2026-05-11 24
cs.CL 2605.10114

SkillRAE: Agent Skill-Based Context Compilation for Retrieval-Augmented Execution

SkillRAE's two-stage skill graph and compact context improve task performance by 11.7%.

Xiangcheng Meng, Shu Wang, Yixiang Fang

2026-05-11 53
cs.CL 2605.09630

Scratchpad Patching: Decoupling Compute from Patch Size in Byte-Level Language Models

Scratchpad Patching method achieves near-baseline quality at 16-byte patches with 16x compute efficiency.

Lin Zheng, Vasilisa Bashlovkina, Timothy Dozat et al.

2026-05-11 4
cs.CL 2607.14118

Budgeted Subset Refinement for Execution-Aware LLM Research Ideation

Budgeted Subset Refinement improves execution-aware LLM research ideation, with MMR-k showing best results.

Micah Zhang

2026-05-09 2
cs.CL 2605.07725

SOD: Step-wise On-policy Distillation for Small Language Model Agents

SOD uses step-wise reweighted OPD to lift a 0.6B agent to 26.13% on AIME 2025.

Qiyong Zhong, Mao Zheng, Mingyang Song et al.

2026-05-08 39
cs.CL 2605.07110

Securing Computer-Use Agents: A Unified Architecture-Lifecycle Framework for Deployment-Grounded Reliability

Proposes a unified architecture-lifecycle framework to enhance CUA reliability in real environments, analyzing perception, decision, execution layers and stages of creation, deployment, operation, maintenance.

Zejian Chen, Zhanyuan Liu, Chaozhuo Li et al.

2026-05-08 47
cs.CL 2605.06276

Linear Semantic Segmentation for Low-Resource Spoken Dialects

Proposes a robust linear semantic segmentation model for low-resource spoken dialects, outperforming baselines with over 1000 multi-genre samples.

Kirill Chirkunov, Younes Samih, Abed Alhakim Freihat et al.

2026-05-07 53
cs.CL 2605.04831

StoryAlign: Evaluating and Training Reward Models for Story Generation

StoryAlign evaluates reward models for story generation using StoryRMB; StoryReward achieves 66.3% accuracy on StoryRMB.

Haotian Xia, Hao Peng, Yunjia Qi et al.

2026-05-06 1
Prev 1 ... 14 15 16 17 18 19 20 ... 90 Next

© 2026 GptGet.net - Paper Insights Platform

Paper List Submit Paper Help GptGet Home