cs.CL 2304.06556

Are LLMs All You Need for Task-Oriented Dialogue?

This study evaluates pre-trained LLMs in task-oriented dialogue, showing limited belief state tracking but effective dialogue guidance, improved by true belief states and few-shot examples.

Vojtěch Hudeček, Ondřej Dušek

2023-04-13 46
cs.CV 2304.05497

Revisiting Single-gated Mixtures of Experts

Single-gate MoE achieves comparable efficiency and accuracy to complex models, outperforming non-mixture baselines.

Amelie Royer, Ilia Karmanov, Andrii Skliar et al.

2023-04-12 39
cs.LG 2304.05055

A Comprehensive Survey on Deep Graph Representation Learning

This survey systematically reviews deep graph representation learning architectures, paradigms, and applications, highlighting GNN innovations and future challenges.

Wei Ju, Zheng Fang, Yiyang Gu et al.

2023-04-11 351 citations 32
cs.HC 2304.03442

Generative Agents: Interactive Simulacra of Human Behavior

A novel architecture combining large language models with dynamic memory, reflection, and planning enables believable human-like agent behavior with long-term coherence.

Joon Sung Park, Joseph C. O'Brien, Carrie J. Cai et al.

2023-04-07 48
cs.CV 2304.02643

Segment Anything

SAM model supports prompt-based image segmentation with over 1 billion masks, achieving state-of-the-art zero-shot performance.

Alexander Kirillov, Eric Mintun, Nikhila Ravi et al.

2023-04-06 34