cs.CV 2409.16709

Pose-Guided Fine-Grained Sign Language Video Generation

Proposed a Pose-Guided Motion Model to generate fine-grained, motion-consistent sign language videos, significantly improving detail and temporal consistency.

Tongkai Shi, Lianyu Hu, Fanhua Shang et al.

2024-09-25 22
cs.AI 2409.18807

LLM With Tools: A Survey

Proposes a standardized framework for tool integration in LLMs, combining fine-tuning and in-context learning to enhance complex task performance.

Zhuocheng Shen

2024-09-24 49
cs.RO 2409.14562

DROP: Dexterous Reorientation via Online Planning

DROP employs online sampling predictive control with vision-based pose estimation to reorient objects, achieving performance comparable to RL methods without extensive training.

Albert H. Li, Preston Culbertson, Vince Kurtz et al.

2024-09-23 27
cs.CL 2409.13265

Towards LifeSpan Cognitive Systems

Proposes LSCS architecture combining model parameters, explicit memory, knowledge graphs, and text storage for lifelong experience absorption and accurate recall.

Yu Wang, Chi Han, Tongtong Wu et al.

2024-09-20 52