GptGet
Features PaperForge Apps Papers Blog Contact AI Chat 中文
Sort: Latest Popular Citations
All Artificial Intelligence Computation and Language Computer Vision Information Retrieval Machine Learning Machine Learning (Stats) Neural and Evolutionary Computing Robotics
cs.LG 2410.09240

nach0-pc: Multi-task Language Model with Molecular Point Cloud Encoder

nach0-pc integrates point cloud encoder with T5, enabling multi-task 3D molecular generation with high efficiency.

Maksim Kuznetsov, Airat Valiev, Alex Aliper et al.

2024-10-12 45
cs.CL 2410.08971

Extra Global Attention Designation Using Keyword Detection in Sparse Transformer Architectures

EGAD enhances Longformer with keyword-based global attention, boosting long-range dependency modeling for abstractive summarization.

Evan Lucas, Dylan Kangas, Timothy C Havens

2024-10-12 51
cs.LG 2410.08925

An Overview of Prototype Formulations for Interpretable Deep Learning

HyperPG outperforms Euclidean prototypes on CUB-200-2011 with simplified training.

Maximilian Xiling Li, Korbinian Franz Rudolf, Paul Mattes et al.

2024-10-11 48
cs.LG 2410.14716

A Systematic Survey on Large Language Models for Algorithm Design

This paper systematically reviews the roles of large language models in algorithm design, including optimizer, predictor, extractor, and designer, across three stages.

Fei Liu, Yiming Yao, Ping Guo et al.

2024-10-11 28
cs.LG 2410.08432

MYCROFT: Towards Effective and Efficient External Data Augmentation

Mycroft combines feature distance and gradient matching to efficiently select informative data subsets from private sources, boosting model performance with minimal data sharing.

Zain Sarwar, Van Tran, Arjun Nitin Bhagoji et al.

2024-10-11 47
cs.CV 2410.08202

Mono-InternVL: Pushing the Boundaries of Monolithic Multimodal Large Language Models with Endogenous Visual Pre-training

Mono-InternVL embeds visual experts into a pre-trained LLM with EViP, achieving superior multi-modal performance, surpassing 13 benchmarks.

Gen Luo, Xue Yang, Wenhan Dou et al.

2024-10-11 48
cs.CL 2410.08197

From Exploration to Mastery: Enabling LLMs to Master Tools via Self-Driven Interactions

DRAFT employs self-driven feedback to iteratively refine tool documentation, enhancing LLM understanding and utilization.

Changle Qu, Sunhao Dai, Xiaochi Wei et al.

2024-10-11 44
cs.CV 2410.08189

SG-Nav: Online 3D Scene Graph Prompting for LLM-based Zero-shot Object Navigation

SG-Nav uses online 3D scene graphs and hierarchical reasoning to boost zero-shot object navigation SR by over 10%.

Hang Yin, Xiuwei Xu, Zhenyu Wu et al.

2024-10-11 54
cs.RO 2410.07864

RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

RDT-1B model uses diffusion and Transformer to learn bimanual manipulation with multimodal data.

Songming Liu, Lingxuan Wu, Bangguo Li et al.

2024-10-10 22
cs.CV 2410.07838

Minority-Focused Text-to-Image Generation via Prompt Optimization

Proposes MinorityPrompt, an online prompt optimization method enhancing low-likelihood sample generation in diffusion models.

Soobin Um, Jong Chul Ye

2024-10-10 34
cs.CV 2410.07171

IterComp: Iterative Composition-Aware Feedback Learning from Model Gallery for Text-to-Image Generation

IterComp leverages multi-model preferences and iterative feedback to enhance compositional text-to-image generation, outperforming SOTA methods.

Xinchen Zhang, Ling Yang, Guohao Li et al.

2024-10-10 32
cs.CV 2410.07296

ReinDiffuse: Crafting Physically Plausible Motions with Reinforced Diffusion Model

ReinDiffuse combines diffusion models with reinforcement learning, achieving 29% and 34% improvements in FID on HumanML3D and KIT-ML, respectively, for physically plausible motion.

Gaoge Han, Mingjiang Liang, Jinglei Tang et al.

2024-10-10 30
cs.CV 2410.06756

DreamMesh4D: Video-to-4D Generation with Sparse-Controlled Gaussian-Mesh Hybrid Representation

DreamMesh4D combines mesh and Gaussian points for high-quality video-to-4D generation, achieving superior spatial-temporal consistency.

Zhiqi Li, Yiming Chen, Peidong Liu

2024-10-09 31
cs.CL 2410.10878

Herald: A Natural Language Annotated Lean 4 Dataset

Herald: Translates Mathlib4 to natural language using dual augmentation, enhancing LLM performance in mathematical reasoning.

Guoxiong Gao, Yutong Wang, Jiedong Jiang et al.

2024-10-09 47
cs.LG 2410.06665

Revisiting Multi-Permutation Equivariance through the Lens of Irreducible Representations

Using irreducible representations and Schur’s lemma, this paper systematically characterizes permutation-equivariant layers, simplifying derivations for DeepSets, graph networks, and weight spaces.

Yonatan Sverdlov, Ido Springer, Nadav Dym

2024-10-09 51
cs.CL 2410.06511

TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

TorchTitan supports 4D parallelism, boosting Llama 3.1 training by 65.08% with modular design and hardware co-optimization.

Wanchao Liang, Tianyu Liu, Less Wright et al.

2024-10-09 57
cs.CL 2410.06153

AgentSquare: Automatic LLM Agent Search in Modular Design Space

Proposes MoLAS and AgentSquare, using module evolution and recombination to optimize LLM agents with a 17.2% performance boost.

Yu Shang, Yu Li, Keyu Zhao et al.

2024-10-08 37
cs.CV 2410.06055

RepLDM: Reprogramming Pretrained Latent Diffusion Models for High-Quality, High-Efficiency, High-Resolution Image Generation

RepLDM achieves high-quality, efficient high-resolution image generation via attention guidance and progressive upsampling.

Boyuan Cao, Jiaxin Ye, Yujie Wei et al.

2024-10-08 3
stat.ML 2410.05602

Amortized Control of Continuous State Space Feynman-Kac Model for Irregular Time Series

ACSSM combines multi-marginal Doob transform and stochastic optimal control for irregular time series modeling.

Byoungwoo Park, Hyungi Lee, Juho Lee

2024-10-08 53
cs.RO 2410.05582

Gen-Drive: Enhancing Diffusion Generative Driving Policies with Reward Modeling and Reinforcement Learning Fine-tuning

Proposes Gen-Drive, a diffusion-based generative driving policy with reward modeling and RL fine-tuning, achieving state-of-the-art planning scores on nuPlan.

Zhiyu Huang, Xinshuo Weng, Maximilian Igl et al.

2024-10-08 36
Prev 1 ... 329 330 331 332 333 334 335 ... 573 Next

© 2026 GptGet.net - Paper Insights Platform

Paper List Submit Paper Help GptGet Home