CrossDepth: Geometry-Constrained Attention for Generalizable Multi-View Surround Depth Estimation
CrossDepth improves multi-view depth estimation accuracy and consistency using geometry-constrained attention.
Samer Abualhanud, Max Mehltretter
CrossDepth improves multi-view depth estimation accuracy and consistency using geometry-constrained attention.
Samer Abualhanud, Max Mehltretter
Introduced IIns-GAN for generating labeled wireless signals, enhancing model training.
Yuxiao Li, Keke Hu, Santiago Mazuelas et al.
EDGE framework synthesizes tool-calling data using dynamic graphs, enhancing performance.
Dain Kim, Eungi Cho, Kyumin Kim et al.
Evaluates LLM explanations' necessity and sufficiency using behavioral evidence, finding limited correlation with model behavior.
Urja Pawar, Rajitha Ramanayake, Nabeel Kemal et al.
Ref-GeNVS is a training-free method for generating reflection-consistent novel views in mirror scenes.
GeonU Kim, Shin Dong-Yeon, Tae-Hyun Oh
Study finds widespread verbatim retrieval in LLMs on molecular regression benchmarks, affecting prediction accuracy.
Matthias Busch, Marius Tacke, Sviatlana V. Lamaka et al.
Partial balayage and logarithmic-time Mellin analysis yield the weak-(1,1) bound O(√n log n).
Daniel Spector, Cody B. Stockdale
Using ACT, this study examines visual distractors' impact on imitation policies, significantly improving UR3e robustness.
Vivek Chavan, Pengtao Xie, Yahuan Shi et al.
CUA-Universe uses App-Forge, Task-Weave, and Path-Steer to create hybrid GUI+CLI environments, enhancing efficiency and success rates.
Haoting Shi, Wenhao Wang, Weicheng Fang et al.
Decompile-Diverge detects behavioral divergence in LLM decompilers, revealing a 13% divergence rate.
Chang Liu, Edward Raff, Kristopher Micinski
Proposes a neuro-symbolic framework combining VLA control and task graphs for long-horizon vision-language-action manipulation.
Vivek Chavan, Yahuan Shi, Oliver Heimann et al.
Introduces a two-level framework using reasoning distillation and product-type test-time training to enhance scalable recommendation, achieving AUC of 0.924.
Siliang Liu, Mohammad Ghasemi, Sapan Patel et al.
Developed a humanoid robot prototype for multimodal HRI, achieving 96% gesture recognition accuracy.
Thang Tran Viet, Thanh Nguyen Canh, Huy Uong Gia et al.
The study examines memory portability during model upgrades, finding fixed-schema knowledge graphs remain stable.
Ankit Goyal, Jaideep Ray
Introduced a Hessian-based variational continuation method to automate finding periodic orbits in double pendulum systems.
Leo Yao, Ziming Liu, Max Tegmark
Proposed H-BAC framework combines quantization and knowledge distillation, achieving 95.13% accuracy and 54.5x model compression.
Mahadev Sunil Kumar, Bhavika Gondi, Desaisetty Venkata Satya Sai Swapnith et al.
Toolkit for measuring contextual individuation in Transformer models using bridge forms.
José Luciano Verçosa Marques, Frederico Jorge Heitmann, Daniel Omar Perez et al.
Behavior Trees lack adaptability in robotic systems, needing enhancements for dynamic environments.
Mehran Rostamnia, Gianluca Filippone, Ricardo Caldas et al.
EGF generates categorical graphs via continuous embeddings, achieving 0.150 FCD on QM9.
Ethan Ma, Zihan Wang, Chris Siu Yeung Chow et al.
FIRE-LIVWO achieves robust LiDAR-Inertial-Visual-Wheel Odometry with mmWave radar enhancement, average localization error of 5.677m.
Kun Hu, Menggang Li, Kaidi Wu et al.