cs.AI 2606.13707

Orchestra-o1: Omnimodal Agent Orchestration

Orchestra-o1 framework enhances multimodal agent collaboration, achieving 10.3% accuracy improvement on the OmniGAIA benchmark.

Fan Zhang, Vireo Zhang, Shengju Qian et al.

2026-06-10 4
cs.AI 2606.11173

The Role of Feedback Alignment in Self-Distillation

This paper introduces feedback alignment in self-distillation, comparing three feedback types; structure-aligned critique outperforms others with +16.11% accuracy.

Semih Kara, Oğuzhan Ersoy

2026-06-10 240
cs.AI 2606.11078

A History-Aware Visually Grounded Critic for Computer Use Agents

Proposes HiViG, a history-aware visually grounded test-time framework, boosting GUI task success rates by 5.8% (Qwen3-VL-32B) and 9% (Gemini-3-Flash) through macro-action history and visual error verification.

Jaewoo Lee, Zaid Khan, Archiki Prasad et al.

2026-06-10 269
cs.AI 2606.20659

Skill Coverage: A Test Adequacy Metric for Agent Skills

Proposes skill coverage metric based on trajectory detection of skill behavior constraints, achieving 38.66%-45.51% coverage; improves failed task recovery by 16%.

Boyin Tan, Xiaowei Huang, Youcheng Sun

2026-06-09 45
cs.AI 2606.06741

OpenSkill: Open-World Self-Evolution for LLM Agents

OpenSkill framework enhances LLM agents' skill transfer in open-world settings without supervision, achieving top automated pass rates.

Zhiling Yan, Dingjie Song, Hanrong Zhang et al.

2026-06-05 5