CONSONANCE.for your information
Monday, 5 October 2026frenvi

Worth reading closely

01 — agents 23 upvotes

Omni-Decision: Evidence-Ledger Planning for Omni-Modal Agents

QUESTION — How can accumulated observation noise in the conversation history of omni-modal agents be mitigated?

Omni-modal agents struggle with multi-step planning because noisy observations accumulate in the conversation history. The authors propose Omni-Decision, an agent built on evidence-ledger planning that replaces growing dialogue histories with an explicit ledger tracking confirmed evidence and conflicts. A critic filters raw observations, passing only usable content so the planner operates on a compact context. Through supervised fine-tuning and decision-level reinforcement learning, Omni-Decision achieves 81.4% accuracy on OmniGAIA at roughly 43% of Gemini-3.1-Pro's cost per question, and ties with top end-to-end models on WorldSense long-video understanding.

Achieves state-of-the-art accuracy of 81.4% on OmniGAIA at approximately 43% of Gemini-3.1-Pro's cost per question.

Scores 65.0% on WorldSense long-video understanding, tying with the strongest end-to-end model.

MBJinX · 24 Sept 2026 read the original ↗
↑