CONSONANCE.for your information
Monday, 5 October 2026frenvi

Worth reading closely

01 — agents 3 upvotes

OmniHarness: Harnessing Generalizable Visual Generation via Symbolic Policy Learning

QUESTION — How can visual generation generalizability be improved in MLLMs and multi-agent systems without fine-tuning model weights?

The authors introduce OmniHarness, a framework utilizing symbolic policy learning to improve generalizable visual generation in multimodal large language models and multi-agent systems. The framework abstracts verified executions into reusable symbolic policies, incorporates intermediate verification for failure recovery, and engages in self-directed inquiry through practice tasks near capability limits. While model parameters remain fixed, execution feedback continually refines the policies. Evaluated across six benchmarks, three MLLM backbones, and three visual agent frameworks, OmniHarness achieves a 95.0% resolve rate on ComfyBench's Creative tasks, outperforming the strongest baseline by 27.5 percentage points, and its frozen policy snapshots serve as plug-and-play enhancements for existing systems.

OmniHarness achieves a 95.0% resolve rate on ComfyBench's Creative tasks, exceeding the strongest baseline by 27.5 percentage points.

Experiments span across six benchmarks, three MLLM backbones, and three visual agent frameworks.

Model parameters remain fixed while execution feedback continually refines the policies.

BrandonLiu · 13 Sept 2026 read the original ↗
↑