AutoDesign: Meta-Harness Optimization for Long-Horizon Agentic Design
Abstract
AutoDesign uses a meta-harness optimizer to recursively improve a code agent for structured media generation, achieving state-of-the-art results on paper-to-poster synthesis.
Transforming multimodal sources into condensed and structured media outputs can be fundamentally conceptualized as a long-horizon agentic process centered on a model-harness system. While an ideal harness system should align with human design priors and accumulate reusable experience through empirical exploration to drive recursive self-improvement, existing paradigms remain static and fall short of this capability. In this paper, we present AutoDesign, a framework that aligns with human design priors, where a meta-harness optimizer guides a code agent to recursively improve harness based on rollout feedback. To instantiate and evaluate this framework, we focus on the academic paper-to-poster generation task and introduce PosterBench, comprising a 100-paper Main Track spanning five disciplines and PosterBench-mini, a shared 10-paper subset for controlled evaluation. On the PosterBench Main Track, AutoDesign achieves the highest score of 78.32, surpassing the closed-source commercial system Claude Design by 7.45 points. Across seven controlled code-agent-model configurations, integrating the learned DesignHarness consistently improves performance, increasing the average PosterBench Score from 54.99 to 67.39 (+12.4%). In a fully autonomous long-horizon loop, it executes 253 tool calls and 11 editing turns within 40 minutes for under $3, reaching average conference-poster quality in human evaluation. A system-blind human study further demonstrates that AutoDesign achieves the highest human preference among evaluated systems.
Community
We also provide live demo at: https://designanything.ai/, though, we recommend to locally install for the best experience.
We also welcome the community to submit issues or propose PR, together, we can continously improve autodesign.
This is an automated message from the Librarian Bot. I found the following papers similar to this paper.
The following papers were recommended by the Semantic Scholar API
- ADIAS: Automated Design of Interactive Agentic Systems (2026)
- Self-Evolving Embodied Agents via Skill-Harness Evolution (2026)
- Hierarchical Self-Improvement: A Framework for Task-Specific Evolvable Agent Harnesses (2026)
- Harness-Aware Self-Evolving: Co-Evolving Model Weights, Harness, and Task Solutions (2026)
- Rethinking the Evaluation of Harness Evolution for Agents (2026)
- Evo-Bench: Can Language Models Improve Agent Harness? (2026)
- OneDayAgent: Towards a Long-Horizon Harness for Autonomous Agents (2026)
Please give a thumbs up to this comment if you found it helpful!
If you want recommendations for any Paper on Hugging Face checkout this Space
You can directly ask Librarian Bot for paper recommendations by tagging it in a comment: @librarian-bot recommend
Get this paper in your agent:
hf papers read 2608.13560 Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash Models citing this paper 0
No model linking this paper
Datasets citing this paper 2
YaxinLuo/PosterBench
Spaces citing this paper 0
No Space linking this paper
Collections including this paper 0
No Collection including this paper
