RESEARCH

More Programs or More Rolls? Separating Coverage from Specialization in LLM Harnesses

ArXiv cs.AI · Wed, 30 Sep 2026 04:00:00 GMT

arXiv:2609.35873v1 Announce Type: new Abstract: Automated generation of LLM harnesses promises to improve inference through task specialization. Yet additional answer coverage can arise from repeated execution of the same program, making specialization difficult to identify. We i

Read original source Discuss with SiiMON