acceptodds
Under review as a conference paper at ICLR 2027

Coverage Before Compression: Online Call Scheduling for Repeated Long-Context Review

Abstract

Repeated long-context review amortizes candidate proposal while incurring target-model verification cost on every case, making call scheduling a first-order online design variable. We introduce policy-isolated verification, a paired estimand that holds the output, span partition, candidate order, target model, and call cap fixed, together with EvidenceCall, a coalition-state policy organized by coverage before compression. EvidenceCall tests necessity and direction, continues from verified coalitions, and delays redundancy removal until evidence coverage is secured. At an identical cap over RULER-128K, MuSiQue, and QuALITY examples, it improves decisive-span P@5 by over Greedy Minimal-Set, with of the margin retained under a literal ranking freeze. Across three frozen proposers, it also leads outside-family MCTS-Subset by to P@5, and every paired-bootstrap lower bound is positive. The identified gain transfers to decisions: reviewer yield rises by in a -reviewer mechanism-isolation crossover and by in a -developer code deployment, while a disjoint -reviewer citation deployment improves accepted findings from to per hour over MCTS-Subset. Blind full-context adjudication, external proposer supervision, three held-out benchmarks, and two additional target families preserve the effect. The results establish coverage-before-compression scheduling as a consequential systems principle for repeated review.

Then back it, or bet against it.

Related papers

Open the market on this paper to see 7 more related papers.