acceptodds
Under review as a conference paper at ICLR 2027

Swap-RoPE: Permutation-Invariant Block Encoding for LLM Judging

Abstract

Pairwise LLM judging repeatedly encodes the same candidates in different positions and contexts. We introduce \name, which combines shared block coordinates with causal attention restricted to a fixed prefix and each block's own history. We establish that a block's hidden states, keys, and values are independent of its competitors and preserved under token-aligned permutations. These properties permit cache reuse across different comparisons. In an all-pairs tournament over candidates, fresh candidate encodings decrease from to under fixed-content and cache-retention assumptions; comparison calls remain quadratic. We fine-tune Qwen3 models at 1.7B, 4B, and 8B and evaluate three mathematical benchmarks. Reported reusable key–value (KV) cache ratios increase from – under standard RoPE to –. Selection accuracy improves over vanilla-position fine-tuning in six of nine settings with four candidates and three of nine with eight candidates. The construction excludes cross-block attention during candidate encoding; verdict consistency and serving speedups require further evaluation.

Then back it, or bet against it.

Related papers

Open the market on this paper to see 7 more related papers.