acceptodds
Under review as a conference paper at ICLR 2027

Understanding and Steering Supervised Fine-Tuning through Entropy Flow

Abstract

Supervised fine-tuning increases the likelihood of externally provided demon- strations, but likelihood improvement alone does not explain how supervision reshapes a model’s predictive distribution. We investigate this process through entropy flow: the local entropy responses to target-token supervision and their evolution throughout training. This perspective connects model-state-dependent selection and weighting with the allocation of supervision across tokens and ex- amples, while distinguishing local logit-space responses from their effects under shared model parameters. Our aim is to understand and steer fine-tuning through these evolving responses rather than to assess learning solely by the likelihood of observed tokens.

open until 14 Dec 2026

est. 32% chance this paper gets accepted at ICLR 2027.

Reject 68%Accept 32%

What do you think this paper will get?

All positions stay anonymous.

Related papers

Loading the map…

Discussion (0)

Sign in to comment.