Steering Conversion Rate vs Model Scale
In-sample separability stays high while causal leverage varies
Steering conversion rate, behavioral lift as a fraction of available headroom, plotted against log parameter count for Qwen, Gemma, and Llama instruct models. Qwen falls to zero at 14 and 32 billion parameters before a partial recovery to 0.179 at 72 billion; Gemma reaches zero at 27 billion; Llama drops from 0.732 at 8 billion to −0.429 at 70 billion, where the star and dashed segment flag degenerate output rather than resistance.