首页 > AI前沿 > Reason in Style: Discovering and Controlling Style in Language Models

Reason in Style: Discovering and Controlling Style in Language Models

arXiv自然语言 2026-10-01 05:15 6 阅读 查看原文

Language models learn content and style jointly, making stylistic variation in their outputs difficult to identify and control.

We study whether recurring styles in model responses can be discovered without supervision and explicitly controlled.

Methodology

We design an algorithm that learns to separate representations of content and style from language models' outputs and validate its effectiveness on math questions in a controlled setting.

By applying this method to over 100K verified traces from nine distinct teacher models, we discover six recurring yet imbalanced styles.

We then fine-tune smaller student models to follow these styles when explicitly conditioned on them, using importance weighting to balance the contribution of the styles represented in the corpus.

Results

This approach improves Pass@$k$ over standard fine-tuning on the same data across six math reasoning benchmarks, demonstrating that we can diversify the style of answers effectively.

We confirm that this also results in strong correspondence between requested and realized styles.

We find that style affects correctness: the probability of solving a problem depends on the style we condition on, and different problems benefit from different styles.

Conclusion

In summary, our results show that stylistic variation in model-generated data can be discovered in an unsupervised way, and made explicit, providing a source of both control and improved reasoning performance.