arXiv 2501.15781

Large Language Models to Diffusion Finetuning

By Edoardo Cetin, Tianyu Zhao, et al.

Published 2025-01-27

Mindmap

Browse the paper's core ideas, clusters, and relationships in a structured outline.

We propose a new finetuning method to provide pre-trained large language models (LMs) the ability to scale test-time compute through the diffusion framework. By increasing the number of diffusion steps, we show our finetuned models achieve monotonically increasing accuracy, directly translating to improved performance across downstream tasks. Furthermore, our finetuned models can expertly answer questions on specifi…

View the original paper on arXiv