2 articles
A breakthrough inference technique called LoPA lets large language models generate text three times faster by processing multiple tokens simultaneously. The plug-and-play method requires no additional training and has shown exceptional results across coding and reasoning tasks.
The new dLLM open-source library simplifies training and deployment of diffusion-based language models by consolidating fragmented workflows into a unified framework with integrated training, evaluation, and scaling tools.
We use cookies to improve your experience on our site and to show you relevant advertising. To find our more, read our privacy policy and cookie policy