Preference Tuning LLMs with Direct Preference Optimization Methods
Hugging Face Blog
· huggingface.co · January 18, 2024
Read on Hugging Face Blog →
Opens the original article in a new tab.
More from the blog
Hugging Face Blog
· huggingface.co
Sep 3, 2026
Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps
MIT News · Artificial Intelligence
· news.mit.edu
Sep 2, 2026
System helps humans predict when self-driving cars will make mistakes
A new method, called CW-Net, translates the reasoning process of an autonomous vehicle’s AI system into understandable concepts that explain its behavior.
MIT News · Artificial Intelligence
· news.mit.edu
Sep 1, 2026
Walter Torous named executive director of MIT Center for Real Estate
The senior lecturer, already director of the degree program, will now oversee all aspects of the center’s activities and operations.
Hugging Face Blog
· huggingface.co
Aug 25, 2026
Granite 4.2 LLMs: How They're Built
A Blog post by IBM Granite on Hugging Face