← Back to all news
Hugging Face Blog August 8, 2023

Fine-tune Llama 2 with DPO

Read the original on Hugging Face Blog →

The Flow has not summarised this story yet — read it at Hugging Face Blog.

  • llms
  • reinforcement-learning
  • fine-tuning

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

Hugging Face Blog
Apr 5, 2023

StackLLaMA: A hands-on guide to train LLaMA with RLHF

llmsreinforcement-learning
More like this →
Hugging Face Blog
Sep 25, 2024

Llama can now see and run on your device - welcome Llama 3.2

llms
More like this →
Hugging Face Blog
Jul 18, 2023

Llama 2 is here - get it on Hugging Face

llms
More like this →
Hugging Face Blog
Aug 25, 2023

Code Llama: Llama 2 learns to code

llms
More like this →
Hugging Face Blog
Nov 7, 2023

Make your llama generation time fly with AWS Inferentia2

llms
More like this →
Hugging Face Blog
Apr 18, 2024

Welcome Llama 3 - Meta's new open LLM

llms
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.0.0 · bb4ee0e