← Back to all news
Hugging Face Blog April 16, 2024

Introducing the LiveCodeBench Leaderboard - Holistic and Contamination-Free Evaluation of Code LLMs

Read the original on Hugging Face Blog →

The Flow has not summarised this story yet — read it at Hugging Face Blog.

  • llms
  • benchmarks

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

Hugging Face Blog
May 4, 2023

StarCoder: A State-of-the-Art LLM for Code

llmsbenchmarks
More like this →
Hugging Face Blog
Apr 9, 2024

CodeGemma - an official Google release for code LLMs

llms
More like this →
Sebastian Raschka
Oct 5, 2025

Understanding the 4 Main Approaches to LLM Evaluation (From Scratch)

Multiple-Choice Benchmarks, Verifiers, Leaderboards, and LLM Judges with Code Examples

By Sebastian Raschka, PhD
llmsbenchmarks
More like this →
Hugging Face Blog
Dec 1, 2023

Open LLM Leaderboard: DROP deep dive

llmsbenchmarks
More like this →
Hugging Face Blog
Feb 20, 2024

Introducing the Open Ko-LLM Leaderboard: Leading the Korean LLM Evaluation Ecosystem

llmsbenchmarks
More like this →
arXiv AI
Jun 3

FLARE: Fine-Grained Diagnostic Feedback for LLM Code Refinement

arXiv:2606. 03852v1 Announce Type: cross Abstract: Large language models often generate code with bugs.

By Yinsheng Yao, Hongxiang Zhang, Weixi Tong, Tianyi Zhang
llms
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.0.0 · bb4ee0e