Baike.dev
All toolsAI codingTrendingOpen sourceNewsSubmit
Log in
Back to tool/Back to issues
#1467·llama

Verify evals on Papers with Code

Author: NielsRoggeCreated Sep 14, 2026Updated Sep 14, 2026

Hi,

Niels here from the open-source team at Hugging Face.

I've made the following papers and their evaluation results available on Papers with Code:

  • LLaMA: Open and Efficient Foundation Language Models — 15 paper-native evaluations.
  • Llama 2: Open Foundation and Fine-Tuned Chat Models — 7 paper-native evaluations.

The Llama 2 70B result currently ranks fourth on TriviaQA.

Would it be possible to verify these results and let me know if any score, model name, benchmark protocol, or openness metadata should be corrected?

You can also edit the task, methods, project page, and GitHub URL directly from each paper page using your Hugging Face account.

Kind regards,

Niels

Source: meta-llama/llama

View original on GitHubView discussion on GitHub