Verify evals on Papers with Code
Author: NielsRoggeCreated Sep 14, 2026Updated Sep 14, 2026
Hi,
Niels here from the open-source team at Hugging Face.
I've made the following papers and their evaluation results available on Papers with Code:
- LLaMA: Open and Efficient Foundation Language Models — 15 paper-native evaluations.
- Llama 2: Open Foundation and Fine-Tuned Chat Models — 7 paper-native evaluations.
The Llama 2 70B result currently ranks fourth on TriviaQA.
Would it be possible to verify these results and let me know if any score, model name, benchmark protocol, or openness metadata should be corrected?
You can also edit the task, methods, project page, and GitHub URL directly from each paper page using your Hugging Face account.
Kind regards,
Niels
Source: meta-llama/llama