EleutherAI/lm-evaluation-harness

reproduce llama 3 evals

Open

#2,557 opened on Dec 10, 2024

View on GitHub
 (8 comments) (2 reactions) (0 assignees)Python (3,306 forks)auto 404
good first issuevalidation

Repository metrics

Stars
 (12,755 stars)
PR merge metrics
 (Avg merge 15d 7h) (11 merged PRs in 30d)

Description

llama 3.{1,2,3} have released most of their eval details here in HF and also some in this repo. Would be great if we can upstream them here. I'm adding some in #2556

Contributor guide