|
4 ماه پیش | |
---|---|---|
.. | ||
inference | 9 ماه پیش | |
llm_eval_harness | 4 ماه پیش | |
README.md | 6 ماه پیش |
lm-evaluation-harness
, a tool to evaluate Llama models including quantized models focusing on quality. We also included a recipe that calculates Llama 3.1 evaluation metrics Using lm-evaluation-harness
and instructions that calculate HuggingFace Open LLM Leaderboard v2 metrics.