|  | преди 5 месеца | |
|---|---|---|
| .. | ||
| evals_synthetic_data | преди 5 месеца | |
| inference | преди 8 месеца | |
| llm_eval_harness | преди 8 месеца | |
| README.md | преди 9 месеца | |
lm-evaluation-harness, a tool to evaluate Llama models including quantized models focusing on quality. We also included a recipe that calculates Llama 3.1 evaluation metrics Using lm-evaluation-harness and instructions that calculate HuggingFace Open LLM Leaderboard v2 metrics.