Add evaluation results to llama2-7b-sandbox

#5
Owner

Add structured evaluation results from Open LLM Leaderboard benchmarks (MMLU, HellaSwag, ARC, GSM8K, TruthfulQA, Winogrande)

pwxcdct changed pull request status to closed

Sign up or log in to comment