Awesome-LLM-Eval: a curated list of tools, datasets/benchmark, demos, leaderboard, papers, docs and models, mainly for Evaluation on LLMs. 一个由工具、基准/数据、演示、排行榜和大模型等组成的精选列表,主要面向基础大模型评测,旨在探求生成式AI的技术边界.
Explore a collection of tools and resources for evaluating large language models.
Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.
Get an email alert on its next release or when it starts trending — never miss the moment.
Free · no card · unsubscribe anytimeAwesome-LLM-Eval: a curated list of tools, datasets/benchmark, demos, leaderboard, papers, docs and models, mainly for Evaluation on LLMs. 一个由工具、基准/数据、演示、排行榜和大模型等组成的精选列表,主要面向基础大模型评测,旨在探求生成式AI的技术边界.
Awesome-LLM-Eval has 651 stars on GitHub. It has been forked 80 times. It has been in active development since 2023. Awesome-LLM-Eval is available under the MIT license. Its main topics are awsome-list, awsome-lists, benchmark, bert.
Awesome-LLM-Eval: a curated list of tools, datasets/benchmark, demos, leaderboard, papers, docs and models, mainly for Evaluation on LLMs. 一个由工具、基准/数据、演示、排行榜和大模型等组成的精选列表,主要面向基础大模型评测,旨在探求生成式AI的技术边界.
Awesome-LLM-Eval is an open-source project. It is released under the MIT license.
Yes. Awesome-LLM-Eval is free and open source — you can use, modify and self-host it.
Awesome-LLM-Eval is available under the MIT license.
Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.
[](https://olud.ai/project/onejune2018-awesome-llm-eval.html)
Measured from GitHub topics shared by both projects, weighted by how rare each topic is.