Open-source AI project: Awesome-LLM-Eval: a curated list of tools, datasets/benchmark, demos, leaderboard, papers, docs and models, mainly for Evaluation on LLMs. 一个由工具、基准/数据、演示、排行榜和大模型等组成的精选列表,主要面向基础大模型评测,旨在探求生成式AI的技术边界
Direct answer
Awesome-LLM-Eval is best for ai research. Source-backed open-source listing awaiting editorial review.
Trust signals
Dera’s take
Use Awesome-LLM-Eval to turn into .
Source-backed open-source listing awaiting editorial review. Best for AI Research.
Skip it if you need more control, a lower price, or a different output style.
Tasks where this tool may help.
We are still adding guides for this tool. Try a related task or check back later.
FAQ