Awesome-LLM-Eval
Awesome-LLM-Eval: a curated list of tools, datasets/benchmark, demos, leaderboard, papers, docs and models, mainly for Evaluation on LLMs. 一个由工具、基准/数据、演示、排行榜和大模型等组成的精选列表,主要面向基础大模型评测,旨在探求生成式AI的技术边界.
- Type: Framework
- Imported from GitHub
- Popularity: 658 GitHub stars
- License: MIT
- Source: https://github.com/onejune2018/Awesome-LLM-Eval
- Repository: https://github.com/onejune2018/awesome-llm-eval
- Tags: awsome-list, awsome-lists, benchmark, bert, chatglm, chatgpt, dataset, evaluation, gpt3, large-language-model, leaderboard, llama
- Updated: 2026-10-01
More frameworks
- open-webui — User-friendly AI Interface (Supports Ollama, OpenAI API, ...)
- langchain — The agent engineering platform.
- awesome-llm-apps — 100+ AI Agents, Agent Skills and RAG Apps - Free and Open Source.
- graphify — Turn any codebase, with its docs, SQL schemas, configs, and PDFs, into a queryable knowledge graph. A /graphify skill for Claude Code, Curs…
- ragflow — RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with Agent capabilities to create…
- PaddleOCR — Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PD…
- crawl4ai — Open-source web crawler and scraper for LLMs and AI agents: any website into clean, LLM-ready Markdown. Run it yourself, or use Crawl4AI Cl…
- hello-agents — 📚 《从零开始构建智能体》——从零开始的智能体原理与实践教程