LLM-Eval-Survey
Visit ToolLLM-eval-survey is an academic research tool that provides a comprehensive collection of papers and resources for evaluating large language models. It serves as an official GitHub page for a survey paper on LLM evaluation.
LLM-eval-survey is an academic research tool that provides a comprehensive collection of papers and resources for evaluating large language models. It serves as an official GitHub page for a survey paper on LLM evaluation.
About
LLM-eval-survey is the official GitHub page for the survey paper "A Survey on Evaluation of Large Language Models." It functions as a central repository for researchers and practitioners, offering a curated collection of papers and resources focused on the evaluation of large language models (LLMs). The repository organizes papers by various evaluation aspects, including natural language processing tasks (understanding, sentiment analysis, text classification, inference, reasoning, generation, summarization, dialogue, translation, question answering), robustness, ethics, biases, trustworthiness, and applications in social science, natural science, engineering, medical, and agent domains. It also provides updates on new paper versions and welcomes community contributions.
Capabilities
Pricing & Plans
Open Source
Free
FAQs