From d494afbb14a1339a59da3c475352f8824034d521 Mon Sep 17 00:00:00 2001 From: zh-hanlabs <2257696485@qq.com> Date: Sun, 27 Sep 2026 19:21:15 +0800 Subject: [PATCH] docs: update ChatEval links after move to thunlp Signed-off-by: zh-hanlabs <2257696485@qq.com> --- README.md | 4 ++-- 1 file changed, 2 insertions(+), 2 deletions(-) diff --git a/README.md b/README.md index dd03c101..cde9d6f1 100644 --- a/README.md +++ b/README.md @@ -138,8 +138,8 @@ python agentverse_command/main_simulation_gui.py --task simulation/db_diag https://github.com/OpenBMB/AgentVerse/assets/11704492/c633419d-afbb-47d4-bb12-6bb512e7af3a -#### [Text Evaluation (ChatEval)](https://github.com/chanchimin/ChatEval) -In the context of the text evaluation scenario, we recommend users explore the [ChatEval](https://github.com/chanchimin/ChatEval) repo. They've implemented a multi-agent referee team on AgentVerse to assess the quality of text generated by different models. When given two distinct pieces of text, roles within ChatEval can autonomously debate the nuances and disparities, drawing upon their assigned personas, and subsequently provide their judgments. Experiments indicate that their referee team, enriched with diverse roles specified in [config.yaml](#2-configuring-the-agents), aligns more closely with human evaluations. This demo is built upon the [Fastchat](https://github.com/lm-sys/FastChat) repo, and we'd like to express our appreciation for their foundational work. +#### [Text Evaluation (ChatEval)](https://github.com/thunlp/ChatEval) +In the context of the text evaluation scenario, we recommend users explore the [ChatEval](https://github.com/thunlp/ChatEval) repo. They've implemented a multi-agent referee team on AgentVerse to assess the quality of text generated by different models. When given two distinct pieces of text, roles within ChatEval can autonomously debate the nuances and disparities, drawing upon their assigned personas, and subsequently provide their judgments. Experiments indicate that their referee team, enriched with diverse roles specified in [config.yaml](#2-configuring-the-agents), aligns more closely with human evaluations. This demo is built upon the [Fastchat](https://github.com/lm-sys/FastChat) repo, and we'd like to express our appreciation for their foundational work. https://github.com/OpenBMB/AgentVerse/assets/75533759/58f33468-f15b-4bac-ae01-8d0780019f85