Multi-agent-as-judge: Aligning llm-agent-based automated evaluation with multi-dimensional human evaluation
Unknown
This paper proposes a multi-agent evaluation framework that aligns LLM-agent-based automated evaluation with multi-dimensional human evaluation for real-world text generation.