https://openreview.net/forum?id=7UqQJUKaLM
xFinder: Large Language Models as Automated Evaluators for Reliable Evaluation | OpenReview
The continuous advancement of large language models (LLMs) has brought increasing attention to the critical issue of developing fair and reliable methods for...
large language models