Pre trained evaluator in EMNLP 2022 paper Towards a Unified Multi Dimensional Evaluator for Text Generation Introduction Multi dimensional evaluation is the dominant paradigm for human evaluation in Natural Language Generation (NLG), i.e., evaluating the generated text from multiple explainable dimensions, such as coherence and fluency. However, automatic evaluation in NLG is still dominated by similarity based metrics (e.g., ROUGE, BLEU), but they are not sufficient to portray the difference between the advanced generation models. Therefore, we propose UniEval to bridge this gap so that a more comprehensive and fine grained evaluation of NLG systems can be achieved. Pre trained Evaluator unieval fact is the pre trained evaluator for the factual consistency detection task. It can evaluate the model output and predict a consistency score. Usage Please refer to our GitHub repository.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy