Generative Interpretation: Toward Human-Like Evaluation for Educational Question-Answer Pair Generation
Generative Interpretation: Toward Human-Like Evaluation for Educational Question-Answer Pair Generation
Hyeonseok Moon,Jaewook Lee,3 Authors,Heu-Jeoung Lim
2024 · DBLP: conf/eacl/MoonLEPSL24
Findings · 7 Citations
TLDR
Through experimental analysis, it is revealed that GI outperforms existing evaluation methods in terms of human alignment, and even shows comparable performance with GPT3.5, only with BART-large.
