Ike, E. S., Takeuchi, J., Joublin, F., Ceravola, A., & Tanti, M. (2025). Automating Dialogue Evaluation: LLMs Vs Human Judgment.
Artificial Intelligence in HCI. HCII 2025, Lecture Notes in Artificial Intelligence,
15820, 353-372. Cham: Springer.
https://doi.org/10.1007/978-3-031-93415-5_21