When asked, large language models~(LLMs) like ChatGPT claim that they can assist with relevance judgments but it is not clear whether automated judgments can reliably be used in evaluations of retrieval systems. In this perspectives paper, we discuss possible ways for~LLMs to support relevance judgm...
Research Assistant
AI chat, annotations, notes & similar papers
No comments yet
Be the first to share your thoughts!