This paper develops a new framework for evaluating and selecting document sets for AI agents, considering the interactions between documents, and proposes a training-free method that achieves the best downstream generation performance with fewer documents and search rounds.
Firehose
Filtered to Papers, tagged “rubric-based evaluation” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News
Browse by tag
This paper introduces EduPanel, a machine learning model that evaluates the quality of teaching videos in a more nuanced way than existing methods, and shows that it can provide reliable and interpretable assessments that complement human expertise.