Papers

Filtered to rubric-based evaluation · clear filter

Browse by term

continual learning 33reinforcement learning 21large language models 9vision-language models 8language models 6benchmarking 5autoregressive models 4diffusion transformers 4generative models 4multimodal learning 4robotics 4video generation 4benchmarks 3computer vision 3policy optimization 3self-distillation 3vision-language-action models 3world modeling 3agent-based systems 2autonomous agents 2calibration 2coding agents 2diffusion models 2foundation models 2image synthesis 2in-context learning 2knowledge graphs 2LLMs 2multimodal evaluation 2multimodal large language models 2

Matching papers

Beyond Relevance-Centric Retrieval: Rubric-Oriented Document Set Selection and Ranking

27 upvotes · 22 JUL 2026 · Kailin Jiang, Lei Liu, Jian Xi et al.

This paper develops a new framework for evaluating and selecting document sets for AI agents, considering the interactions between documents, and proposes a training-free method that achieves the best downstream generation performance with fewer documents and search rounds.

EduPanel: A Three-Agent LLM Judge for Teaching Videos -- Reliability, Complementarity, and Human Trust Calibration

4 upvotes · 20 JUL 2026 · Jia-Kai Dong, Yi-Cheng Lin, Hung-yi Lee

This paper introduces EduPanel, a machine learning model that evaluates the quality of teaching videos in a more nuanced way than existing methods, and shows that it can provide reliable and interpretable assessments that complement human expertise.