Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Feedback-to-Rubrics: Can We Learn Expert Criteria from Inline Comments?

About

Large language models (LLMs) are increasingly used for writing and review support, but their usefulness depends on context-dependent criteria, such as expert preferences or organization-specific conventions, that are often tacit, undocumented, and difficult to elicit directly. We propose a problem setting for learning reusable natural-language rubrics from accumulated inline comments on artifacts such as human-written or LLM-generated drafts. Our method infers rubrics from these comments and iteratively refines them by observing comment-wise mismatches between rubric-conditioned predictions and reference comments. We evaluate the proposed method in real-world review settings and in controlled settings with reference rubrics. These results show that inline comments can be distilled into reusable rubrics that support comment prediction, rubric understanding, and automatic artifact revision.

Kotaro Yoshida, So Kuroki, Yuki Imajuku, Taishi Nakamura, Ryunosuke Iwai, Haruki Goda, Takuya Akiba• 2026

Related benchmarks

TaskDatasetResultRank
Comment PredictionResearch Proposal Review (test)
Review Score3.42
6
Comment PredictionEssay Review (test)
Essay Review Score8.71
6
Comment PredictionHealthBench (test)
Medical Chat Annotation Score7.31
6
Comment PredictionExpertLongBench (test)
Bio Score2.75
6
Showing 4 of 4 rows

Other info

Follow for update