grassr: rater reliability on binary outcomes, from rating matrix to Report Card
The data: a rater panel | The problem: the label tracks the study, not the raters | What grassr reports instead | Case studies | Three labels, one position | Two panels the fixed bands call identical | A divergent panel and the per-rater route | Two raters | Layered access | Building blocks | Reference resolution and the clamp guard | The machinery | Extending the calibration | References