AdaRubric: Adaptive Dynamic Rubric Evaluator for Agent Trajectories
-
Updated
Aug 24, 2026 - Python
AdaRubric: Adaptive Dynamic Rubric Evaluator for Agent Trajectories
A local-first, evidence-linked assignment planner.
Open-source self-hosted web tool for evaluating Agent Skills with rubric scores, Deep Review, and improvement suggestions.
A survey of rubrics across the evolving LLM landscape.
Hermes Agent skill for grading SOUL.md identity files with a research-backed rubric and public-safe field guide.
Open Rubric System: Scaling Reinforcement Learning with Pairwise Adaptive Rubric
Run Pro Codex skill stack for project-folder execution, academic delivery, audit, revision, and paper-to-PPT workflows.
Reward model engineering harness for evolutionary rubric search, deployable RM artifacts, online scoring, and RL experiment lineage.
A Claude Code skill that adds a rubric-based eval layer to any agent project. Framework-agnostic — generates rubric, test cases, judge prompt, and harness. Returns a weighted score plus a judge-leniency signal.
Export grades from assignment using advanced grading methods in excel format
Rubric-driven AI homework grading system built as a Claude Code Skill. Score student submissions with CoT reasoning, bias mitigation, and PDCA quality cycle.
Evaluate Claude Code and Codex skill directories with deterministic rubric checks and graded, fixable reports.
AskBench: LLM question-asking/clarification benchmark & dataset with evaluation and training code (paper: arXiv 2602.11199).
Universal quality evaluation plugin for Claude Code �� 7-dimension scoring (correctness, completeness, adherence, efficiency, safety), configurable rubrics, threshold blocking, auto-hooks & /judge command.
To associate your repository with the rubric topic, visit your repo's landing page and select "manage topics."