A Python library and a set of instruments for measuring reward models, training runs and the graders in between. It is built around one rule: a reading is either evidence carrying a trust level that was computed rather than claimed, or a refusal carrying a remedy.

Mohammed Suhail B Nadaf

The conditions this was built under

No independent reproduction has been recorded.

Where to go from here