{% extends "base.html" %} {% block title %}Judges — SimpleAudit{% endblock %} {% block content %}
A judge is how conversations are graded: the criteria (what to evaluate), the output format (severity, score, yes/no or checklist) and the probe prompt the auditor follows. Start from one of SimpleAudit's judges, or write your own criteria. You pick the model that grades on New Experiment. Editing a judge saves a new version; runs keep the version they used.
{{ j.description }}
{% endif %}Nothing matches this filter.
{% else %}Set up your judges
Start with SimpleAudit's judges: safety (its default), harm, helpfulness, factuality, abstention, checklist and more. You can edit their criteria or clone them later.