prompt-eval-rubric: Score Your Agent's Outputs Without Paying for Another LLM Call

A Python library with weighted 0.0-1.0 scoring rubrics for LLM outputs. No LLM needed to run the rubrics.

Read Original

Related