← map
○  learning

LLM-as-judge

—  ·  evaluation

Using a model to grade model output. Cheap and scalable, and quietly vulnerable to position bias and self-preference.

Nothing written yet. It's on the map because I want to understand it — the node exists so I can see the gap.