In my work on LLM as a judge, I prefer to use LLM decisions as features in a downstream classic ML model for the final decision. It works really well

https://softwaredoug.com/blog/2025/01/21/llm-judge-decision-...

Cool approach, thanks for sharing.

FWIW I also liked this (very different) post of yours: https://softwaredoug.com/blog/2026/07/13/who-want-to-be

Thank you!