How AI is applied across API Evangelist and APIs.io. Read my AI disclosure →
API Evangelist API Evangelist
Discovery
Learnings
Guidance
Toolbox
Alignment
API Evangelist LLC

Grading the Machine: Using LLM as a Judge to Monitor AI Agents in Production

calendar_today June 15, 2026 person State Farm Engineering domain state-farm

By Alex Ginglen Imagine you’re in charge of building a generative AI agent chatbot. You’ve done countless user testing sessions before deploying to production. You’ve even written custom regression tests using LLM as a Judge and BERTScore that runs in your CI/CD pipeline.

open_in_new Read original post