Back to archive
By Ivan Vydrin12 min read25 August 2026< 50 views

Evaluating AI Agents on Databricks

When the correct output is a set rather than a value, you cannot assert it. Code scorers, LLM judges, and where each one belongs in an evaluation.