
Lillian Zhao
How Cresta Grades Its Agents: Evaluators in an LLM World
Learn how Cresta evaluates AI Agents at scale using LLM judges, deterministic checks, simulations, and human calibration to catch failures before launch and build trust in high-stakes environments.
The Data Comes First: Mining Real Conversations for Test Coverage
Learn how Cresta mines historical data, synthetic customers, knowledge-base Q&A pairs, and post-launch feedback to build test coverage that reflects how customers actually behave.
How Cresta Grades Its Agents: Evaluators in an LLM World
Learn how Cresta evaluates AI Agents at scale using LLM judges, deterministic checks, simulations, and human calibration to catch failures before launch and build trust in high-stakes environments.
The Data Comes First: Mining Real Conversations for Test Coverage
Learn how Cresta mines historical data, synthetic customers, knowledge-base Q&A pairs, and post-launch feedback to build test coverage that reflects how customers actually behave.