Escrito por

Lillian Zhao

Responsable principal de producto
No se han encontrado resultados.

How Cresta Grades Its Agents: Evaluators in an LLM World

Learn how Cresta evaluates AI Agents at scale using LLM judges, deterministic checks, simulations, and human calibration to catch failures before launch and build trust in high-stakes environments.

Más información
Ingeniería
All
No se han encontrado resultados.

The Data Comes First: Mining Real Conversations for Test Coverage

Learn how Cresta mines historical data, synthetic customers, knowledge-base Q&A pairs, and post-launch feedback to build test coverage that reflects how customers actually behave.

Más información
Ingeniería
All
No se han encontrado resultados.

Why AI Agent Evaluations Fail — and How the Swiss-Cheese Model Prevails

Learn about Cresta's forward-deployed team and how they approach building and iterating on AI agents.

Más información
Ingeniería
All
Este autor no tiene guías
No se han encontrado resultados.

How Cresta Grades Its Agents: Evaluators in an LLM World

Learn how Cresta evaluates AI Agents at scale using LLM judges, deterministic checks, simulations, and human calibration to catch failures before launch and build trust in high-stakes environments.

Más información
Ingeniería
All
No se han encontrado resultados.

The Data Comes First: Mining Real Conversations for Test Coverage

Learn how Cresta mines historical data, synthetic customers, knowledge-base Q&A pairs, and post-launch feedback to build test coverage that reflects how customers actually behave.

Más información
Ingeniería
All
No se han encontrado resultados.

Why AI Agent Evaluations Fail — and How the Swiss-Cheese Model Prevails

Learn about Cresta's forward-deployed team and how they approach building and iterating on AI agents.

Más información
Ingeniería
All
Este autor no tiene guías