Media Summary: Every AI team wants to ship faster than ever — but no one wants to ship something that breaks in production. This video introduces a new series on testing AI Discover why selecting the right metrics is crucial for designing and developing high-performing, safe AI applications.  ...

How To Evaluate Agents Galileo - Detailed Analysis & Overview

Every AI team wants to ship faster than ever — but no one wants to ship something that breaks in production. This video introduces a new series on testing AI Discover why selecting the right metrics is crucial for designing and developing high-performing, safe AI applications.  ... Learn Eval Engineering in this free, 5-part, hands-on course presented by ​90% of AI

Photo Gallery

How to Evaluate Agents: Galileo’s Agentic Evaluations in Action
ElevenLabs Voice Agent Observability with Galileo | Multi-Turn Evaluation Tutorial
Evaluating Multi Agent Systems
Meet Galileo: The Evaluation, Observability & Guardrails Stack for AI
Taming Rogue AI Agents with Observability-Driven Evaluation — Jim Bennett, Galileo
How to evaluate agents in practice
AI Agent evaluation: A complete guide to measuring performance
AI Agent Evaluation | Pratik Bhavsar, Galileo
The agent evaluation revolution
How to Choose the Right AI Evaluation Metrics (with Galileo)
Evaluating and Debugging Non-Deterministic AI Agents
Evaluation Agents: Exploring the Next Frontier of GenAI Evals
View Detailed Profile
How to Evaluate Agents: Galileo’s Agentic Evaluations in Action

How to Evaluate Agents: Galileo’s Agentic Evaluations in Action

Evaluating

ElevenLabs Voice Agent Observability with Galileo | Multi-Turn Evaluation Tutorial

ElevenLabs Voice Agent Observability with Galileo | Multi-Turn Evaluation Tutorial

See how to monitor and

Evaluating Multi Agent Systems

Evaluating Multi Agent Systems

Multi-

Meet Galileo: The Evaluation, Observability & Guardrails Stack for AI

Meet Galileo: The Evaluation, Observability & Guardrails Stack for AI

Every AI team wants to ship faster than ever — but no one wants to ship something that breaks in production.

Taming Rogue AI Agents with Observability-Driven Evaluation — Jim Bennett, Galileo

Taming Rogue AI Agents with Observability-Driven Evaluation — Jim Bennett, Galileo

LLM

How to evaluate agents in practice

How to evaluate agents in practice

Evaluating Agents

AI Agent evaluation: A complete guide to measuring performance

AI Agent evaluation: A complete guide to measuring performance

Evaluating

AI Agent Evaluation | Pratik Bhavsar, Galileo

AI Agent Evaluation | Pratik Bhavsar, Galileo

Pratik Bhavsar, from

The agent evaluation revolution

The agent evaluation revolution

This video introduces a new series on testing AI

How to Choose the Right AI Evaluation Metrics (with Galileo)

How to Choose the Right AI Evaluation Metrics (with Galileo)

Discover why selecting the right metrics is crucial for designing and developing high-performing, safe AI applications. @erinmikail ...

Evaluating and Debugging Non-Deterministic AI Agents

Evaluating and Debugging Non-Deterministic AI Agents

Evaluate

Evaluation Agents: Exploring the Next Frontier of GenAI Evals

Evaluation Agents: Exploring the Next Frontier of GenAI Evals

From simple RAG to advanced

Observability in AI apps. Eval Engineering for AI Developers, lesson 2 - add observability to AI

Observability in AI apps. Eval Engineering for AI Developers, lesson 2 - add observability to AI

Learn Eval Engineering in this free, 5-part, hands-on course presented by @jimbobbennett ​90% of AI