Hey Phoenix community 👋
We’re hosting a webinar on Benchmarking AI agents & tool use with Harbor and Arize Phoenix
Elizabeth H. from the Phoenix team will join us for the webinar and walk through an end-to-end agent benchmark with Harbor, then analyze the results in Phoenix.
You’ll learn how to:
• compare configurations against the same tasks before you ship
• turn a real agent workflow into a repeatable test with a measurable final state
• run the same task set across different agent, model, and tool configurations
• compare scores, infrastructure failures, and traces in Phoenix to diagnose regressions and prioritize improvements
Oct 8: 11am PT / 2pm ET
Format: 30 min content + 15 min Q&A
Register here: luma.com/arizeai-benchmarking-ai-agents-with-phoenix?…