All resources

Everything we’ve published — page 3.

Guide

Arize alternatives: How the top AI observability and evaluation tools compare

Compare Arize alternatives including LangSmith, Langfuse, Braintrust, Helicone, and Fiddler on evals, tracing, self-hosting, and production depth. See…

Read the guide
Blog

Where agent evals are going: Agent-as-a-Judge

Agents changed what failure looks like, and the evaluation layer has to change with them. Why agent-as-a-judge is…

Read the post
Customer Story

How TheFork uses agent evals and tracing to improve AI in production

TheFork’s AI team on catching regressions early, finding hidden latency, and building the feedback loops needed to improve…

Read the story
Case Studies

How Tripadvisor is building the AI product development lifecycle for agentic travel

Tripadvisor VP of Data and AI Rahul Todkar on building a production AI lifecycle for traditional ML and…

Read the story
Case Studies

How Booking.com scales AI observability with Arize

How Booking.com built a unified AI observability stack with Arize for agentic GenAI workflows and traditional ML —…

Read the story
Case Studies

How LG Uplus is building better AI customer service agents with evaluation-driven development

How LG Uplus uses Arize AX to build evaluation-driven AI contact center agents — combining production traces, user…

Read the story
Customer Story

How Handshake deployed and scaled 15+ LLM use cases in under 6 months with Arize AX

See how Handshake scaled 15+ production LLM use cases in six months with Arize AX for tracing, evals,…

Read the story
Guide

How to build agent evals from traces

Evals are tests for AI; traces are logs for AI. This tutorial shows how to read agent traces,…

Read the guide
Blog

How Uber evaluates AI agents at production scale

A background comment about pizza exposed a failure that Uber’s offline evaluations had missed. The incident helped reveal…

Read the post
Guide

AI model lifecycle management: 7 stages, controls, and tools

The seven stages of AI model lifecycle management, what to version at each gate, which tools own which…

Read the guide
Blog

Arize and Dynatrace: Making the World’s AI Work

Today we are announcing the signing of a definitive agreement for the acquisition of Arize by Dynatrace to…

Read the post
Post

AI agent guardrails vs. evals: How to build more reliable agent systems

Guardrails constrain what an agent can do in code; evals judge whether it performed well. Learn how both…

Read the post

Don’t ship vibes.

Arize gives AI teams observability and evals to understand and improve agent performance.