-
Prompt EngineeringHow to ship a local LLM that matches frontier LLMs with evals and prompt engineering
Most production AI features don't need a frontier model. Here's how capability evals and prompt engineering can help ship a local SLM that matches frontier-model quality with lower… RL Nabors May 26, 2026 15 min read -
Prompt EngineeringPrompt templates as configs, not code
This post was written in April 2026. Cloud products, feature maturity, and recommended patterns change over time, so readers should treat these examples as directional guidance. For teams… Dat Ngo April 30, 2026 21 min read -
Prompt EngineeringTop 5 AI Prompt Management Tools for 2026
Every new AI release rises or falls on how people experience it, and prompts play a major role in shaping that experience. A short sentence can write code,… Aryan Kargwal January 1, 2026 16 min read -
Prompt EngineeringCLAUDE.md Best Practices for Claude Code
In our last post on Prompt Learning (our prompt optimization feature), we optimized Cline, a powerful coding agent, through its system prompt. This time, we used it on… Priyan Jindal November 20, 2025 9 min read -
Prompt EngineeringGEPA vs Prompt Learning: Benchmarking Different Prompt Optimization Approaches
In June 2025, Andrej Karpathy introduced Software 3.0: the notion that software development is shifting from programming through code to prompting through natural language. When building programs, the… Priyan Jindal November 17, 2025 11 min read -
Prompt Engineering8 Top Prompt Testing & Optimization Tools (2026)
The best prompt testing and optimization tools for LLMs and multi-agent systems in 2026, compared including features, evals, and how to choose. If we were to give the… Trent Fowler October 28, 2025 18 min read -
Prompt EngineeringOptimizing Coding Agent Rules (./clinerules) for Improved Accuracy
Coding agents have become the focal point of modern software development. Tools like Cursor, Claude Code, Codex, Cline, Windsurf, Devin, and many more are revolutionalizing how engineers write… Priyan Jindal October 14, 2025 10 min read -
Prompt EngineeringEvidence-Based Prompting Strategies for LLM-as-a-Judge: Explanations and Chain-of-Thought
When LLMs are used as evaluators, two design choices often determine the quality and usefulness of their judgments: whether to require explanations for decisions, and whether to use… Sri Chavali Elizabeth Hutton Aparna Dhinakaran August 20, 2025 8 min read -
Prompt EngineeringNew In Arize AX: Prompt Learning, Arize Tracing Assistant, and Multiagent Visualization
July was a big month for Arize AX, with updates to make AI and agent engineering much easier. From prompt learning to new skills for Alyx and OpenInference… Sanjana Yeddula August 7, 2025 5 min read -
Prompt EngineeringPrompt Learning: Using English Feedback to Optimize LLM Systems
Applications of reinforcement learning (RL) in AI model building has been a growing topic over the past few months. From Deepseek models incorporating RL mechanics into their training… Jason Lopatecki Aparna Dhinakaran Priyan Jindal Aman Khan July 18, 2025 15 min read -
Prompt EngineeringArize Observe 2025 – Product Releases
Arize Observe 2025 brought a wealth of new product releases, including a redesigned copilot, agent eval options, and state-of-the-art prompt optimization techniques. Check them all out below! Copilot… John Gilhuly June 25, 2025 7 min read -
Prompt EngineeringNew in Arize: Realtime Trace Ingestion, Prompt Playground Upgrades & More
In May, we expanded access to realtime trace ingestion across all Arize AX tiers, making it easier than ever to monitor LLM performance live. We also rolled out… Sally-Ann DeLucia June 4, 2025 2 min read -
Prompt EngineeringNew in Arize: Bigger Datasets, Better Evaluations, and Expanded CV Support
April was a big month for Arize, with updates designed to make building, evaluating, and managing your models and prompts even easier. From larger dataset runs in Prompt… Sally-Ann DeLucia April 28, 2025 2 min read -
Prompt EngineeringPrompt Optimization Techniques
LLMs are powerful tools, but their performance is heavily influenced by how prompts are structured. The difference between an effective and ineffective prompt can determine whether a model… Sri Chavali March 17, 2025 9 min read -
Prompt EngineeringPrompt Management from First Principles
How we built a holistic prompt management system that preserves developer freedom Unlike traditional software, where code execution follows predictable paths, LLM applications are inherently non-deterministic. Their behavior… Xander Song Mikyo King March 7, 2025 5 min read
Don’t ship vibes.
Arize gives AI teams observability and evals to understand and improve agent performance.