@spenserskates
Today we're launching Agent Analytics Every team shipping an AI agent has the same blind spot. Offline evals pass, you ship, and then you have no idea what's happening in production. AI fails silently. Users ask a question and get different answers. They all look 'engaged' in a classic dashboard. You don't know who got a great response and who got a terrible one. Agent Analytics solves it: - Every session scored out of the box on task completion, response quality, friction, safety, and negative feedback - Topic clustering across thousands of conversations, so you know if a failure hits 1 user or 10,000 - Eval agents that watch for regressions, and if you want will file a Linear ticket or the pull request themselves - Agent quality sits next to product data, so 'payment scheduling fails 31%' becomes 'which renewals did that cost us?' The Economist got their agent to a 96.9% task success rate and cut weekly failures 84%. Included on every plan. Free tier included. https://t.co/K5vqrlMLvd