Log inSign up
Arize AI
1,846 posts
Arize AI profile banner
@arizeai

Arize AI

@arizeai
The AI engineering platform for teams shipping reliable AI agents and LLM applications. Also home to @ArizePhoenix.
San Francisco, CA
arize.com
Joined January 2020
162
Following
5,052
Followers
RepliesRepliesRepostsRepostsMediaMedia

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • @arizeai
    Arize AI
    @arizeai
    13h
    Did you know LLM-as-a-judge works for multi-modal as well as text? We've just released a hands-on guide with a video walkthrough that shows you how to build an evaluator for images.
    Image
    Evaluate Receipt Agents with an Image Judge - Arize AX Docs
    From arize.com
  • @arizeai
    Arize AI
    @arizeai
    Sep 1
    Token dashboards tell you what you spent. They don't tell you if quality moved. ICYMI from Cost Alongside Quality: we showed how Arize AX puts cost next to evals across coding agents and shipped apps, and how a managed agent investigates a spend spike. Resources below 🧵
    6
  • @arizeai
    Arize AI
    @arizeai
    Aug 27
    Cost Alongside Quality: Proving the ROI of Your Coding Agents and AI Apps
  • @arizeai
    Arize AI
    @arizeai
    Aug 27
    If your agent needs 85 MCP turns to answer a SQL-shaped question, the problem may be the interface you gave it. Retrieval is great when a coding agent needs to inspect a few traces. It gets clumsy when the task is really about counting, filtering, joining, or aggregating across
    Image
    7
  • @arizeai
    Arize AI
    @arizeai
    Aug 25
    Better models don’t fix every agent failure. As models get more capable, the engineering around them matters even more: context, prompts, evals, feedback loops, and the way you measure the agent’s behavior. We spoke with @stuart__sy from @OpenAI about what developers should
    Image
    00:00
    4
Image
REPLAY
@arizeai
Arize AI
@arizeai
Cost Alongside Quality: Proving the ROI of Your Coding Agents and AI Apps
Advertisement
Advertisement