1. X
  2. Cleanlab
Log inSign up
Cleanlab
701 posts
Cleanlab profile banner
user avatar

Cleanlab

@CleanlabAI
Cleanlab makes AI agents reliable. Detect issues, fix root causes, and apply guardrails for safe, accurate performance.
San Francisco
cleanlab.ai
Joined October 2021
233
Following
2,434
Followers
RepliesRepliesMediaMedia
  • Pinned
    user avatar
    Cleanlab
    @CleanlabAI
    Nov 18, 2025
    🚀 New from Cleanlab: Expert Guidance AI agents running multi-step workflows can fail in tiny, trust-breaking ways. Expert Guidance lets teams fix these behaviors with simple human feedback, instantly. ✈️In one airline workflow: 76% → 90% after only 13 guidance entries.
    Image
    00:00
  • user avatar
    Cleanlab
    @CleanlabAI
    Jan 28
    We're thrilled to join forces with @joinHandshake, where we'll be able to scale our team's pioneering work to inflect change with the world's leading AI labs. Hear more from our CEO and Co-founder, @cgnorthcutt, to learn about our next chapter.
    user avatar
    Curtis G. Northcutt
    @cgnorthcutt
    Jan 28
    News: @joinHandshake acquires @CleanlabAI! This "ten-year old job marketplace" has quietly become a top human data lab for AI--building an AI research org, acquiring top AI talent, and advancing Cleanlab tech and research to lead data foundations for frontier AI. 1 of 4
    Image
    00:00
  • user avatar
    Cleanlab
    @CleanlabAI
    Dec 3, 2025
    We discovered how to cut the failure rate of any AI agent on Tau²-Bench, the #1 benchmark for customer service AI. Agents often fail in multi-turn, tool-use tasks due to a single bad LLM output (reasoning slip, hallucinated fact, misunderstanding, wrong tool call, etc). We
    Image
  • user avatar
    Cleanlab
    @CleanlabAI
    Nov 10, 2025
    The “Year of the Agent” just got pushed back. Out of 1,837 enterprise leaders, most are struggling with stack churn + reliability. ⚙️ 70% rebuild every 90 days 😬 Less than 35 % are happy with their infrastructure 🤖 Most “agents” still aren’t really acting yet
    Image
  • user avatar
    Cleanlab
    @CleanlabAI
    Oct 30, 2025
    🚧 Even the best AI models still hallucinate. OpenAI’s recent paper on Why Language Models Hallucinate shows why this problem persists, especially in domain-specific settings. For teams implementing guardrails, we put together a short walkthrough: youtu.be/i_6fjKgboFg?si…

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
Advertisement
Advertisement