1. X
  2. TurningPoint AI
Log inSign up
TurningPoint AI
15 posts
Image
user avatar
TurningPoint AI
@TurningPointAI
We are a compact and hardcore research team focused on harnessing the power of Multimodal Reasoning. #Google #UCLA #UMD #PennState
Earth
turningpoint-ai.com
Joined June 2024
29
Following
81
Followers
RepliesRepliesMediaMedia
  • user avatar
    TurningPoint AI
    @TurningPointAI
    Jul 10
    Many monitors are trained & evaluated on prompt-elicited hacking trajectories, where models are explicitly asked to exploit the reward signal. But the real test is whether they catch the training-time hacks that naturally emerge during RL training without hacking instructions.
    Image
  • user avatar
    TurningPoint AI
    @TurningPointAI
    Jun 4
    We’re excited to share our new work on CVPR 2026, Understanding Reward Hacking in Text-to-Image Reinforcement Learning. Reinforcement learning is becoming an increasingly important tool for post-training text-to-image generation models. But as we optimize these models with
    Image
    Made with AI
  • user avatar
    TurningPoint AI
    @TurningPointAI
    Mar 6, 2025
    Many thanks to @_akhaliq for sharing our work on the first multimodal Aha moment with a 2B non-SFT model. Join us on our journey into multimodal reasoning and stay tuned for more cool research at @TurningPointAI turningpoint-ai.com!
    user avatar
    AK
    @_akhaliq
    Mar 4, 2025
    VisualThinker-R1-Zero R1-Zero's Aha Moment on just a 2B non-SFT Model VisualThinker-R1-Zero is a replication of DeepSeek-R1-Zero in visual reasoning. Successfully observe the emergent “aha moment” and increased response length in visual reasoning on just a 2B non-SFT models
    Image
  • user avatar
    TurningPoint AI
    @TurningPointAI
    Feb 28, 2025
    🚀 We’re excited to share our latest work! Welcome to the first successful "aha moment" on multimodal reasoning. "Aha moment" is featured by improved response length & performance. It emerges during RL of an unaligned base model on multimodal tasks. Aha moment for language
    Image
    GIF
  • user avatar
    TurningPoint AI
    @TurningPointAI
    Jun 29, 2024
    We made Multimodal LLMs safe, but have they also become oversensitive? "Every time I try, it uses all tokens just refusing." - @artilectium "This isn’t safety. It's a nanny state." - @krishnanrohit Concerned AI safety has gone too far? you’re not alone! Explore MOSSBench by
    Image

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
Advertisement
Advertisement