After Terminal-Bench, we’re excited to introduce Frontier Bench — a new coding benchmark built from real-world software engineering tasks contributed by the community.
One thing has become strikingly clear: it’s getting really hard to find coding tasks that frontier AI agents
- Presenting UniT at CVPR!🚀 Introducing our fresh work at Stanford and Meta MSL: UniT — Unified Multimodal Chain-of-Thought Test-time Scaling What if a single model could generate an image, look at it, think about what's wrong, and fix it — all by itself? That's exactly what UniT does. 🧵👇
- J2A brought together researchers, founders, creators, and industry practitioners working across cinematic video generation, world models, controllable creation, editing, VFX, post-production, and creative evaluation. Huge thanks to our speakers from @runwayml, @GoogleDeepMind,The first @Journey2Awards workshop at #CVPR2026 is officially wrapped 🎬 When we proposed @Journey2Awards, the question was simple: AI video is moving fast — but what does it take to move from impressive clips to movie-grade production?
- The first @Journey2Awards workshop at #CVPR2026 is officially wrapped 🎬 When we proposed @Journey2Awards, the question was simple: AI video is moving fast — but what does it take to move from impressive clips to movie-grade production?
- so excited that 1st @Journey2Awards workshop is happening tomorrow at @CVPR! We will explore how generative AI and multimodal LLMs can support the full film production pipeline. J2A emphasizes on studio-caliber cinematic video generation, and the integration of generative AI

