Inline preview
Generated images, video frames, and speech previews display directly in your terminal. Visual previews use the Kitty graphics protocol where supported, including images returned as JPEG or WebP.
Generate text, images, video, and audio. Evaluate typed questions. Composable commands, shell pipelines, and hundreds of models for people and agents.
Run the same prompt across multiple models in parallel. Compare outputs side by side to find the best result. Combine with -n to generate multiple per model.
Use AI SDK evaluation models to ask focused questions about the same input in one call. Get typed answers and probabilities your scripts can use directly.
Pipe text in as context, turn images into video, transcribe audio, or send typed judgments to jq. Compose AI with the commands you already use.
Access text, image, video, speech, transcription, and evaluation models from OpenAI, Anthropic, Google, Black Forest Labs, ByteDance, and more through Vercel AI Gateway.
Generate content and make structured decisions in scripts, CI pipelines, agent toolchains, or your terminal.
Generated images, video frames, and speech previews display directly in your terminal. Visual previews use the Kitty graphics protocol where supported, including images returned as JPEG or WebP.
Predictable behavior for scripts and agents. Selected records on stdout, generated artifacts in files or pipes, and JSON metadata for automation.
Models are fetched directly from the AI Gateway — no hardcoded lists to maintain. Use short names or full provider/model IDs.
No config files, no init command, no setup wizard. Set an API key environment variable and start generating. Defaults work out of the box.