Video generation and editing model for fast, conversational text- and image-to-video workflows
Video generation and editing model for fast, conversational text- and image-to-video workflows. Supports a context window of 131,072 tokens. Supports extended reasoning. Input priced at $1.5 per million tokens. Weights are not publicly released; access is via the provider API.
api
paid
No benchmark results have been added yet.