Fast Gemini workhorse for multimodal apps where latency and price matter
Fast Gemini workhorse for multimodal apps where latency and price matter. Supports a context window of 1,048,576 tokens. Supports extended reasoning. Supports tool calling. Input priced at $0.3 per million tokens. Weights are not publicly released; access is via the provider API.
api
paid