Efficient model for low-latency assistance, extraction, and routine automation
Efficient model for low-latency assistance, extraction, and routine automation. Supports a context window of 32,768 tokens. Supports tool calling. Input priced at $0.15 per million tokens. Weights are not publicly released; access is via the provider API.
Upstage
api
paid
No benchmark results have been added yet.