Xiaomi's native omni-modal model (text / image / video / audio understanding) with Pro-level agentic performance at ~ha…
Xiaomi's native omni-modal model (text / image / video / audio understanding) with Pro-level agentic performance at ~half the inference cost and a 1M-token context — built for cost-efficient, perception-rich agent workflows. File Support: Text, Markdown, Image, Video and PDF files Context window: 1M tokens
Novita AI
web
subscription
Others in the same category, ranked by how often they are opened.