Gemini‑2.5‑Pro‑TTS is Google’s highest‑quality text‑to‑speech model preview, designed for complex workflows like podcas…
Gemini‑2.5‑Pro‑TTS is Google’s highest‑quality text‑to‑speech model preview, designed for complex workflows like podcasts, audiobooks, and customer support; it delivers expressive, accent‑ and style‑controllable single‑ or multi‑speaker speech, supporting over 23 languages, and built for state‑of‑the‑art output with the most powerful model architecture. Notes: - Text and style prompt limited to 4,000 bytes each (8,000 bytes combined) - Max output duration: approximately 10 minutes - Multi-speaker requires SpeakerName: text format (example: Alice: Hi! Bob: Hello, must be on new lines) - The model auto-detects the input language. The Language setting is a hint to help choose the right voice/accent, the model may override it if the text is in a different language. This bot supports optional parameters for additional customization.
EmpirioLabs AI
web
free
Others in the same category, ranked by how often they are opened.