GenAiHub

Last checked 8 October 2026 — responded normally.

Gemini-2.5-Flash-TTS

Gemini‑2.5‑Flash‑TTS is Google’s low‐latency text‑to‑speech model that converts text input into audio output, supportin…

Live
Open / InstallLast updated October 8, 2026

Description

Gemini‑2.5‑Flash‑TTS is Google’s low‐latency text‑to‑speech model that converts text input into audio output, supporting both single‑ and multi‑speaker voices with controllable style, accent, and expressive tone — ideal for applications like podcasts, audiobooks, and conversational voice systems. Notes: - Text and style prompt limited to 4,000 bytes each (8,000 bytes combined) - Max output duration: approximately 10 minutes - Multi-speaker requires SpeakerName: text format (example: Alice: Hi! Bob: Hello, must be on new lines) - The model auto-detects the input language. The Language setting is a hint to help choose the right voice/accent, the model may override it if the text is in a different language. This bot supports optional parameters for additional customization.

Community Metrics

Views
0
Avg Rating
N/A
Ratings
0
Likes
0
Comments
—

Author

EmpirioLabs AI

Platform

web

Pricing model

free

Categories

  • Productivity

Tags

  • poe
  • empiriolabs-ai
  • text

Capabilities

  • Text input
  • Text generation
  • By EmpirioLabs AI

Information

TypeAgent
SourcePoe
AddedJune 18, 2026

Comments

?

Others in the same category, ranked by how often they are opened.