Articles ยท Gemini AI Tools
Generate speech with Gemini and label it as generated audio
A text box that returns a voice is easy to treat as a microphone. It is not. When you press Generate speech, DevNestro sends the text to the configured TTS model (gemini-3.1-flash-tts-preview by default) and asks for an AUDIO response with a prebuilt voice such as Kore or Puck. The clip is generated. DevNestro will not call it a human recording, an official narrator, or a substitute for a licensed voice actor. The page shows a generated-speech note on every result.
Why this is worth doing in the browser
This page is separate from /tools/text-to-speech, which still uses local Kokoro voices in the browser. Mixing those two stories is how a visitor pastes a confidential script and believes it never left the tab. Gemini TTS is capability-gated and uses the text rate limit. If the host key is empty or the TTS model is unavailable on that key, the page stays usable: you can write text, pick a voice, see the cloud disclosure, and the request fails closed instead of inventing audio.
How to use the tool
Paste text. Pick a voice and an optional delivery style. Confirm the text may be sent to Google Gemini. Generate speech. Play, read any model note, and download. Clear / Start Over drops the text and the clip. Nothing is stored as a library.
- Do not paste confidential scripts unless you accept Google processing.
- Treat the voice as generated media, not a real person.
- A refused request may be a safety block, a missing model, or a rate limit.
- Keep the local Text to Speech tool for files that must stay in the browser.
Privacy
The text is sent to Google Gemini. DevNestro does not keep a generated-speech library. This is not processed locally. Free-tier training disclosure follows the configured tier.
Want generated speech with an honest cloud label? Open Gemini Text to Speech, send a short phrase you are willing to share, and download only what you will attribute as AI.