Overview
TTS 1.5 Mini converts written text into natural-sounding speech in roughly 120 milliseconds, making it one of the fastest synthesis options available. Whether you need a voiceover draft, a product demo narration, or a spoken notification for an app, you paste the text, choose a voice, and get back a clean audio file in seconds. Available on Picasso IA, it covers 15 languages so multilingual projects no longer require separate recording sessions or different tools for each locale. The result is a workflow where you can iterate through multiple takes in the time it once took to prepare a single recording.
How It Works
- Paste up to 2,000 characters of text into the input field. You can include break tags for timed pauses, emotion markers like [happy] or [sad], and non-verbal sounds like [laugh] or [sigh] to shape the delivery.
- Select a voice from the preset list (Ashley, Dennis, Alex, and others) or enter a custom voice ID if you have a cloned voice saved.
- Choose your audio format: MP3, WAV, OGG Opus, or FLAC. Pick a sample rate from 8,000 Hz up to 48,000 Hz to match the technical spec of your project.
- Adjust the speaking rate if you need faster or slower delivery, and set the temperature to control how expressive or neutral the voice sounds.
- Turn text normalization on, off, or leave it on auto so numbers, dates, and abbreviations are read out naturally.
- Click generate. TTS 1.5 Mini processes the input and returns your audio file in around 120 milliseconds.
Frequently Asked Questions
Do I need programming skills or technical knowledge to use this?
No, just open TTS 1.5 Mini on Picasso IA, adjust the settings you want, and hit generate.
Is it free to try?
Yes, you can run TTS 1.5 Mini without any account setup or payment required to get started. Submit your text, pick a voice, and download the file.
How long does it take to get results?
The model targets around 120 milliseconds of latency from request to audio output. For most inputs, the file is ready almost as soon as you click generate.
What output formats are supported?
TTS 1.5 Mini exports audio in MP3, WAV, OGG Opus, and FLAC. You can also select from seven sample rate options, from 8,000 Hz to 48,000 Hz, to match the technical requirements of your platform.
Can I customize the voice or speaking style?
Yes. Pick from preset voice names or supply a custom cloned voice ID. The temperature parameter controls expressiveness: lower values give a consistent, neutral tone; higher values add more variation. The speaking rate slider lets you slow down or speed up the narration.
What languages does TTS 1.5 Mini support?
It supports 15 languages, so you can produce multilingual audio content from a single tool without switching between services.
Where can I use the audio files I download?
The output files are clean with no added watermarks, so you can drop them directly into video edits, podcasts, mobile apps, e-learning modules, or any project that needs spoken audio.