Voice generation crossed the threshold where output stops sounding synthetic, and that changed both what the tools are useful for and what questions they raise.
This category covers the practical side first. Which tools sound natural across a long read rather than a demo sentence, how well voice cloning holds up from a short sample, multilingual output and accent handling, pronunciation control for names and technical terms, and where the per-character pricing becomes the limiting factor.
It also covers the parts that are less comfortable. Consent and licensing around cloned voices differ by tool and by jurisdiction, and the terms are worth reading before you build anything on top of them. Members here have run into that in real projects.
Music generation, sound design and audio cleanup tools fit here too. If your question is about turning recorded speech into text rather than generating it, Transcription is the better home.
Popular Discussions
The most-engaged audio & voice discussions in the WhatAI community.
I have a repetitive strain injury that makes extended typing painful. I have managed it for years through ergonomic equipment and pacing but some tasks are harder to pace than others. Documentation wr...
By paigeking
May 12, 2026
Audio & Voice
✓ Reviewed for community standards
❤ 1
👁 8
💬 6
🚩 Report
I am not a musician. I cannot play an instrument, I do not know music theory and I have never used a DAW. I have had song ideas floating around in my head for years with nowhere to put them. Suno AI c...
By nora_d
Apr 19, 2026
Audio & Voice
✓ Reviewed for community standards
❤ 2
👁 7
💬 5
🚩 Report
I produce music as a hobby and I occasionally need to work with stems from existing recordings, isolate a vocal, pull out a bassline to sample, remove the drums from a reference track to analyze the a...
By riley.parke
Apr 11, 2026
Audio & Voice
✓ Reviewed for community standards
❤ 1
👁 7
💬 5
🚩 Report
I create content for a brand that operates across several markets including Brazil, Japan and Germany. We have been exploring AI-generated music for branded content and social media and I want to writ...
By zachmason
May 16, 2026
Audio & Voice
✓ Reviewed for community standards
❤ 1
👁 8
💬 4
🚩 Report
Most people know ElevenLabs as the voice cloning and text-to-speech tool. That is still there and still excellent. But they have added image and video generation to the platform and I do not think eno...
By henry_allen
May 17, 2026
Audio & Voice
✓ Reviewed for community standards
❤ 1
👁 6
💬 4
🚩 Report
Recent Discussions
The latest audio & voice conversations from the community.
The 2026 content creator guide https://www.youtube.com/watch?v=XcwKeRWqV64 covers ElevenLabs in a way that makes clear how much the platform has expanded beyond its original voice generation positioni...
By finn108
Jul 7, 2026
Audio & Voice
✓ Reviewed for community standards
❤ 0
👁 8
💬 2
🚩 Report
We built WhatAI Music, a YouTube channel of AI-assisted soundtracks for focus, study, and deep work. It went live this week, and rather than pretend everyone will love it, we want to have the actual a...
By riley76
Jul 6, 2026
Audio & Voice
⭐ Trusted Member
❤ 1
👁 3
💬 0
🚩 Report
The LALAL.AI Andromeda network overview https://www.youtube.com/watch?v=4tupyfWGC2s covers the sixth-generation AI engine and what specifically the generation claim means for separation quality. Train...
By dan.cook
Jul 4, 2026
Audio & Voice
✓ Reviewed for community standards
❤ 1
👁 8
💬 2
🚩 Report
The LALAL.AI Voice Cloner overview https://www.youtube.com/watch?v=Lpy1pbXbBc0 covers the core capabilities: generating natural-sounding voice clones from uploaded audio recordings, supporting all lan...
By abby_hall
Jul 3, 2026
Audio & Voice
✓ Reviewed for community standards
❤ 1
👁 5
💬 0
🚩 Report
The Suno beginner tutorial https://www.youtube.com/watch?v=72R1NjNaUnE covers the organisational and creation workflow in a combination that makes the first session productive rather than exploratory....
By isaac59
Jul 3, 2026
Audio & Voice
✓ Reviewed for community standards
❤ 0
👁 4
💬 2
🚩 Report