Guide
TL;DR: An evergreen guide to AI audio workflows—voice cloning, TTS, STT, stem splitting, noise reduction—with links to in-depth how‑tos and product pages.
Explore the full landscape of AI audio. Learn the fundamentals, compare approaches, and jump straight into hands‑on tools.
Create songs, rap, instrumentals and samples from prompts.
Create and use custom voices in 200+ languages.
Accurate transcripts with diarization and timestamps.
Split songs into a detailed multitrack — vocals, individual drums, bass frequencies, and more.
Clean up audio with deep learning denoising.
Separate multi‑speaker conversations with clarity.
Short, task-oriented lessons written against each studio as it actually is — the same words you will see next to the controls.
From your first prompt to a finished track: how to describe a style the studio understands, what each control actually changes, and how to fix a take you nearly like.
Manuscript to finished chapters — casting a narrator, keeping a long read consistent, and exporting masters that pass a retailer's spec check.
Turn a recording of several people into one clean track per speaker — how many voices to declare, when to let it decide, and what the labelled transcript is for.
Split, isolate, or exclude — choosing the right mode for what you actually need, why the source recording decides the result, and what to do with the parts.
From typed text to a performance you'd keep — directing delivery, cloning a voice with consent, and holding one character across a long read.
Turn a recording you already have into a different voice — what makes a conversion convincing, why the source recording decides most of it, and whose voice you may actually use.
Getting a clean recording out of a noisy one — what the quality setting really changes, how to hear when cleanup has gone too far, and the problems no amount of processing will fix.
Getting a transcript you can trust — standard against premium accuracy, labelling who spoke, and the export formats that fit where it's going.
Coming soon: objective comparisons and audio samples.