Skip to content

Workflow directory

Private AI audio tools for the exact job you need to do

The tools directory gathers cleanup utilities, speech-to-text, subtitle, creator voice, audio effects and editing, measurement and conversion, and Ebook & Document Audio workflows in one place. Use it when you know the task you need, not just the model you want.

35 focused tools 6 workflow families Free to use Β· no signup or API key

01 Β· Prepare

Tools

Clean text, fix punctuation, strip subtitle noise, and prepare scripts before generation or export.

4 tools

02 Β· Transcribe

Transcription & Subtitles

Private browser workflows for speech to text, captions, and subtitle exports.

5 tools

03 Β· Create

Creator Voice Tools

Voice-over workflows for creators who want browser-based narration and repurposing assets.

3 tools

04 Β· Polish & Edit

Voiceover Polish & Audio Editing

Shape voice tone, boost levels, reduce background noise, trim silence, and splice takes together β€” rendered offline in the browser.

16 tools
πŸ”Š

Volume Booster

Make a quiet recording louder and normalise inconsistent levels

🎼

Pitch Shifter

Transpose a recording up or down while keeping its original length

⏩

Tempo & Speed Changer

Slow down or speed up audio while holding the original pitch

πŸ›οΈ

Add Reverb to Audio

Place a dry recording in a room, hall, plate, or cathedral

🎚️

Bass Booster

Add low-end weight with a shelf and a mud cut, not a blunt volume bump

πŸŽ›οΈ

Audio Equalizer

Ten bands from 32 Hz to 16 kHz, with presets and a spectrum view

βͺ

Reverse Audio

Flip a recording backwards, with an optional pre-reverse reverb swell

βœ‚οΈ

Audio Trimmer

Drag a selection on the waveform, add fades, and export just that part

πŸ”—

Audio Joiner

Merge files in order, with optional normalization, crossfades, and gaps

🎧

8D Audio

Rotate a track around the listener with stereo panning and space

⚑

Nightcore Maker

Speed a track up with a matching pitch lift β€” the classic nightcore sound

⏫

Sped Up Song Maker

Raise the tempo and hold the pitch β€” fast but not squeaky

πŸŒ™

Slowed + Reverb Maker

Slow a track down and wash it in a long reverb tail

🌴

Vaporwave Maker

Slow a track to roughly 65% with a hazy, softened top end

🎭

Voice Changer

Sixteen character presets, each one still fully editable

🧽

Noise Remover

Learn the noise floor of a recording, then gate it out band by band

05 Β· Analyze & Convert

Audio Lab & Analysis Tools

Detect tempo, musical key, and loudness, convert formats, trim silence, and generate calibration test signals.

4 tools

06 Β· Listen

Ebook & Document Audio

Turn EPUB, PDF, and TXT reading material into speech for accessibility, review, and long-form listening.

3 tools

Choose the Tool by the Output You Need

These pages are not interchangeable keyword versions of one generator. Each one opens a different component or a focused view of a shared engine. Start from the artifact you need to leave with, then check the limits before moving sensitive or production material into the workflow.

Needed outputStart hereWhat it changesReview before use
Plain text without web or formatting noiseText CleanerRemoves selected URLs, email addresses, HTML, emoji, invisible characters, or extra spacingImportant symbols and contact details may be removed
A narration draft with more deliberate punctuationPunctuation EnhancerApplies deterministic punctuation and capitalization rulesRules can misread names, dialogue, code, or specialist text
Spoken lines extracted from SRT or VTTSubtitle CleanerStrips selected cue notation, labels, sounds, tags, and duplicatesSRT/VTT output uses placeholder timing and needs retiming
A structured creator or audiobook scriptScript FormatterAdds line breaks, planning markers, pauses, and an approximate durationLabels are not speaker detection and may be read aloud
TXT transcript or SRT/VTT captionsAudio to TextRuns Whisper transcription with segment or optional word timingNames, numbers, speakers, and timestamps need human review
A louder, level-consistent copy of a recordingVolume BoosterMeasures peak and RMS, applies gain, and limits the outputPeak normalization is not LUFS loudness compliance
A cleaned recording with the background noise reducedNoise RemoverLearns the noise floor from the quietest window and gates below itSteady noise only β€” clicks and overlapping speech survive
A tempo or key change on a finished trackTempo ChangerPhase-vocoder stretch that preserves pitch, or shifts it on requestPhase-vocoder artefacts appear on dense percussion
One file cut into the section you needAudio TrimmerWaveform selection with sample-accurate start and end pointsCuts need a short fade or they click
The tempo, key, and level of an unknown fileBPM & Key FinderOnset-histogram tempo, chroma key estimation, and loudness factsTempo and key are estimates and can lock onto a subdivision
A file at a different sample rate or channel countAudio ConverterResamples, folds to mono or widens to stereo, then re-encodesLinear resampling is not mastering grade; exports are WAV or MP3 only
Narration audio for a video timelineYouTube Voice GeneratorTurns a reviewed script into downloadable speechIt does not create video, clone a person, publish, or guarantee monetization
A resumable book with paragraph audioEbook to AudioParses supported files, stores a local bookshelf, and generates selected chaptersScanned PDFs need OCR; storage and phonemization boundaries apply

What β€œPrivate Browser Workflow” Means

Text Cleaner, Punctuation Enhancer, Subtitle Cleaner, and Script Formatter apply their rules in page JavaScript. Audio to Text runs Whisper inference on the selected media in the browser. TTS engines synthesize audio locally after their assets load. This avoids an upload-first account dashboard, but it does not mean the whole website is disconnected from the network.

The site, analytics, and model hosts receive ordinary web requests. Non-English Kokoro sends entered text to the documented phonemization service before local synthesis. First-time model use needs a download, and browser storage may retain models, preferences, ebook text, or generated paragraph audio until cleared. The Privacy Policy explains these paths and deletion options.

A Safer Production Order

Keep an unchanged source copy. Work on a short representative sample, select only the cleanup rules you need, compare the output, and then test the final TTS or STT engine. Check names, numbers, quotations, meaningful sound labels, and rights-sensitive material before exporting. Local processing does not make an inaccurate transcript correct or grant permission to reuse someone else's book, captions, recording, or voice identity.

For production delivery, listen to the complete audio or compare the full transcript with the source. Save the reviewed text alongside the exported file so later corrections are possible. Clear site data when local retention is inappropriate, remembering that exported files must be deleted separately from the operating system. Report reproducible tool failures through the Contact page with private content removed.

Need the raw engine workspace instead?

Go straight to the TTS or Whisper workspace when you already know which engine-driven tool you want.