Workflow directory

Private AI audio tools for the exact job you need to do

The tools directory gathers cleanup utilities, speech-to-text, subtitle, creator voice, and Ebook & Document Audio workflows in one place. Use it when you know the task you need, not just the model you want.

15 focused tools 4 workflow families Free to use Β· no signup or API key

01 Β· Prepare

Tools

Clean text, fix punctuation, strip subtitle noise, and prepare scripts before generation or export.

4 tools

Sponsored

Ads help keep OfflineTTS free to use.

02 Β· Transcribe

Transcription & Subtitles

Private browser workflows for speech to text, captions, and subtitle exports.

5 tools

03 Β· Create

Creator Voice Tools

Voice-over workflows for creators who want browser-based narration and repurposing assets.

3 tools

04 Β· Listen

Ebook & Document Audio

Turn EPUB, PDF, and TXT reading material into speech for accessibility, review, and long-form listening.

3 tools

Choose the Tool by the Output You Need

These pages are not interchangeable keyword versions of one generator. Each one opens a different component or a focused view of a shared engine. Start from the artifact you need to leave with, then check the limits before moving sensitive or production material into the workflow.

Needed outputStart hereWhat it changesReview before use
Plain text without web or formatting noiseText CleanerRemoves selected URLs, email addresses, HTML, emoji, invisible characters, or extra spacingImportant symbols and contact details may be removed
A narration draft with more deliberate punctuationPunctuation EnhancerApplies deterministic punctuation and capitalization rulesRules can misread names, dialogue, code, or specialist text
Spoken lines extracted from SRT or VTTSubtitle CleanerStrips selected cue notation, labels, sounds, tags, and duplicatesSRT/VTT output uses placeholder timing and needs retiming
A structured creator or audiobook scriptScript FormatterAdds line breaks, planning markers, pauses, and an approximate durationLabels are not speaker detection and may be read aloud
TXT transcript or SRT/VTT captionsAudio to TextRuns Whisper transcription with segment or optional word timingNames, numbers, speakers, and timestamps need human review
Narration audio for a video timelineYouTube Voice GeneratorTurns a reviewed script into downloadable speechIt does not create video, clone a person, publish, or guarantee monetization
A resumable book with paragraph audioEbook to AudioParses supported files, stores a local bookshelf, and generates selected chaptersScanned PDFs need OCR; storage and phonemization boundaries apply

What β€œPrivate Browser Workflow” Means

Text Cleaner, Punctuation Enhancer, Subtitle Cleaner, and Script Formatter apply their rules in page JavaScript. Audio to Text runs Whisper inference on the selected media in the browser. TTS engines synthesize audio locally after their assets load. This avoids an upload-first account dashboard, but it does not mean the whole website is disconnected from the network.

The site, analytics, and model hosts receive ordinary web requests. Non-English Kokoro sends entered text to the documented phonemization service before local synthesis. First-time model use needs a download, and browser storage may retain models, preferences, ebook text, or generated paragraph audio until cleared. The Privacy Policy explains these paths and deletion options.

A Safer Production Order

Keep an unchanged source copy. Work on a short representative sample, select only the cleanup rules you need, compare the output, and then test the final TTS or STT engine. Check names, numbers, quotations, meaningful sound labels, and rights-sensitive material before exporting. Local processing does not make an inaccurate transcript correct or grant permission to reuse someone else's book, captions, recording, or voice identity.

For production delivery, listen to the complete audio or compare the full transcript with the source. Save the reviewed text alongside the exported file so later corrections are possible. Clear site data when local retention is inappropriate, remembering that exported files must be deleted separately from the operating system. Report reproducible tool failures through the Contact page with private content removed.

Need the raw engine workspace instead?

Go straight to the TTS or Whisper workspace when you already know which engine-driven tool you want.