aboutsummaryrefslogtreecommitdiff
diff options
context:
space:
mode:
authorhistoria <historiavg@proton.me>2026-08-16 20:34:05 -0400
committerhistoria <historiavg@proton.me>2026-08-16 20:34:05 -0400
commit2bb8f721092fe2f4d6ae6e425181094c1f0da57a (patch)
tree49de85f0b8b8197a8eb73f989a2bf26109044635
parent12caa5275c5d904c1451ea9b6ef9e7ad241d969a (diff)
downloadtts-audiobook-generator-2bb8f721092fe2f4d6ae6e425181094c1f0da57a.tar.gz
readme
-rw-r--r--README.md2
1 files changed, 1 insertions, 1 deletions
diff --git a/README.md b/README.md
index f03aa07..ecf4d04 100644
--- a/README.md
+++ b/README.md
@@ -91,7 +91,7 @@ python audiobook_converter.py --voice-clone --voice-sample path/to/reference.wav
The reference .wav should be ~10-15 seconds with a minimum of 3 seconds and maximum of 60 seconds. Longer is not better. ~15 seconds is ideal.
-Omit `--voice-sample-text` and Whisper will be used automatically to transcribe the reference audio (`faster_whisper` or `whisper`). If no Whisper backend is installed, it falls back to x-vector-only cloning.
+Whisper will be used automatically to transcribe the reference audio (`faster_whisper` or `whisper`). If no Whisper backend is installed, it falls back to x-vector-only cloning.
To skip automatic transcription explicitly, pass `--no-transcription`. This should be worse, but in my experience may give a preferable flatter tone to certain voices.