TÀI LIỆU

Hướng Dẫn Sử Dụng

Tất cả mọi thứ bạn cần biết để chuyển văn bản thành giọng nói chuyên nghiệp với Omnivox.cc

🚀

Bắt đầu nhanh

Tạo audio đầu tiên chỉ trong 3 bước đơn giản

Bước 01
📝

Nhập văn bản

Mở Studio và nhập nội dung muốn chuyển thành giọng nói

Bước 02
🎤

Chọn giọng đọc

Chọn ngôn ngữ, giọng từ thư viện hoặc dùng giọng AI mặc định

Bước 03
⬇️

Tạo và tải về

Nhấn Tạo giọng, nghe preview rồi tải file WAV về máy

💡
Đăng nhập bằng GoogleOmnivox.cc yêu cầu đăng nhập Google. Tài khoản mới nhận 1 credit miễn phí. Nâng cấp gói để nhận thêm credits và mở khóa tính năng nâng cao.
👤

Tài khoản & Credits

Hệ thống credits, các gói dịch vụ và giới hạn sử dụng

1

Đăng nhập bằng Google

Nhấn Đăng nhập với Google. Tài khoản tạo tự động — không cần điền thông tin.

2

Credits là gì?

Mỗi lần tạo audio tiêu tốn 1 credit. Tài khoản mới có 1 credit miễn phí. Nâng cấp gói để nhận thêm.

3

Chọn gói phù hợp

Xem chi tiết tại Bảng giá. Thanh toán qua chuyển khoản (Sepay) hoặc thẻ quốc tế (LemonSqueezy).

GóiCreditsTối đa ký tự/lầnLưu audio
Free1 credit1.000 ký tự72 giờ
Starter200 credits5.000 ký tự7 ngày
Pro800 credits15.000 ký tự30 ngày
Business3.000 credits50.000 ký tự90 ngày
Voice Clone & Voice Design từ gói ProTính năng nhân bản và thiết kế giọng yêu cầu gói Pro trở lên.
🎙️

Studio TTS

Chuyển văn bản thành giọng nói với nhiều tùy chọn ngôn ngữ và giọng đọc

1

Nhập văn bản

Gõ hoặc dán nội dung. Giới hạn: Free 1.000 · Starter 5.000 · Pro 15.000 · Business 50.000 ký tự. Văn bản dài tự chia segment và ghép thành 1 file liên tục.

2

Chọn ngôn ngữ

Chọn Tiếng Việt, English, 中文 hoặc để Tự động — AI tự nhận diện ngôn ngữ.

3

Chọn giọng đọc

Chọn giọng từ thư viện. Giọng đã clone hoặc thiết kế cũng xuất hiện ở đây.

4

Tùy chỉnh thông số (nâng cao)

Chỉnh Tốc độ, Số bướcGuidance Scale. Xem mục Thông số nâng cao.

5

Tạo và tải về

Nhấn Tạo giọng. Mất 5–30 giây. Sau đó nhấn ▶ Phát để nghe hoặc ↓ Tải về để lưu WAV.

Mẹo: Văn bản dàiHệ thống tự chia theo câu và ghép nối liền mạch. Không cần chia thủ công.
🔊

Nhân bản giọng nói

Tạo bản sao giọng nói từ một đoạn audio mẫu

⚠️
Yêu cầu audio mẫuRõ ràng, không tiếng ồn, không nhạc nền. Lý tưởng: 15–60 giây. Định dạng: WAV, MP3, M4A, FLAC.
1

Chuẩn bị audio mẫu

File có giọng rõ, nền im lặng. Giọng nói liên tục (đọc đoạn văn). Chất lượng audio ảnh hưởng trực tiếp đến clone.

2

Tải lên audio mẫu

Mở Nhân bản giọng, kéo thả hoặc chọn file. Hệ thống hiển thị sóng âm. Nghe lại để xác nhận.

3

Điền thông tin giọng

Nhập Tên giọng, chọn Ngôn ngữ, Giới tính, Độ tuổi và mô tả ngắn.

4

Chế độ hiển thị

🌍 Công khai để chia sẻ với cộng đồng, hoặc 🔒 Riêng tư chỉ mình bạn dùng.

5

Lưu và sử dụng

Nhấn Lưu giọng. Giọng xuất hiện ngay trong Studio TTS với nhãn Clone.

ℹ️
Nhập transcript (tùy chọn)Nếu biết nội dung đã nói, nhập vào ô "Transcript" để tăng độ chính xác clone.

Thiết kế giọng nói AI

Tạo giọng hoàn toàn mới bằng mô tả — không cần audio mẫu

1

Chọn ngôn ngữ và giới tính

Chọn ngôn ngữ: Tiếng Việt, English hoặc 中文. Chọn Nam hoặc Nữ làm nền tảng.

2

Viết mô tả giọng (Instruct)

Ví dụ: Giọng nữ trẻ miền Nam, nhẹ nhàng ấm áp, tốc độ vừa phải. Càng chi tiết càng chính xác.

3

Thử với văn bản mẫu

Nhập 1–2 câu và nhấn ▶ Thử giọng. Chỉnh lại mô tả và thử lại — miễn phí không giới hạn.

4

Lưu vào thư viện

Khi hài lòng, nhập tên và nhấn Lưu giọng. Giọng xuất hiện trong Studio với nhãn Design.

Mẹo viết instruct hiệu quảBao gồm: Giới tính + Độ tuổi + Vùng miền + Cảm xúc + Tốc độ. Ví dụ: "Giọng nam trung niên Hà Nội, trang trọng, tốc độ chậm, rõ từng từ".
🎬

Lồng tiếng video

Tải video từ YouTube/TikTok/Facebook hoặc upload file, dịch và lồng tiếng AI tự động

1

Chọn video nguồn

Mở Lồng tiếng. Dán link YouTube/TikTok/Facebook để tải, hoặc upload trực tiếp file (Video: MP4, MKV, MOV, AVI, WEBM · Audio: MP3, M4A, WAV, FLAC — tối đa 500MB).

2

Tải video (nếu dán link)

Chọn chất lượng hoặc 🎵 Chỉ audio nếu không cần hình. Nhấn ⬇️ Tải video rồi 🎙️ Lồng tiếng ngay.

3

Cài đặt giọng lồng tiếng

Chọn Ngôn ngữ đích, Giới tính giọngNgười nói trong video: Nhiều người (AI tự nhận diện từng người nói, chọn giọng phù hợp riêng) hoặc 1 người (1 giọng cho cả video).

4

Tùy chỉnh giọng (tuỳ chọn)

Viết Mô tả giọng tùy chỉnh hoặc chọn 1 Giọng clone đã lưu thay cho giọng mặc định. Chọn preset Master Audio (Podcast/Audiobook/Broadcast) để hậu kỳ chuyên nghiệp.

5

Xem & dịch thử transcript

Nhấn Transcript để xem văn bản gốc AI nhận dạng. Nhấn nút Dịch để xem trước bản dịch sang ngôn ngữ đích trước khi lồng tiếng cả video.

6

Bắt đầu lồng tiếng

Nhấn Bắt đầu lồng tiếng. Pipeline: nhận dạng giọng nói → dịch → tổng hợp TTS từng câu → ghép theo timeline → xuất MP4. Thời gian xử lý tỉ lệ theo độ dài video.

Tải video bị lỗi hoặc chậm?Dùng nút 💻 Tải xuống Local — công cụ nhỏ chạy ngay trên máy bạn, tải video bằng IP nhà riêng (ổn định hơn IP server dùng chung). Transcript/Dịch/Lồng tiếng vẫn dùng credits tài khoản Omnivox.cc như bình thường.
🗂️
Chế độ Hàng loạtDán nhiều link cùng lúc vào tab 🗂️ Hàng loạt, chọn hành động (Download / Tách transcript / Lồng tiếng) áp dụng chung, rồi nhấn ▶ Chạy tất cả để xử lý liên tiếp không cần thao tác từng video.
Thao tácChi phí
Tải video gốc1–5 credits (tuỳ chất lượng, audio-only rẻ nhất)
Lồng tiếng2 credits / phút video
📊

Dashboard người dùng

Xem lịch sử, quản lý giọng đã lưu và theo dõi hoạt động

📋

Tổng quan

Thống kê request hôm nay, tổng audio đã tạo, giọng đã lưu và biểu đồ 7 ngày.

⏱️

Lịch sử audio

Lưu theo gói: Free 72h · Starter 7 ngày · Pro 30 ngày · Business 90 ngày.

🔊

Giọng của tôi

Quản lý giọng clone hoặc thiết kế. Đổi chế độ công khai/riêng tư bất cứ lúc nào.

Audio tự xóa theo TTL của góiFree: 72h · Starter: 7 ngày · Pro: 30 ngày · Business: 90 ngày. Nhớ tải về trước khi hết hạn.

Truy cập Dashboard

Nhấn icon 👤 ở góc phải header. Hoặc vào trực tiếp /dashboard.html.

Phát lại audio

Nhấn nút bên cạnh mỗi entry để nghe trực tiếp trong trang.

⚙️

Thông số nâng cao

Hiểu rõ các tham số để tối ưu chất lượng audio

Tham sốMặc địnhGiải thích
Tốc độ (speed)1.00.8 = chậm rõ, 1.2 = nhanh hơn. Khuyến nghị: 0.85–1.1
Số bước (num_step)32Nhiều hơn = tự nhiên hơn, lâu hơn. Khuyến nghị: 20–40
Guidance Scale2.0Cao hơn = chính xác hơn nhưng kém tự nhiên. Khuyến nghị: 1.5–3.0
Ngôn ngữTự độngĐể Tự động khi văn bản hỗn hợp. Chọn cụ thể khi chỉ dùng 1 ngôn ngữ.
🎧

Preset: Podcast

Tốc độ 0.95 · Bước 32 · Guidance 2.0

📖

Preset: Đọc sách

Tốc độ 0.88 · Bước 40 · Guidance 1.8

Preset: Nhanh

Tốc độ 1.1 · Bước 20 · Guidance 2.5

Câu hỏi thường gặp

Giải đáp nhanh các thắc mắc phổ biến

Tôi có cần tạo tài khoản không?
Có. Đăng nhập Google, chọn tài khoản là xong. Tài khoản mới nhận 1 credit miễn phí.
1 credit bằng bao nhiêu?
Mỗi lần nhấn Tạo giọng tốn 1 credit, bất kể văn bản dài hay ngắn.
Audio của tôi bị xóa khi nào?
Free: 72 giờ · Starter: 7 ngày · Pro: 30 ngày · Business: 90 ngày. Dashboard hiển thị thời gian còn lại.
Giới hạn độ dài văn bản?
Free 1.000 · Starter 5.000 · Pro 15.000 · Business 50.000 ký tự/lần. Tự chia segment và ghép lại.
Hỗ trợ những ngôn ngữ nào?
Tốt nhất: Tiếng Việt, English, Tiếng Trung. Chế độ Tự động nhận diện nhiều ngôn ngữ khác.
Giọng clone có dùng được trong Studio?
Có. Sau khi lưu, xuất hiện trong Giọng của tôi ở thư viện Studio.
Tại sao xử lý lâu?
Lần đầu trong ngày có thể mất 30–60 giây khởi động GPU (cold start). Các lần sau nhanh hơn (~5–15 giây).
Định dạng file output là gì?
WAV 16-bit, 24000 Hz. Chất lượng cao không nén. Có thể convert sang MP3 bằng công cụ bên ngoài.

Sẵn sàng bắt đầu?

Đăng nhập bằng Google, nhận 1 credit miễn phí và tạo audio ngay bây giờ.

▶ Mở Studio ngay
DOCUMENTATION

User Guide

Everything you need to know to turn text into professional speech with Omnivox.cc

🚀

Quick Start

Create your first audio in just 3 simple steps

Step 01
📝

Enter text

Open Studio and paste or type the content you want converted to speech

Step 02
🎤

Choose a voice

Select a language and voice from the library, or use the default AI voice

Step 03
⬇️

Generate & download

Click Generate, preview the audio, then download your WAV file

💡
Sign in with GoogleOmnivox.cc requires a Google account. New accounts receive 1 free credit to try immediately. Upgrade your plan to unlock more credits and advanced features.
👤

Account & Credits

The credit system, plans, and usage limits

1

Sign in with Google

Click Sign in with Google. Your account is created automatically — no form to fill out.

2

What are credits?

Each generation costs 1 credit regardless of text length. New accounts get 1 free credit. Upgrade to get more.

3

Choose a plan

View details at Pricing. Pay via international card (LemonSqueezy) or bank transfer.

PlanCreditsMax chars/requestAudio storage
Free1 credit1,000 chars72 hours
Starter200 credits5,000 chars7 days
Pro800 credits15,000 chars30 days
Business3,000 credits50,000 chars90 days
Voice Clone & Voice Design require Pro or aboveCloning and AI voice design are only available on Pro and higher plans.
🎙️

Studio TTS

Convert text to speech with a wide range of languages and voices

1

Enter your text

Type or paste content. Limits: Free 1,000 · Starter 5,000 · Pro 15,000 · Business 50,000 chars. Long text is auto-split into segments and merged into one continuous audio file.

2

Select language

Choose Vietnamese, English, Chinese, or leave on Auto — the AI detects the language automatically.

3

Choose a voice (optional)

Pick from the voice library. Your cloned and designed voices also appear here.

4

Adjust advanced settings

Tweak Speed, Steps, and Guidance Scale. See the Advanced Params section for details.

5

Generate & download

Click Generate. Processing takes 5–30 seconds. Then hit ▶ Play to preview or ↓ Download to save the WAV.

Tip: Long textThe system automatically splits by sentence and stitches the audio seamlessly. No need to split manually.
🔊

Voice Cloning

Create a voice clone from any audio sample

⚠️
Audio sample requirementsClear speech, no background noise, no music. Ideal length: 15–60 seconds. Formats: WAV, MP3, M4A, FLAC.
1

Prepare your sample

Use a file with clear speech and silent background. Continuous speech (reading a paragraph) works best. Sample quality directly affects clone quality.

2

Upload the sample

Open Voice Clone, drag-and-drop or click to select. The system shows a waveform — listen back to confirm quality.

3

Fill in voice details

Enter a Voice name, select Language, Gender, Age, and a short description.

4

Set visibility

Choose 🌍 Public to share with all users, or 🔒 Private for personal use only. You can change this later in Dashboard.

5

Save & use

Click Save voice. It appears immediately in Studio TTS under the Clone label.

ℹ️
Add a transcript (optional)If you know what was said in the sample, enter it in the "Transcript" field to significantly improve clone accuracy.

AI Voice Design

Create a brand-new voice by description — no audio sample needed

1

Select language & gender

Choose the primary language: Vietnamese, English, or Chinese. Select Male or Female as a base.

2

Write a voice description (Instruct)

Example: Young female voice, warm and gentle, clear pronunciation, moderate pace. More detail = more accurate result.

3

Preview with sample text

Enter 1–2 sentences and click ▶ Preview voice. Tweak the description and retry — it's free and unlimited.

4

Save to library

Happy with the result? Enter a name and click Save voice. It appears in Studio under the Design label.

Instruct writing tipsInclude: Gender + Age + Region/Accent + Emotion + Pace + Style. Example: "Middle-aged male British English, formal, slow and clear, suitable for narration".
🎬

Video Dubbing

Download from YouTube/TikTok/Facebook or upload a file, translate, and auto-dub with AI

1

Choose a source video

Open Dub. Paste a YouTube/TikTok/Facebook link to download, or upload a file directly (Video: MP4, MKV, MOV, AVI, WEBM · Audio: MP3, M4A, WAV, FLAC — up to 500MB).

2

Download the video (if pasting a link)

Pick a quality, or 🎵 Audio only if you don't need the picture. Click ⬇️ Download, then 🎙️ Dub now.

3

Configure the dubbed voice

Choose Target language, Voice gender, and Speakers in video: Multiple (AI detects each speaker and picks a matching voice for each) or Single (one voice for the whole video).

4

Customize the voice (optional)

Write a Custom voice description or pick a saved Cloned voice instead of the default. Choose a Master Audio preset (Podcast/Audiobook/Broadcast) for professional post-processing.

5

Review & preview the translation

Click Transcript to see the AI-recognized original text. Click Translate to preview the translation into the target language before dubbing the whole video.

6

Start dubbing

Click Start dubbing. Pipeline: speech recognition → translation → per-sentence TTS synthesis → timeline assembly → MP4 export. Processing time scales with video length.

Download failing or slow?Use the 💻 Download Local button — a small tool that runs on your own machine, downloading via your residential IP (more reliable than a shared server IP). Transcript/Translate/Dub still use your Omnivox.cc account credits as usual.
🗂️
Batch modePaste multiple links at once in the 🗂️ Batch tab, pick an action (Download / Extract transcript / Dub) to apply to all, then click ▶ Run all to process them one after another without handling each video manually.
ActionCost
Download original video1–5 credits (depends on quality, audio-only is cheapest)
Dub2 credits / minute of video
📊

User Dashboard

View history, manage saved voices, and track your activity

📋

Overview

Today's request count, total audio generated, saved voices, and a 7-day activity chart.

⏱️

Audio history

Stored by plan: Free 72h · Starter 7 days · Pro 30 days · Business 90 days.

🔊

My Voices

Manage cloned or designed voices. Toggle public/private visibility at any time.

Audio auto-deletes per plan TTLFree: 72h · Starter: 7d · Pro: 30d · Business: 90d. Download before expiry if you need to keep the file.

Accessing Dashboard

Click the 👤 icon in any page header. Or go directly to /dashboard.html.

Replaying audio

Click the small button next to any history entry to play it inline.

⚙️

Advanced Parameters

Fine-tune audio quality for your specific use case

ParameterDefaultDescription
Speed1.00.8 = slower & clearer, 1.2 = faster. Recommended: 0.85–1.1
Steps (num_step)32More steps = more natural, longer processing. Recommended: 20–40
Guidance Scale2.0Higher = more accurate pronunciation but less natural. Recommended: 1.5–3.0
LanguageAutoLeave on Auto for mixed text. Set explicitly for single-language content for best accuracy.
🎧

Preset: Podcast

Speed 0.95 · Steps 32 · Guidance 2.0 — Natural, easy listening

📖

Preset: Audiobook

Speed 0.88 · Steps 40 · Guidance 1.8 — Slow and clear

Preset: Quick

Speed 1.1 · Steps 20 · Guidance 2.5 — Good for testing

Frequently Asked Questions

Quick answers to the most common questions

Do I need to create an account?
Yes. Sign in with Google — it takes just a few seconds. New accounts get 1 free credit to try immediately.
How much is 1 credit?
Every time you click Generate costs 1 credit, regardless of text length.
When is my audio deleted?
Free: 72 hours · Starter: 7 days · Pro: 30 days · Business: 90 days. The Dashboard shows time remaining for each file.
What is the text length limit?
Free 1,000 · Starter 5,000 · Pro 15,000 · Business 50,000 characters per request. Long text is auto-split and merged.
Which languages are supported?
Best support: Vietnamese, English, and Mandarin Chinese. Auto mode can detect and handle several other languages.
Can I use a cloned voice in Studio?
Yes. After saving, cloned and designed voices appear under My Voices in the Studio library.
Why does processing take long?
The first request of the day may take 30–60 s to cold-start the GPU. Subsequent requests in the same session are much faster (~5–15 s).
What is the output audio format?
WAV 16-bit, 24,000 Hz — uncompressed, studio quality. You can convert to MP3 or other formats with any external tool.

Ready to get started?

Sign in with Google, claim your free credit, and create audio right now.

▶ Open Studio