A WAV recording can be uploaded directly and used as the dataset for a cloud RVC training run. The output is a trained .pth model—not a renamed WAV—plus a matching retrieval .index and complete ZIP package.
When you have both an original WAV and a compressed copy of the same recording, the WAV is normally the better training source because it avoids an additional lossy encoding stage.
Active trainings
Selected job
Training details
Working
Model—
Training strength—
Elapsed time—
Stage 1 of 6Receiving audio
Epoch 0 of 2000%
Epoch progress comes directly from the running RVC trainer.
The audio you sent
PTH
Latest training checkpoint
An inference-ready early model is available while training continues.
Your first training is free - then pay only for the cloud training you use. Every account also includes five voice-conversion download minutes.
Your first training is free. After that, one subscription covers every tool on this page with nothing to meter - and every model you train is still a file you download and keep. Cancel any time; it stays yours.
Up to 500 epochs for every balance-funded training.
An exact quote first based on your audio and selected epochs.
Training balance
Pay only when you train
One balance for every NiceVois training. Your balance never expires. For more voice-conversion downloads, choose a monthly plan below.
20% off your first balance purchase. Applied automatically at checkout.
Free to start
Your first training
$0with a free account
Create a free account and your first voice training is free. It also comes with five voice-conversion download minutes. Stem separation, noise removal and reverb removal are free for everyone, with nothing to sign up for.
Checking eligibility...Taster pack
$5 training balance
$4.99$4.99non-expiring
One-time, not a subscription. At current quotes, $5 covers up to 11 one-minute tests at 100 epochs, 4 five-minute runs at 180 epochs, or 2 ten-minute runs at 180 epochs. The exact cost and balance left are shown before training.
Starter pack
$10 training balance
$8.99$8.99non-expiring
Pay different amounts for short tests and longer, stronger training.
Best value
$27.50 training balance
$24.99$24.99includes $2.51 extra
Best for creators training several voices or comparing epoch settings.
Monthly conversion plans
More voice-conversion download minutes
Use the included free minutes first. Subscribe only when you need more conversion downloads; monthly plans also include renewable training credit.
Pricing
Pay once, or subscribe
Buy one or three trainings, each with no limits. Subscribe to the monthly plan if you use the tools often.
25% off your first month. Regular monthly price after that.
Train one voice
$5.00once
One training run, any length, up to the full epoch range. Five minutes of conversion downloads included. No subscription, nothing to cancel.
Train three voices
$10.00once
Three training runs, any length, up to the full epoch range. Fifteen minutes of conversion downloads included. No subscription, nothing to cancel.
Unlimited everything
Every tool, one price
$20.00$20.00per month
Then $20.00 per month.
Every tool, unlimited: voice training, voice conversion, stem separation and cleanup.
Starter
For occasional projects
$8.99$8.99per month
Then $8.99 per month.
$5 monthly training credit and 30 voice-conversion download minutes.
Creator
For regular cover makers
$14.99$14.99per month
Then $14.99 per month.
$10 monthly training credit and 90 voice-conversion download minutes.
Pro
For high-volume work
$24.99$24.99per month
Then $24.99 per month.
$20 monthly training credit and 240 voice-conversion download minutes.
Secure checkout by Lemon SqueezyCancel any time from your accountYour model files are yours to keep
Before you pay
Do I need an account to pay?
No. Pay with any email. Sign in with that same email on any device and everything you paid for is there.
What does unlimited actually mean?
No minutes to count on any tool, every output included, up to 500 epochs per training run. Fair use applies.
How do I cancel?
One click from your account, no email needed. You keep everything until the end of the month you paid for.
What if a training fails?
You are not charged for it. A failed or cancelled run releases what it reserved, and an upload that never arrives does not spend your free run.
Who sees my card?
Lemon Squeezy, the merchant of record. NiceVois never sees card details, and tax is handled at checkout.
Do I keep my voices if I leave?
Yes. Every model is a .pth, .index and ZIP you download, and they work anywhere RVC runs.
What happens to my audio?
It is processed for your job and never shared. Source audio is removed after validation. Models trained on a paid run stay until you delete them; a model from the free run is kept for 7 days, and its expiry date is shown on the model itself.
Can I hear it before I pay?
Yes. Your first training is free, and every finished model comes with a listening proof before you decide anything.
Wrong charge or an unused month? Email info@nicevois.com with the checkout email and date; it is reviewed with Lemon Squeezy and returned to the original payment method. Refund policy · Privacy
Secure checkout is being prepared.
Persistent cloud storage
Your model library
Your voice models are files you own. Download them once and no platform can take them away.
Loading…
☁
Paid model packages stay saved until you delete them.
Your .pth, .index, ZIP, and private listening preview remain available after closing the browser or restarting your computer. A model from the free training run is kept for 7 days and shows its expiry date; download it before then. Full-resolution source audio and temporary training files are removed after validation.
— used
⌁Loading your models
Checking private cloud storage…
What makes a WAV useful for RVC?
The container alone does not guarantee good training data. A clean WAV with one voice, consistent volume, and little echo is useful. A WAV containing a full music mix, clipping, several speakers, or aggressive noise-reduction artifacts is still a weak dataset.
Use the original recording or clean isolated vocal whenever possible.
Remove silence only when it dominates the file; natural pauses are not inherently harmful.
Remove other speakers, backing vocals, and instrumental bleed.
Keep natural variation in pitch, words, pace, and expression.
Do you need to resample or split the WAV yourself?
No. The cloud pipeline reads the uploaded source, slices it into workable segments, normalizes and resamples the training material as required, extracts RMVPE pitch and HuBERT features, and then trains the pinned RVC v2 model. Manual splitting is unnecessary for this interface.
Verified WAV workflow
The project’s real end-to-end proof used a 33,189,998-byte WAV. It passed browser upload, private dataset preparation, cloud training, model validation, and named .pth download. A later 100-epoch run produced 53 prepared segments and a 57,590,877-byte inference model that loaded in the native RVC synthesizer.
These numbers verify compatibility and artifact production; they are not a promise that every recording will produce the same perceptual quality or runtime. See the testing methodology.
WAV upload checklist
Check
Good sign
Warning sign
Speaker
One consistent target voice
Interviewer, duet, or backing vocals
Room
Dry and close
Long echo or distant microphone
Level
Clear without clipping
Crackling or flattened peaks
Background
Quiet or cleanly isolated
Music, crowd, traffic, or hum
Use your WAV as the training source
The uploader checks the duration, recommends an epoch count, and shows each cloud stage.