Yes. An M4A uploads directly and trains an RVC .pth model - no conversion step on your side. M4A is the format an iPhone Voice Memo, a WhatsApp voice note, and most phone recorders produce, which makes it the most common real starting point for a voice dataset.
M4A is a container holding AAC audio. AAC is lossy, but at the bitrates phones record at it holds voice detail well - usually better than an MP3 of the same size. You do not need to export it to WAV first.
Active trainings
Selected job
Training details
Working
Model—
Training strength—
Elapsed time—
Stage 1 of 6Receiving audio
Epoch 0 of 2000%
Epoch progress comes directly from the running RVC trainer.
The audio you sent
PTH
Latest training checkpoint
An inference-ready early model is available while training continues.
Your first training is free - then pay only for the cloud training you use. Every account also includes five voice-conversion download minutes.
Your first training is free. After that, one subscription covers every tool on this page with nothing to meter - and every model you train is still a file you download and keep. Cancel any time; it stays yours.
Up to 500 epochs for every balance-funded training.
An exact quote first based on your audio and selected epochs.
Training balance
Pay only when you train
One balance for every NiceVois training — voice models and image models alike. Your balance never expires. For more voice-conversion downloads, choose a monthly plan below.
20% off your first balance purchase. Applied automatically at checkout.
Free to start
Your first training
$0with a free account
Create a free account and your first voice training is free. It also comes with five voice-conversion download minutes, and ten minutes each on the stem separator, noise remover and reverb remover.
Checking eligibility...Taster pack
$5 training balance
$4.99$4.99non-expiring
One-time, not a subscription. At current quotes, $5 covers up to 11 one-minute tests at 100 epochs, 4 five-minute runs at 180 epochs, or 2 ten-minute runs at 180 epochs. The exact cost and balance left are shown before training.
Starter pack
$10 training balance
$8.99$8.99non-expiring
Pay different amounts for short tests and longer, stronger training.
Best value
$27.50 training balance
$24.99$24.99includes $2.51 extra
Best for creators training several voices or comparing epoch settings.
Monthly conversion plans
More voice-conversion download minutes
Use the included free minutes first. Subscribe only when you need more conversion downloads; monthly plans also include renewable training credit.
Pricing
Pay once, or subscribe
Buy one training or three, with nothing to cancel. The monthly plan covers every tool if you use them often.
25% off your first month. Regular monthly price after that.
Train one voice
$5.00once
One training run, any length, up to the full epoch range. Five minutes of conversion downloads included. No subscription, nothing to cancel.
Train three voices
$10.00once
Three training runs, any length, up to the full epoch range. Fifteen minutes of conversion downloads included. No subscription, nothing to cancel.
Unlimited everything
Every tool, one price
$20.00$20.00per month
Then $20.00 per month.
Every tool, unlimited: voice training, voice conversion, stem separation and cleanup.
Starter
For occasional projects
$8.99$8.99per month
Then $8.99 per month.
$5 monthly training credit and 30 voice-conversion download minutes.
Creator
For regular cover makers
$14.99$14.99per month
Then $14.99 per month.
$10 monthly training credit and 90 voice-conversion download minutes.
Pro
For high-volume work
$24.99$24.99per month
Then $24.99 per month.
$20 monthly training credit and 240 voice-conversion download minutes.
Secure checkout by Lemon SqueezyCancel any time from your accountYour model files are yours to keep
Before you pay
Do I need an account to pay?
No. Pay with any email. Sign in with that same email on any device and everything you paid for is there.
What does unlimited actually mean?
No minutes to count on any tool, every output included, up to 500 epochs per training run. Fair use applies.
How do I cancel?
One click from your account, no email needed. You keep everything until the end of the month you paid for.
What if a training fails?
You are not charged for it. A failed or cancelled run releases what it reserved, and an upload that never arrives does not spend your free run.
Who sees my card?
Lemon Squeezy, the merchant of record. NiceVois never sees card details, and tax is handled at checkout.
Do I keep my voices if I leave?
Yes. Every model is a .pth, .index and ZIP you download, and they work anywhere RVC runs.
What happens to my audio?
It is processed for your job and never shared. Source audio is removed after validation; finished models stay until you delete them.
Can I hear it before I pay?
Yes. Your first training is free, and every finished model comes with a listening proof before you decide anything.
Wrong charge or an unused month? Email info@nicevois.com with the checkout email and date; it is reviewed with Lemon Squeezy and returned to the original payment method. Refund policy · Privacy
Secure checkout is being prepared.
Persistent cloud storage
Your model library
Your voice models are files you own. Download them once and no platform can take them away.
Loading…
☁
Finished model packages stay saved until you delete them.
Your .pth, .index, ZIP, and private listening preview remain available after closing the browser or restarting your computer. Full-resolution source audio and temporary training files are removed after validation.
— used
⌁Loading your models
Checking private cloud storage…
What phone audio does well, and where it hurts
Phone recordings are convenient and usually close to the mouth, which is exactly what training wants. The risk is everything else the room contributed.
Record indoors, away from fans, traffic, and open windows - the model learns room noise as part of the voice.
Hold the phone a consistent distance away; large volume swings teach the model an inconsistent timbre.
Turn off any “voice enhancement” or noise-suppression mode if your recorder offers one. Aggressive suppression chews holes in consonants.
Prefer the original file over one that has travelled through a messaging app twice - each re-encode costs detail.
Voice notes from messaging apps
A WhatsApp or iMessage voice note is usable, but it is encoded for small size rather than fidelity. If you have the same words recorded in the Voice Memos app, use that instead. When a voice note is all you have, gather several minutes of it rather than relying on one short clip - length partly compensates for a compressed source.
What the service does with the M4A
The official RVC training workflow reads supported source audio through FFmpeg, decodes it into the representation preprocessing requires, and extracts the features used for training. NiceVois runs that workflow in a pinned cloud environment and validates the exported model before showing the downloads. Nothing about the M4A container needs handling on your side.
M4A versus the alternatives
Question
M4A (AAC)
WAV
Can this tool upload it?
Yes
Yes
Compression
Lossy, but efficient at voice bitrates
Uncompressed
Typical source
Phone recorder, voice note, AAC export
Desktop recording, DAW export
File size
Small
Large
Best choice
Use when the recording was made on a phone
Prefer when you recorded straight to your computer
What you receive
The finished model is available as a named .pth, a matching .index, and a ZIP containing both. The public workspace retains the completed files for the displayed seven-day window, and you can delete an inactive job from your private model library at any time.
Have a voice memo?
Upload the M4A directly - the service prepares the training dataset for you.