N NiceVois Private workspace
RVC model training · No local GPU required

Train an RVC voice model online

Upload clean voice audio and receive a validated .pth, feature .index, and complete model package. No software installation or notebook setup.

New model

Start from an audio file

1 · Configure
.pth

This name is used for the .pth, .index, and complete ZIP package.

Standard is selected automatically.
200 epochs

Recommended: Add an audio file for a duration-based recommendation.

Clean audio produces a better model. Use a dry voice recording with minimal music, echo, or background noise.

Persistent cloud storage

Your model library

Loading…
Finished model packages stay saved until you delete them.

Your .pth, .index, and ZIP remain available after closing the browser or restarting your computer. Source audio and temporary training files are removed after validation.

Loading your models

Checking private cloud storage…

Direct, portable output

From audio to an RVC model in three steps

This is model training, not a file-format rename. The service prepares your audio, runs RVC training on a cloud GPU, validates the artifacts, and returns files you can keep.

  1. 1Upload clean voice audio

    Choose a supported file, name the finished model, and use the duration-based epoch recommendation.

  2. 2Follow real training progress

    See dataset preparation, GPU queueing, live epochs, validation, and any available checkpoint.

  3. 3Download the complete model

    Take the named .pth, the retrieval .index, or a ZIP containing the complete portable package.

Explore NiceVois

Tools and guidance for the complete RVC workflow

Go straight to the input you have, or learn how to prepare, train, and use the files you receive.

Common questions

Before you train

Is this literally converting audio into a .pth file?

It is training an RVC voice model from the audio. A .pth is the resulting model-weight file, so the operation requires dataset preparation and GPU training rather than a simple format conversion.

Which audio formats can I upload?

The current interface accepts WAV, MP3, FLAC, M4A, and OGG. Clean, dry voice audio with minimal background sound produces the most reliable result.

How much audio should I use?

The public beta currently accepts up to five minutes. Use as much clean, single-speaker material as the limit allows; recordings shorter than one minute may not produce a reliable model.

Do I need an NVIDIA GPU or local RVC installation?

No. Training runs on an external cloud GPU. Once training has started, you can close the page or shut down your computer and return later.

What can I download?

You can download the named .pth by itself, the .index by itself, or a ZIP package containing both. When available, an early inference-ready checkpoint can also be downloaded while the final run continues.

How long are beta files stored?

Public-beta model files are available for seven days. Successful source audio and temporary training files are removed after validation. Download the package before its displayed expiry date.