Skip to content
Tools / Voice Transformer
Voice Transformer icon

Voice Transformer

Voice swap with emotion preserved

1 free skill. Paid skills start at $0.005. The final price is shown before running. Failed paid calls do not charge.

What is verified

Catalog facts and aggregate ToolRouter calls. Success uses all recorded calls, including caller errors in the total, and is not a controlled benchmark.

2Maintained skills
7Integrated providers
$0.005Paid calls from
2026-03-22Tool updated

Voice Transformer converts a recorded audio clip into a completely different voice while keeping the original emotion, pacing, and delivery intact. You provide source audio and a target voice ID, and the AI rebuilds the speech in that new voice without losing what made the original performance work.

This is fundamentally different from text-to-speech — there's no transcript required, and nuances like rising inflection, pause length, and emotional intensity all carry through to the output. It also includes noise removal to clean up ambient recordings before transformation.

What you can do

  • transform_voice — convert audio from one voice to another, preserving pacing, emotion, and delivery style; includes optional background noise removal

Who it's for

Podcast producers dubbing interviews into different voices for privacy. Video creators matching a brand voice across multiple presenters. Game developers generating character voice variants from a single performance. Localization teams adapting audio without re-recording. Anyone who wants to anonymize a recording or match a specific brand voice.

How to use it

  1. Use the Voice Generator's list_voices skill to browse available target voices and find the voice ID you want
  2. Run transform_voice with the source audio URL and target voice ID
  3. Adjust stability for consistent vs. expressive output, and similarity_boost to stay closer to the target voice
  4. Enable remove_background_noise if the recording has ambient sound
  5. Use seed for reproducible results across multiple takes

Getting started

Have the audio URL ready (MP3 or WAV at a publicly accessible address). Use list_voices in the Voice Generator tool to find your target voice ID before running the transformation.

Permissions and setup

  • ElevenLabs API Key (secret): Optional: use your own ElevenLabs key instead of the platform default Official setup
  • fal.ai API Key (secret): Optional: use your own fal.ai key instead of the platform default Official setup
  • Prodia API Token (secret): Optional: use your own Prodia token instead of the platform default Official setup
  • Higgsfield API Key (secret): Optional: use your own Higgsfield key instead of the platform default Official setup
  • Photalabs API Key (secret): Optional: use your own Photalabs key instead of the platform default Official setup
  • Google AI API Key (secret): Optional: use your own Google AI key instead of the platform default Official setup
  • OpenRouter API Key (secret): Optional: use your own OpenRouter key instead of the platform default Official setup
Transform VoicePricing: paid

Convert speech audio into a different voice while preserving the original emotion, delivery, and pacing. Provide a source audio URL and a target voice ID to produce a transformed audio file with fine-grained control over stability, similarity, and style.

Returns: Transformed audio file path (auto-uploaded), target voice ID, model ID, noise removal status, output format, and file size
List ModelsPricing: free

List available models for this tool, sorted by popularity. Returns provider details and pricing.

Returns: List of available models with pricing and provider info
Loading reviews...

Loading activity...

v0.022026-03-22
  • Added subtitle, expanded description, and agent instructions
v0.012026-03-20
  • Initial release

Copy these instructions to use Voice Transformer in Claude, ChatGPT, Copilot, and more.

Related Tools

Related Categories

Frequently Asked Questions

What does voice transformation actually change?

It swaps the voice while keeping the pacing, emotion, and delivery of the original recording.

Do I need a target voice ID?

Yes. Pick a voice with `list_voices`, then pass that `voice_id` into `transform_voice`.

Can it clean up background noise too?

Yes. Turn on background-noise removal when the source audio has ambient noise or room tone.

Is it language-agnostic?

The default model is English-only, so that is the safest assumption unless the tool output says otherwise.