Agent Skills

elevenlabs-voice-changer

ElevenLabs voice changer - transform any voice to a different voice while preserving speech content and emotion via inference.sh CLI. Models: eleven_multilingual_sts_v2 (70+ languages), eleven_english_sts_v2. Capabilities: speech-to-speech, voice transformation, accent change, voice disguise. Use for: content creation, voice acting, privacy, dubbing, character voices. Triggers: voice changer, speech to speech, voice transformation, change voice, voice swap, voice conversion, voice disguise, elev

Install

npx skills add https://github.com/inference-sh/skills --skill elevenlabs-voice-changer
SKILL.md

Install the belt CLI skill: npx skills add belt-sh/cli

ElevenLabs Voice Changer

Transform any voice into a different voice via inference.sh CLI.

Voice Changer

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

# Transform voice
belt app run elevenlabs/voice-changer --input '{"audio": "https://recording.mp3", "voice": "aria"}'

Available Models

Model ID Best For
Multilingual STS v2 eleven_multilingual_sts_v2 70+ languages (default)
English STS v2 eleven_english_sts_v2 English-optimized

Voice Options

Same 22+ premium voices as ElevenLabs TTS:

Voice Style
george British, authoritative (default)
aria American, conversational
alice British, confident
brian American, conversational
charlie Australian, natural
daniel British, commanding
jessica American, expressive
sarah American, friendly
adam, bella, bill, callum, chris, eric, harry, laura, liam, lily, matilda, river, roger, will Various styles

Examples

Basic Voice Transformation

# Change voice to British male
belt app run elevenlabs/voice-changer --input '{
  "audio": "https://my-recording.mp3",
  "voice": "george"
}'

# Change voice to American female
belt app run elevenlabs/voice-changer --input '{
  "audio": "https://my-recording.mp3",
  "voice": "aria"
}'

Choose Output Format

belt app run elevenlabs/voice-changer --input '{
  "audio": "https://recording.mp3",
  "voice": "daniel",
  "output_format": "mp3_44100_192"
}'

English-Optimized Model

belt app run elevenlabs/voice-changer --input '{
  "audio": "https://english-speech.mp3",
  "voice": "brian",
  "model": "eleven_english_sts_v2"
}'

Workflow: Voice-Over Replacement

# 1. Record yourself reading the script (any quality mic)
# 2. Transform to professional voice
belt app run elevenlabs/voice-changer --input '{
  "audio": "https://my-rough-recording.mp3",
  "voice": "george"
}' > professional.json

# 3. Add to video
belt app run infsh/video-audio-merger --input '{
  "video_file": "video.mp4",
  "audio_file": "<professional-audio-url>"
}'

Workflow: Character Voices

# Record one actor, create multiple characters
# Character 1: British narrator
belt app run elevenlabs/voice-changer --input '{
  "audio": "https://actor-line1.mp3",
  "voice": "george"
}' > char1.json

# Character 2: Young female
belt app run elevenlabs/voice-changer --input '{
  "audio": "https://actor-line2.mp3",
  "voice": "lily"
}' > char2.json

# Character 3: Casual male
belt app run elevenlabs/voice-changer --input '{
  "audio": "https://actor-line3.mp3",
  "voice": "charlie"
}' > char3.json

Use Cases

  • Content Creation: Transform your voice for videos and podcasts
  • Voice Acting: Create multiple characters from one performance
  • Privacy: Anonymize voice in recordings
  • Dubbing: Replace voices in video content
  • Accessibility: Convert to preferred voice characteristics
  • Prototyping: Test different voices before hiring talent

Related Skills

# ElevenLabs TTS (generate from text instead)
npx skills add inference-sh/skills@elevenlabs-tts

# ElevenLabs voice isolator (clean audio first)
npx skills add inference-sh/skills@elevenlabs-voice-isolator

# ElevenLabs dubbing (translate to other languages)
npx skills add inference-sh/skills@elevenlabs-dubbing

# Full platform skill (all apps)
npx skills add inference-sh/skills@infsh-cli

Browse all audio apps: belt app list --category audio

Related skills

video-editgenmedia-labs715KEdit existing video on RunComfy — this skill is a smart router that matches the user's intent to the right edit model in the RunComfy catalog. Picks Wan 2.7 Edit-Video (general restyle / background swap / packaging swap, identity + motion preservation), Kling 2.6 Pro Motion Control (transfer precise motion from a reference video to a target character), or Lucy Edit Restyle (lightweight identity-stable restyle / outfit swap). Bundles each model's documented prompting patterns so the skill gets shai-video-generationgenmedia-labs714KGenerate AI videos on RunComfy via the `runcomfy` CLI — a smart router across the full video-model catalog: HappyHorse 1.0 (Arena #1, native in-pass audio), Wan-AI Wan 2-7 (open weights, audio-driven lip-sync), ByteDance Seedance v2 / 1-5 / 1-0 (multi-modal cinematic), Kling 3.0 / 2-6, Google Veo 3-1, MiniMax Hailuo 2-3, ByteDance Dreamina 3-0. Covers text-to-video (t2v), image-to-video (i2v), and Veo's video-extend endpoint. The skill picks the right model for the user's intent (Arena-#1 qualitai-musicgenmedia-labs714KGenerate AI music on RunComfy via the `runcomfy` CLI — a smart router across the music-model catalog. Routes to ElevenLabs AI Music Generation (premium 44.1 kHz stereo vocal tracks, 5 s–5 min, $0.0083/s) and ACE Step / ACE Step 1.5 (StepFun-AI open-weights, tag-driven composition, multilingual lyrics, $0.0002–0.0003/s, ~27× cheaper), plus ACE Step audio-inpaint (regenerate a time range inside an existing track) and ACE Step audio-outpaint (extend a track before or after). Picks the right model fimage-to-videogenmedia-labs713KAnimate any still image on RunComfy — this skill is a smart router that matches the user's intent to the right i2v model in the RunComfy catalog. Picks HappyHorse 1.0 I2V (Arena #1, native audio, identity preservation) for general animations, Wan 2.7 with `audio_url` for custom-voiceover lip-sync, or Seedance 2.0 Pro for multi-modal animation from image + reference video + reference audio. Bundles each model's documented prompting patterns so the caller gets sharper output without burning iterat

Search skills and MCP servers

Fuzzy search across 23,137 skills and servers