elevenlabs-tts — independently scanned and version-tracked by SaferSkills.
SaferSkills independently audited elevenlabs-tts (Agent Skill) and scored it 100/100 (green). The audit ran 55 deterministic rules across Security, Supply Chain, Maintenance, Transparency, and Community; it found 0 high-severity and 0 lower-severity findings. The full rule-by-rule trace and per-finding evidence are below. Free, methodology-open.
Findings & checks · 0 flagged
Every scanned point with the score it earned and what moved between them.
First recorded scan — no prior version to compare against.
The primary manifest — the file an agent reads to learn what this artifact does.
Install the belt CLI skill: npx skills add belt-sh/cliPremium text-to-speech with 22+ voices via inference.sh CLI.

Requires inference.sh CLI (belt). Install instructionsbelt login
# Generate speech with ElevenLabs
belt app run elevenlabs/tts --input '{"text": "Hello, welcome to our product demo.", "voice": "aria"}'| Model | ID | Best For | Latency |
|---|---|---|---|
| Multilingual v2 | eleven_multilingual_v2 | Highest quality, 32 languages | ~250ms |
| Turbo v2.5 | eleven_turbo_v2_5 | Balance of speed & quality | ~150ms |
| Flash v2.5 | eleven_flash_v2_5 | Ultra-low latency | ~75ms |
| Voice | Style |
|---|---|
aria | American, conversational |
alice | British, confident |
bella | American, warm |
jessica | American, expressive |
laura | American, professional |
lily | British, soft |
sarah | American, friendly |
| Voice | Style |
|---|---|
george | British, authoritative |
adam | American, deep |
bill | American, mature |
brian | American, conversational |
callum | Transatlantic, intense |
charlie | Australian, natural |
chris | American, casual |
daniel | British, commanding |
eric | American, friendly |
harry | American, young |
liam | American, articulate |
matilda | American, warm |
river | American, confident |
roger | American, authoritative |
will | American, bright |
belt app run elevenlabs/tts --input '{"text": "Welcome to our quarterly earnings presentation.", "voice": "george"}'# Highest quality
belt app run elevenlabs/tts --input '{
"text": "This is our premium multilingual model with the best quality.",
"voice": "aria",
"model": "eleven_multilingual_v2"
}'
# Ultra-fast for real-time applications
belt app run elevenlabs/tts --input '{
"text": "Flash model for low-latency applications.",
"voice": "brian",
"model": "eleven_flash_v2_5"
}'belt app run elevenlabs/tts --input '{
"text": "Fine-tune the voice characteristics for your use case.",
"voice": "bella",
"stability": 0.3,
"similarity_boost": 0.9,
"style": 0.4
}'| Parameter | Range | Effect |
|---|---|---|
stability | 0-1 | Higher = more consistent, lower = more expressive |
similarity_boost | 0-1 | Higher = closer to original voice character |
style | 0-1 | Higher = more style exaggeration |
use_speaker_boost | true/false | Enhances speaker clarity |
# High-quality MP3
belt app run elevenlabs/tts --input '{
"text": "High quality audio output.",
"voice": "daniel",
"output_format": "mp3_44100_192"
}'| Format | Description |
|---|---|
mp3_44100_128 | MP3 at 44.1kHz, 128kbps (default) |
mp3_44100_192 | MP3 at 44.1kHz, 192kbps |
pcm_16000 | Raw PCM at 16kHz |
pcm_22050 | Raw PCM at 22.05kHz |
pcm_24000 | Raw PCM at 24kHz |
pcm_44100 | Raw PCM at 44.1kHz |
ElevenLabs supports 32 languages including English, Spanish, French, German, Italian, Portuguese, Chinese, Japanese, Korean, Arabic, Hindi, Russian, and more.
# Spanish
belt app run elevenlabs/tts --input '{
"text": "Hola, bienvenidos a nuestra presentación.",
"voice": "aria",
"model": "eleven_multilingual_v2"
}'
# French
belt app run elevenlabs/tts --input '{
"text": "Bonjour, bienvenue à notre démonstration.",
"voice": "alice",
"model": "eleven_multilingual_v2"
}'# 1. Generate voiceover
belt app run elevenlabs/tts --input '{
"text": "Introducing the future of AI-powered content creation.",
"voice": "george"
}' > voiceover.json
# 2. Create talking head video
belt app run bytedance/omnihuman-1-5 --input '{
"image_url": "https://portrait.jpg",
"audio_url": "<audio-url-from-step-1>"
}'# ElevenLabs multi-speaker dialogue
npx skills add inference-sh/skills@elevenlabs-dialogue
# ElevenLabs voice changer
npx skills add inference-sh/skills@elevenlabs-voice-changer
# ElevenLabs sound effects
npx skills add inference-sh/skills@elevenlabs-sound-effects
# All TTS models (Kokoro, DIA, Chatterbox, Inworld TTS, and more)
npx skills add inference-sh/skills@text-to-speech
# Full platform skill (all 250+ apps)
npx skills add inference-sh/skills@infsh-cliBrowse all audio apps: belt app store --category audio
~30 seconds. Free. No account. Every finding cites a rule and a line of evidence.