elevenlabs

Convert text into speech and generate sound effects via the ElevenLabs API.

Updated Apr 8, 2026
One-click install
npx skills add https://github.com/aimentor606/aether --skill elevenlabs-aimentor606
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill.
Skill: elevenlabs
Source: https://github.com/aimentor606/aether/tree/main/core/kortix-master/opencode/skills/GENERAL-KNOWLEDGE-WORKER/elevenlabs
Command: npx skills add https://github.com/aimentor606/aether --skill elevenlabs-aimentor606

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) components.

What problem does it solve?

Convert written text and prompts into high-quality spoken audio and sound effects so teams can produce narration, voiceovers, and personalized audio without studio recording or manual editing.

Core Features & Use Cases

  • Text-to-Speech: Generate natural, multilingual speech from text with selectable models and output formats.
  • Voice Cloning: Create custom voices from audio samples and reuse them for personalized messages or branded narration.
  • Batch Processing & SFX: Convert entire documents into audio files, split by paragraphs, and create ambient or effect audio from prompts.
  • Use Case: Turn a product guide into an audiobook, create podcast intros in a branded voice, or generate accessibility narration for documents and slides.

Quick Start

Generate a narrated MP3 of the file report.md using voice Rachel and save it as report_narration.mp3.

Frequently Asked Questions about elevenlabs

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I convert a text document into natural-sounding speech for narration?▼

To convert text into natural-sounding speech, this tool processes written documents and outputs high-quality spoken audio files. It supports selectable models and voice tuning parameters to produce narration without manual editing.

Can I generate audio from multiple text files at once for batch processing?▼

Yes, batch processing is supported for audio generation. You can convert entire documents into audio files, automatically splitting the text by paragraphs to produce multiple standard audio segments.

Do I need an ElevenLabs API key to generate voiceovers and sound effects?▼

Yes, an ElevenLabs API key is required to authenticate requests for generating voiceovers and sound effects. You must provide this key to access the text-to-speech, voice cloning, and audio generation features.

How does voice cloning work for creating personalized audio?▼

Voice cloning works by creating custom voices from provided audio samples. Once cloned, these voices can be reused for personalized messages or branded narration across your text-to-speech outputs.

What is the best way to generate ambient sound effects from a text prompt?▼

The best way to generate sound effects from a prompt is using the built-in SFX feature. It converts descriptive text prompts into ambient or effect audio, supplementing your standard text-to-speech generation.

Can I use text-to-speech to create accessibility narration for slides and documents?▼

Yes, you can use text-to-speech to create accessibility narration for slides and documents. It converts written content into high-quality spoken audio, eliminating the need for studio recording or manual editing.