sag

Generate speech from text via the ElevenLabs API with voice selection.

455|34|Updated Mar 9, 2026
One-click install
npx skills add https://github.com/understudy-ai/understudy --skill sag-understudy-ai
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill.
Skill: sag
Source: https://github.com/understudy-ai/understudy/tree/main/skills/sag
Command: npx skills add https://github.com/understudy-ai/understudy --skill sag-understudy-ai

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

This Skill provides a command-line interface for generating speech from text using ElevenLabs, offering a user experience similar to the macOS say command.

Core Features & Use Cases

  • Text-to-Speech Generation: Convert written text into spoken audio using ElevenLabs voices.
  • Voice Selection & Customization: Choose from various ElevenLabs voices and apply specific pronunciation or delivery rules.
  • Use Case: Generate an audio file for a chatbot response in a specific character voice, like a "crazy scientist," for enhanced user engagement.

Quick Start

Use the sag skill to speak the phrase "Hello there" using the default voice.

Frequently Asked Questions about sag

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate speech from text using ElevenLabs TTS?▼

To generate speech from text using ElevenLabs TTS, you provide your text and an ElevenLabs API key, select a voice, and the Skill synthesizes an audio file for local playback. It functions similarly to the macOS say command.

Can I use ElevenLabs voice selection for a specific character like a crazy scientist?▼

Yes, you can use ElevenLabs voice selection to assign specific character voices, such as a crazy scientist, to your audio generation. The Skill supports applying delivery modifications and pronunciation rules to match the persona.

Do I need an ElevenLabs API key to use text-to-speech generation?▼

Yes, you need an ElevenLabs API key to authenticate and use the text-to-speech generation capabilities. The Skill relies on the ElevenLabs TTS API to synthesize spoken audio from your written text.

What is the best way to add voice responses to a chatbot using speech synthesis?▼

The best way to add voice responses to a chatbot using speech synthesis is generating an audio file with a selected ElevenLabs voice. The Skill creates these audio files specifically for integration into conversational agents.

How does the macOS say command UX compare to this ElevenLabs speech synthesis tool?▼

The macOS say command UX provides a simple command-line interface for immediate local playback, which this ElevenLabs speech synthesis tool mimics. It adds high-quality voice selection, pronunciation rules, and audio tag-based delivery modifications.

Are there limitations when applying pronunciation rules in text-to-speech generation?▼

Limitations in text-to-speech generation with pronunciation rules depend entirely on the ElevenLabs API capabilities. The Skill passes your defined delivery modifications and pronunciation rules to the API to generate the audio file.