# ElevenLabs

> Cinematic-quality AI voice cloning + TTS

Source: https://aistack.sh/tool/elevenlabs
Category: Voice
Site: https://elevenlabs.io
Status: active
Last verified by a human: 2026-09-05

The voiceover engine behind most faceless YouTube and audiobook stacks. Voice cloning, multilingual, low-latency streaming.

## Pricing

As published on the tool page: $6/mo Starter · $22/mo Creator

Read from https://elevenlabs.io/pricing on 2026-09-05:

- Free: $0 per month
- Starter: $6 per month
- Creator: $22 per month
- Pro: $99 per month
- Scale: $299 per month
- Business: $990 per month
- Enterprise: custom / contact per month
- Starter Annual: $60 per year
- Creator Annual: $220 per year
- Pro Annual: $990 per year
- Scale Annual: $2990 per year
- Business Annual: $9900 per year

## Alternatives

- openai-voice — https://aistack.sh/tool/openai-voice

## Used in 5 tested stacks

- [Faceless YouTube](https://aistack.sh/stack/faceless-youtube) — Voiceover, $92/mo
- [Faceless YouTube (budget)](https://aistack.sh/stack/faceless-youtube-budget) — Voiceover, $6/mo
- [TikTok character](https://aistack.sh/stack/tiktok-character) — Character voice, $92/mo
- [TikTok character (audio + text)](https://aistack.sh/stack/tiktok-character-audio) — Character voice, $42/mo
- [UGC ad creative](https://aistack.sh/stack/ugc-ads) — Voiceover, $67/mo

## Recent changes

- **Scribe v2 Medical speech recognition model released** (2026-09-11, release) — Scribe v2 Medical is now generally available as a specialized batch speech recognition model for medical and clinical audio. It maintains the same billing rate as Scribe v2 and supports keyterm prompting, entity detection, speaker diarization, and no-verbatim mode. Pass scribe_v2_medical as the model_id parameter. Source: https://elevenlabs.io/docs/changelog
- **Twilio outbound calls now support answering machine detection** (2026-09-07, feature) — Outbound calls through Twilio can now detect whether they reach a human or answering machine. Configure detection mode to trigger early detection or wait for voicemail greetings to complete. Results arrive via a new answering_machine_detection webhook event. Source: https://elevenlabs.io/docs/changelog
- **ElevenAgents WebSocket responses now support file attachments** (2026-09-07, feature) — Agent response WebSocket events include optional attachments with URL, name, and optional MIME type. Enables agents to send files and documents to users during conversations. Source: https://elevenlabs.io/docs/changelog
- **Twilio answering machine detection added to ElevenAgents outbound telephony** (2026-09-07, feature) — Outbound telephony configurations now support Twilio answering machine detection with two modes: early human-or-machine verdict or waiting for voicemail greeting to finish. Detection results are delivered via the new answering_machine_detection webhook event. Source: https://elevenlabs.io/docs/changelog
- **Agent response attachments in ElevenAgents conversations** (2026-09-07, feature) — agent_response WebSocket events now include an optional attachments field, allowing agents to send files with metadata including URL, filename, and MIME type. Source: https://elevenlabs.io/docs/changelog
- **ElevenAgents adds workspace conversation management and Twilio answering machine detection** (2026-09-07, feature) — ElevenAgents gained workspace-wide conversation ticket management with cursor pagination and filtering, dynamic variable filters supporting comparison operators (eq, gt, gte, lt, lte), and Twilio answering machine detection for outbound calls with early or post-greeting detection modes. Agent responses now support file attachments, and knowledge base RAG queries include source URLs. Source: https://elevenlabs.io/docs/changelog
- **Phone keypad (DTMF) input support added to ElevenAgents conversations** (2026-08-31, feature) — Agent conversation configuration now accepts dtmf_input_settings allowing agents to receive phone keypad input with configurable timeout, hash terminator, and optional redaction from logs. Enables IVR-style interaction patterns in agent conversations. Source: https://elevenlabs.io/docs/changelog
- **DTMF phone keypad input support in ElevenAgents conversations** (2026-08-31, feature) — Agent conversation configuration now supports DTMF (phone keypad) input through optional dtmf_input_settings. Developers can configure digit timeout, use # as a terminator, and optionally redact keypad entries from transcripts and logs. Source: https://elevenlabs.io/docs/changelog
- **File attachment history in conversation transcripts** (2026-08-31, feature) — Conversation transcript entries now include file_inputs array containing details about every attached file (ID, name, MIME type, and URL). This enables developers to access complete context about files used during agent conversations. Source: https://elevenlabs.io/docs/changelog
- **DTMF phone keypad input support for agents** (2026-08-31, feature) — ElevenAgents conversations now support DTMF (phone keypad) input with configurable digit timeout and optional # terminator. Developers can also redact keypad entries from transcripts and logs. Source: https://elevenlabs.io/docs/changelog

---

From aistack.sh. Last updated 2026-09-05.
