Skip to content

Repository files navigation

elevenlabs (elevenlabs)

Converts text into speech using a voice of your choice and returns audio.

APIs.json: https://raw.githubusercontent.com/api-evangelist/elevenlabs/refs/heads/main/apis.yml

Timestamps

  • Modified: 2026-05-19

APIs

ElevenLabs Text to Speech API

The ElevenLabs Text to Speech API converts text into lifelike spoken audio with nuanced intonation, pacing, and emotional awareness. It supports multiple output formats including MP3, PCM, and mu-law, and offers a range of models such as Flash v2.5 for ultra-low latency real-time applications and Multilingual v2 for support across 70+ languages. Developers can select from thousands of pre-built voices or use custom cloned voices to generate speech that sounds natural and expressive.

Tags

  • AI
  • Audio
  • Speech Synthesis
  • Text to Speech
  • Voice

Properties

ElevenLabs Speech to Text API

The ElevenLabs Speech to Text API provides state-of-the-art transcription capabilities, converting spoken audio into accurate text. It supports multiple audio formats and languages, enabling developers to build applications that require reliable audio transcription. The API is designed for both real-time and batch processing use cases.

Tags

  • AI
  • Audio
  • Speech to Text
  • Transcription

Properties

ElevenLabs Voice Cloning API

The ElevenLabs Voice Cloning API allows developers to create custom AI voices from audio recordings. Instant Voice Cloning requires as little as 60 seconds of clean audio to generate a usable voice clone, while Professional Voice Cloning produces higher fidelity results from a minimum of 30 minutes of recordings. Cloned voices can then be used with the Text to Speech API for generating speech that closely matches the original speaker.

Tags

  • AI
  • Audio
  • Voice
  • Voice Cloning

Properties

ElevenLabs Voices API

The ElevenLabs Voices API provides management capabilities for the voice library, including listing, retrieving, creating, editing, and deleting voices. Developers can access a library of over 5,000 pre-built voices and manage their own custom voices. The API also supports voice design, allowing creation of new AI voices from text descriptions specifying desired characteristics such as accent, age, and tone.

Tags

  • AI
  • Voice Library
  • Voice Management
  • Voices

Properties

ElevenLabs Sound Effects API

The ElevenLabs Sound Effects API generates cinematic sound effects from text descriptions. Developers can describe the desired sound in natural language and receive high-quality audio output. The API supports audio tags for controlling delivery, emotion, emphasis, pauses, and specific sound effects, making it suitable for game development, film production, and multimedia content creation.

Tags

  • AI
  • Audio Generation
  • Sound Effects

Properties

ElevenLabs Audio Isolation API

The ElevenLabs Audio Isolation API removes background noise from audio recordings, isolating vocal tracks from ambient sounds and interference. This is useful for cleaning up recordings, improving audio quality for podcasts and interviews, and preparing audio files for further processing such as voice cloning or transcription. The API processes audio files and returns cleaned versions with the vocal content preserved.

Tags

  • Audio Isolation
  • Audio Processing
  • Noise Removal

Properties

ElevenLabs Dubbing API

The ElevenLabs Dubbing API enables automatic translation and voice-over of audio and video content into different languages. It preserves the original speaker's voice characteristics while translating the spoken content, supporting seamless localization of multimedia content. The API handles the full dubbing pipeline including transcription, translation, and speech synthesis with lip-sync timing.

Tags

  • Audio
  • Dubbing
  • Localization
  • Translation
  • Video

Properties

ElevenLabs Voice Changer API

The ElevenLabs Voice Changer API performs speech-to-speech conversion, replacing one voice with another while preserving the original speech content, timing, and emotional delivery. Developers can transform audio recordings to sound like a different speaker using any voice from the ElevenLabs library or a custom cloned voice. This is useful for content creation, privacy protection, and character voice generation.

Tags

  • Audio Processing
  • Voice Changer
  • Voice Conversion

Properties

ElevenLabs Music Generation API

The ElevenLabs Music Generation API creates music from text prompts, allowing developers to generate original musical compositions programmatically. Users describe the desired genre, mood, tempo, and instrumentation in natural language and receive generated audio output. The API is designed for applications that need background music, jingles, or custom soundtracks without requiring manual composition.

Tags

  • AI
  • Audio Generation
  • Music

Properties

ElevenLabs Conversational AI API

The ElevenLabs Conversational AI API enables developers to build interactive voice agents that can engage in natural, real-time conversations. It combines speech recognition, language understanding, and speech synthesis into a unified interface supporting multi-turn dialogue across 70+ languages. The API is designed for building customer service agents, voice assistants, and interactive voice response systems with expressive, human-sounding voices.

Tags

  • AI
  • Conversational AI
  • Real-Time
  • Voice Agents

Properties

ElevenLabs Studio API

The ElevenLabs Studio API provides programmatic access to the ElevenLabs Studio project management system. Developers can create, manage, and render long-form audio content projects through the API, organizing text into chapters and assigning different voices to different sections. The Studio is designed for producing audiobooks, podcasts, and other long-form audio content at scale.

Tags

  • Content Management
  • Projects
  • Studio

Properties

Common Properties