ElevenLabs

ElevenLabs

ElevenLabs Music

ElevenLabs Music generates prompt-driven audio tracks up to five minutes long. Reach for it when your application needs background scores, soundtracks, or royalty-free music from text descriptions.

Modalities

Text → Audio

Price

$0.135 / sec

Speed

slow

Performance

Live production data from real requests on Kyma — not synthetic benchmarks.

Rank

#74

of 87 active models

Tokens served

56

all-time

Total requests1
Platform share0.0%

Pricing

Per second of generated video. Failed jobs are refunded in full.

$0.135 / sec

When to use ElevenLabs Music

Updated 2026-07-31

Where this model earns its cost — and where it doesn't.

This model converts text prompts into audio, supporting both lyrical and instrumental requests. It accepts a 2000-token context window and produces audio outputs up to five minutes per generation.

On Kyma, the model runs on the premium tier with slower generation speeds. It supports prompt caching, which bills repeated prefixes at 10% of the standard input rate. Responses return exact generation cost in usage.cost and identify the active model via the X-Kyma-Model header, with automatic failover handling routing if a path degrades.

The model does not support reasoning, vision, or structured outputs. It accepts text-only input and returns audio-only output. It is optimized for batch or asynchronous generation rather than real-time streaming.

Generate background tracks

Create royalty-free audio for videos, games, or applications.

Build custom soundtracks

Produce instrumental or lyrical compositions up to five minutes.

Prototype theme music

Iterate on text prompts to test audio concepts before final production.

Automate content audio

Integrate prompt-driven music generation into media pipelines.

Not ideal for: Do not use this model for real-time audio streaming, speech synthesis, or latency-sensitive interactive voice applications.

Quick start

Up and running in under two minutes.

  1. 1

    Create an API key

    Sign up and grab a key from the dashboard — $0.50 free credit, no card required.

    Get API key →
  2. 2

    Make your first request

    Submit a generation job and poll until it succeeds.

    curl https://kymaapi.com/v1/videos/generations \
      -H "Authorization: Bearer YOUR_API_KEY" \
      -H "Content-Type: application/json" \
      -d '{
        "model": "elevenlabs-music",
        "prompt": "A wave breaking on a rocky shore at golden hour, cinematic",
        "duration": 5
      }'
    # Poll: GET /v1/jobs/{id} until status="succeeded"

FAQ

Common questions about this model.

How much does ElevenLabs Music cost?

$0.135 per sec. Per second of generated video. Failed jobs are refunded in full.

How do I use ElevenLabs Music?

Kyma is OpenAI-compatible: point your SDK's base URL at https://kymaapi.com/v1, use your Kyma API key, and set the model to elevenlabs-music. Signing up is free and includes $0.50 of credit — no card required.

Can I pass lyrics to the prompt?

Yes, the model accepts text prompts describing lyrics or instrumental arrangements to generate audio.

How does prompt caching work for this model?

Kyma applies a 90% discount to repeated prompt prefixes, billing cached input at 10% of the standard rate.

What happens if the generation endpoint fails?

Kyma automatically reroutes the request to a healthy serving path without requiring client-side retries.

Start with $0.50 free credit — no card required.Create account →

More models by ElevenLabs

ModelContextInputOutput
ElevenLabsElevenLabs v3$0.405 / 1K char
ElevenLabsElevenLabs Flash v2.5$0.2025 / 1K char
ElevenLabsElevenLabs Turbo v2.5$0.2025 / 1K char
ElevenLabsElevenLabs Sound Effects$0.027 / call
ElevenLabsElevenLabs Multilingual v2$0.405 / 1K char