> ## Documentation Index
> Fetch the complete documentation index at: https://docs.eversince.ai/llms.txt
> Use this file to discover all available pages before exploring further.

# generate audio

> Generate audio: voiceover (a script read by a voice from list_models type voice, in any of the listed languages), music (a prompt, or lyrics on the models that take them), or sound_effect (a prompt). Voiceover and sound effects finish inside the call; music runs as a job. Either way get_jobs names the item. Taken before the provider runs, refunded when it fails without doing the work. Credits.



## OpenAPI

````yaml /openapi.json post /tools/generate_audio
openapi: 3.1.0
info:
  title: Eversince API
  version: '1'
  description: >-
    Every Eversince tool as one POST call, plus the routes beside them: the tool
    list, account and keys, uploads, webhooks and models. The prose is at
    https://docs.eversince.ai.
servers:
  - url: https://eversince.ai/api/v1
security:
  - bearer: []
tags:
  - name: Discovery
    description: The tool list, the same on the MCP server, the REST API and the CLI.
  - name: Workspace
    description: Jobs, balances, settings, templates, skills and the overview.
  - name: Library
    description: >-
      Items, search, import, reading media, comments, boards, calendars, the
      brand kit and public links.
  - name: Timeline
    description: 'Video editing: timelines, clips, captions, sound, language and rendering.'
  - name: Canvas
    description: 'Still editing: canvases, slides, layers and rendering.'
  - name: Generation
    description: Image, video and audio generation, upscaling, cutouts and models.
  - name: Research
    description: 'What platforms publish: pulling and keeping posts.'
  - name: Account
    description: Balances, keys, workspaces and script sessions.
  - name: Uploads
    description: Files into the library by presigned upload or by URL.
  - name: Webhooks
    description: Job results posted to a URL as they complete.
  - name: Models
    description: Generation models and cost estimates.
paths:
  /tools/generate_audio:
    post:
      tags:
        - Generation
      summary: generate audio
      description: >-
        Generate audio: voiceover (a script read by a voice from list_models
        type voice, in any of the listed languages), music (a prompt, or lyrics
        on the models that take them), or sound_effect (a prompt). Voiceover and
        sound effects finish inside the call; music runs as a job. Either way
        get_jobs names the item. Taken before the provider runs, refunded when
        it fails without doing the work. Credits.
      operationId: generate_audio
      parameters:
        - $ref: '#/components/parameters/IdempotencyKey'
        - $ref: '#/components/parameters/WorkspaceId'
      requestBody:
        required: true
        content:
          application/json:
            schema:
              type: object
              properties:
                kind:
                  type: string
                  enum:
                    - voiceover
                    - music
                    - sound_effect
                  description: What to generate.
                text:
                  description: >-
                    voiceover: the script to speak, read exactly as written:
                    spell numbers, symbols, addresses and names the way they
                    should be said.
                  type: string
                voice_id:
                  description: 'voiceover: voice id from list_models(type: "voice").'
                  type: string
                prompt:
                  description: >-
                    music or sound_effect: the prompt. On google music, timed
                    sections ("[0:00 - 0:10] Intro: soft lo-fi beat") are
                    followed, which is how a track lands on a cut; on minimax a
                    named key and tempo ("E minor, 90 BPM") are followed.
                  type: string
                music_model:
                  description: >-
                    music: the provider, default google. google and minimax take
                    lyrics and set their own length (about 3 and 6 minutes);
                    elevenlabs takes a duration. google follows a timed
                    arrangement in the prompt, minimax a named key and tempo.
                    estimate_cost quotes each.
                  type: string
                  enum:
                    - google
                    - minimax
                    - elevenlabs
                duration:
                  description: >-
                    music (elevenlabs, required, 3-600s) / sound_effect target
                    seconds (0.5-22, omit to let the model choose). Ignored by
                    minimax and google music.
                  type: number
                  exclusiveMinimum: 0
                lyrics:
                  description: 'music (minimax/google): lyrics with structure tags.'
                  type: string
                instrumental:
                  description: >-
                    music (elevenlabs, minimax): instrumental, no vocals.
                    Defaults to true. Pass false for a track with singing, and
                    on minimax the model writes the words when you send none.
                  type: boolean
                language:
                  description: >-
                    voiceover: the language to speak in; the values here are the
                    whole supported set. Every voice speaks all of them, so the
                    voice is chosen for its character and this for its language.
                  type: string
                  enum:
                    - en
                    - es
                    - fr
                    - de
                    - it
                    - pt
                    - ja
                    - ko
                    - zh
                    - ar
                    - hi
                    - ru
                    - nl
                    - pl
                    - sv
                    - da
                    - fi
                    - 'no'
                    - cs
                    - sk
                    - hu
                    - ro
                    - bg
                    - hr
                    - uk
                    - el
                    - tr
                    - th
                    - vi
                    - id
                    - ms
                    - fil
                    - ta
                    - te
                    - ml
                    - kn
                    - bn
                    - gu
                    - mr
                    - pa
                    - he
                    - fa
                    - ur
                    - sw
                    - ha
                    - af
                    - ga
                    - cy
                    - is
                    - ca
                    - gl
                    - sl
                    - et
                    - lv
                    - lt
                    - sr
                    - bs
                    - mk
                    - ka
                    - hy
                    - az
                    - kk
                    - ne
                stability:
                  description: >-
                    voiceover: 0-1, default 0.5. Higher holds one delivery
                    across takes; lower is more expressive and varies more.
                  type: number
                  minimum: 0
                  maximum: 1
                similarity_boost:
                  description: >-
                    voiceover: 0-1, default 0.75. How closely the output tracks
                    the source voice.
                  type: number
                  minimum: 0
                  maximum: 1
                style:
                  description: >-
                    voiceover: 0-1, default 0. Raises expressiveness at some
                    cost to stability.
                  type: number
                  minimum: 0
                  maximum: 1
                speed:
                  description: >-
                    voiceover: 0.7-1.2, default 1. Delivery rate. Use it to fit
                    a line to a shot rather than trimming after.
                  type: number
                  minimum: 0.7
                  maximum: 1.2
                speaker_boost:
                  description: >-
                    voiceover: sharpens resemblance to the source voice, at a
                    small latency cost.
                  type: boolean
                prompt_influence:
                  description: >-
                    sound_effect: 0-1, default 0.3. Higher follows the prompt
                    more literally, lower gives the model more room.
                  type: number
                  minimum: 0
                  maximum: 1
                loop:
                  description: >-
                    sound_effect: generate a seamless loop, for ambience held
                    under a whole scene.
                  type: boolean
                seed:
                  description: >-
                    voiceover: seed for a reproducible take. Ignored by music
                    and sound_effect.
                  type: integer
                idempotency_key:
                  type: string
                  minLength: 1
                  maxLength: 200
                  description: >-
                    A key you mint per dispatch (a UUID). The same key on a
                    retry returns the job already dispatched instead of charging
                    again; a new attempt takes a new key.
                context:
                  description: 'Optional: one sentence shown to the user beside this action.'
                  type: string
                  maxLength: 600
              required:
                - kind
                - idempotency_key
      responses:
        '200':
          description: >-
            Done, or a job started for work that runs longer (follow it with
            get_jobs).
          content:
            application/json:
              schema:
                type: object
                required:
                  - ok
                  - data
                  - meta
                properties:
                  ok:
                    const: true
                  data:
                    type: object
                    additionalProperties: true
                  meta:
                    type: object
                    properties:
                      request_id:
                        type: string
                      cost:
                        $ref: '#/components/schemas/Cost'
                      cloud_processing:
                        $ref: '#/components/schemas/CloudProcessing'
                  media:
                    type: array
                    items:
                      type: object
                      additionalProperties: true
        default:
          $ref: '#/components/responses/Error'
components:
  parameters:
    IdempotencyKey:
      name: Idempotency-Key
      in: header
      required: false
      schema:
        type: string
      description: >-
        Retry a paid dispatch safely: the same key returns the job already
        dispatched instead of charging again.
    WorkspaceId:
      name: x-workspace-id
      in: header
      required: false
      schema:
        type: string
      description: The workspace this one call acts in, when it is not the key's own.
  schemas:
    Cost:
      type: object
      description: AI credits the call cost.
      additionalProperties: true
    CloudProcessing:
      type: object
      description: Cloud processing the call used, and what is left.
      properties:
        ran_on:
          const: cloud
        minutes:
          type: number
        queued_minutes:
          type: number
        left_minutes:
          type: number
        bought_minutes:
          type: number
        note:
          type: string
    Error:
      type: object
      required:
        - ok
        - error
      properties:
        ok:
          const: false
        error:
          type: object
          required:
            - code
            - message
            - retryable
          properties:
            code:
              type: string
            message:
              type: string
              description: What went wrong, in words to act on.
            retryable:
              type: boolean
            suggested_action:
              type: string
              description: The next step that works.
            cloud_processing:
              $ref: '#/components/schemas/CloudProcessing'
  responses:
    Error:
      description: The call was refused or failed.
      content:
        application/json:
          schema:
            $ref: '#/components/schemas/Error'
  securitySchemes:
    bearer:
      type: http
      scheme: bearer
      description: >-
        An API key (es_live_…) from Settings, under API keys, or an OAuth access
        token.

````