Minimax Text to Speech API Docs | fal

fal-ai/minimax/preview/speech-2.5-turbo

Generate fast speech from text prompts and different voices using the MiniMax Speech-02 Turbo model, which leverages advanced AI techniques to create high-quality text-to-speech.

Table of contents

  1. Calling the API
  2. Authentication
  3. Queue
  4. Files
  5. Schema

About

Convert text to speech using MiniMax 2.5 Turbo API.

1. Calling the API

Install the client

The client provides a convenient way to interact with the model API.

npm install --save @fal-ai/client

Setup your API Key

Set FAL_KEY as an environment variable in your runtime.

export FAL_KEY="YOUR_API_KEY"

Submit a request

The client API handles the API submit protocol. It will handle the request status updates and return the result when the request is completed.

import { fal } from "@fal-ai/client";

const result = await fal.subscribe("fal-ai/minimax/preview/speech-2.5-turbo", {
  input: {
    text: "Hello world! This is a test of the text-to-speech system."
  },
  logs: true,
  onQueueUpdate: (update) => {
    if (update.status === "IN_PROGRESS") {
      update.logs.map((log) => log.message).forEach(console.log);
    }
  },
});
console.log(result.data);
console.log(result.requestId);

2. Authentication

The API uses an API Key for authentication. It is recommended you set the FAL_KEY environment variable in your runtime when possible.

API Key

In case your app is running in an environment where you cannot set environment variables, you can set the API Key manually as a client configuration.

import { fal } from "@fal-ai/client";

fal.config({
  credentials: "YOUR_FAL_KEY"
});

3. Queue

Submit a request

The client API provides a convenient way to submit requests to the model.

import { fal } from "@fal-ai/client";

const { request_id } = await fal.queue.submit("fal-ai/minimax/preview/speech-2.5-turbo", {
  input: {
    text: "Hello world! This is a test of the text-to-speech system."
  },
  webhookUrl: "https://optional.webhook.url/for/results",
});

Fetch request status

You can fetch the status of a request to check if it is completed or still in progress.

import { fal } from "@fal-ai/client";

const status = await fal.queue.status("fal-ai/minimax/preview/speech-2.5-turbo", {
  requestId: "764cabcf-b745-4b3e-ae38-1200304cf45b",
  logs: true,
});

Get the result

Once the request is completed, you can fetch the result. See the Output Schema for the expected result format.

import { fal } from "@fal-ai/client";

const result = await fal.queue.result("fal-ai/minimax/preview/speech-2.5-turbo", {
  requestId: "764cabcf-b745-4b3e-ae38-1200304cf45b"
});
console.log(result.data);
console.log(result.requestId);

4. Files

Some attributes in the API accept file URLs as input. Whenever that's the case you can pass your own URL or a Base64 data URI.

Data URI (base64)

You can pass a Base64 data URI as a file input. The API will handle the file decoding for you.

Hosted files (URL)

You can also pass your own URLs as long as they are publicly accessible.

Uploading files

We provide a convenient file storage that allows you to upload files and use them in your requests. You can upload files using the client API and use the returned URL in your requests.

import { fal } from "@fal-ai/client";

const file = new File(["Hello, World!"], "hello.txt", { type: "text/plain" });
const url = await fal.storage.upload(file);

5. Schema

Input

text string* required

Text to convert to speech (max 5000 characters, minimum 1 non-whitespace character)

{
  "text": "Hello world! This is a test of the text-to-speech system.",
  "output_format": "hex"
}

Output

audio File* required

The generated audio file

{
  "audio": {
    "url": "https://fal.media/files/kangaroo/kojPUCNZ9iUGFGMR-xb7h_speech.mp3"
  }
}