minimax-speech-2.6-turbo API — MuAPI: AI Audio & Music

Zintegruj model AI minimax-speech-2.6-turbo za pomocą wydajnych interfejsów API z niskimi opóźnieniami i płatnością za rzeczywiste zużycie.

📝

Overview

About this model

Uzyskaj dostęp programistyczny do modelu minimax-speech-2.6-turbo na platformie MuAPI z natychmiastową skalowalnością i niskimi opóźnieniami.

1Generowanie treści multimedialnych
2Integracja produktów AI
3Zautomatyzowane przepływy pracy
💰

Pricing & Value

Cost analysis

muapiapp$0.65 每次生成

muapiapp is 20-50% more affordable than its competitors while delivering comparable or superior quality.

Fal.ai$0.85 每次生成

Fal.ai 收费 about 20-50% more 每次生成 compared to muapiapp, ensuring muapiapp remains the more cost-effective solution without compromising on 质量.

Replicate$0.85 每次生成

Replicate's 定价 is nearly identical to Fal.ai, making muapiapp a 20-50% 更实惠 option with equal or better 性能.

** Competitor pricing is estimated based on similar model architectures and usage tiers.

⚙️

Technical Details

Configuration schema

Promptstring

Text to convert to speech. Every character is 1 token. Maximum 10000 characters. Use <#x#> between words to control pause duration (0.01-99.99s).

Default ValueWelcome to Minimax-Speech 2.6 by Muapiapp! Get ready for an audio revolution! We are thrilled to introduce a model so realistic, it's virtually indistinguishable from a human voice. You're going to be amazed by its lifelike delivery!
Voice IDEnum (472 options)

Desired voice ID. Use a voice ID you have trained (https://muapi.ai/playground/minimax-voice-clone), or one of the following system voice IDs

Default ValueFriendly_Person
Speedint

Speech speed. Range: 0.5-2.0, where 1.0 is normal speed.

Default Value1
Volumeint

Speech volume. Range: 0.1-10.0, where 1.0 is normal volume.

Default Value1
Pitchint

Speech pitch. Range: -12 to 12, where 0 is normal pitch.

Default Value0
EmotionEnum (7 options)

The emotion of the generated speech.

Default Valuesurprised
English Normalizationboolean

This parameter supports English text normalization, which improves performance in number-reading scenarios.

Default Valuefalse
Sample RateEnum (6 options)

Sample rate of generated sound.

Default Value8000
BitrateEnum (4 options)

Bitrate of generated sound.

Default Value32000
ChannelEnum (2 options)

he number of channels of the generated audio. 1: mono, 2: stereo.

Default Value1
FormatEnum (4 options)

Format of generated sound.

Default Valuemp3
Language BoostEnum (41 options)

Enhance the ability to recognize specified languages and dialects.

Default Valueauto
📖

Implementation Guide

Developer documentation

Wyślij żądanie POST do punktu końcowego z wymaganymi parametrami i kluczem API MuAPI.

Common Questions

Frequently asked