gemini-audio-vision API — MuAPI: Tạo Văn Bản AI

Tích hợp mô hình AI gemini-audio-vision qua API hiệu suất cao và giá cả linh hoạt.

📝

Overview

About this model

Truy cập lập trình vào mô hình gemini-audio-vision trên MuAPI với độ trễ thấp và khả năng mở rộng quy mô tức thì.

1Tạo nội dung đa phương tiện
2Tích hợp sản phẩm AI
3Xử lý quy trình tự động
💰

Pricing & Value

Cost analysis

muapiapp$2.00/M input tokens, $5.00/M output tokens (higher per-run minimum for audio)

Token-based billing with no subscription — pay only for what you use, including the higher token cost of audio input.

Fal.aiKhông khả dụng

Fal.ai ; Gemini API 。

ReplicateKhông khả dụng

Replicate Gemini 。

** Competitor pricing is estimated based on similar model architectures and usage tiers.

⚙️

Technical Details

Configuration schema

Câu lệnh (Prompt)string

Default ValueDescribe what is said and any background sounds in this audio, including speaker changes and tone.
Âm thanh URLstring

URL。

Default Valuehttps://d3adwkbyhxyrtq.cloudfront.net/webassets/audiomodels/sample-audio.mp3
Câu lệnh (Prompt)string

system-level instruction to guide the model's analysis style.

Default ValueRespond with a structured JSON analysis, not prose.
modelEnum (1 options)

Gemini 。

Default Valuegemini-2.5-flash
📖

Implementation Guide

Developer documentation

Gửi yêu cầu POST tới endpoint với thông số mong muốn và khóa API MuAPI của bạn.

Common Questions

Frequently asked