面向性能、成本和高可扩展性优化的自定义工作流与工具模型,覆盖图像和视频任务。
MuAPI 的自定义工具和工作流模型提供优化的性能与路由。MuAPI 将自有模型与其他提供商并列,帮助团队比较能力、价格和适合的集成方式。

$0.010 / 秒
Video Background Remover automatically removes the background from any video, producing a clean cutout of the subject with a transparent or solid-color backdrop. It handles hair, edges, and fine detail with frame-accurate matting, supports videos up to 60 seconds, and can output transparent WebM/MOV, standard MP4, or animated GIF while optionally preserving the original audio.
$0.040 / 秒
LatentSync is a video-to-video model that generates lip sync animations from audio using advanced algorithms for high-quality synchronization.
$0.040 / 秒
Realistic lipsync video - optimized for speed, quality, and consistency.
$0.250 / 秒
Convert any video into 175+ languages with synchronized voice translation, AI-voice cloning, and accurate lip sync. Just upload your video (or provide a link), select a target language, and HeyGen recreates the speech in that language. 0.05$ per second.
$0.065 / 秒
The AI Video Watermark Remover is our flagship model designed to remove Sora 2 watermarks, logos, captions, and unwanted text from videos without compromising quality. Supporting a wide range of formats, it's fast, efficient, and processes with the highest quality.
$0.200 / 秒
Ovi is a unified model that generates synchronized video and audio from textual input. You write a scene description, including dialogue and ambient sounds, and Ovi produces a short video clip (typically ~5 seconds) where visuals and sound align naturally. Videos are generated in 540p resolution.
$0.040 / 秒
Generate realistic lipsync from any audio using VEED's latest model
$0.300 / 秒
Bring your characters and worlds to life with AI Dance Effects — a creative video effect that adds playful, dynamic, and cinematic motion to your generations. AI Dance Effects lets you guide how characters move, react, and express themselves.
$0.200 / 秒
InfiniteTalk Image-to-Video brings still portraits and character photos to life by generating natural, realistic talking videos. You provide a single face image and a dialogue script, and the model animates lip movement, facial expressions, and subtle head gestures to match the speech.
$0.200 / 秒
InfiniteTalk Video-to-Video enhances or transforms existing videos by syncing the subject’s lip movements and facial expressions with new dialogue or speech. Instead of starting from a still image, you provide a video clip, and the model seamlessly reanimates the speaker’s mouth and expressions to match the script.
$0.240 / 秒
The AI Video Upscaler is a powerful tool designed to enhance the resolution and quality of videos. Whether you're working with low-resolution videos that need a boost or aiming to improve the clarity of existing footage, this upscaler leverages advanced machine learning models to deliver high-quality, upscaled videos.
$0.200 / 秒
Ovi is a unified audio–video generation model that can transform a static image plus a descriptive prompt into a short video with synchronized audio. It supports both text-to-video and image-conditioned video inputs. With built-in lip sync, background audio / sound effects, and dialogue support, Ovi brings still visuals to life in cinematic fashion. Videos are generated in 540p resolution.
$0.000 / 秒
Add custom watermark to videos with adjustable position, opacity, and size. Free local processing using FFmpeg.
$0.200 / 秒
Add AI-generated animated captions to any video using Vadoo's caption engine. Supports multiple languages and viral caption themes like Hormozi style. Perfect for social media creators, marketers, and content producers.

$0.100 / 秒
Drive a video's lip movements to match a target audio track, producing a lip-synced video output.

$0.630 / 秒
Generate animated motion graphics videos from a text prompt using AI-generated React/Remotion code rendered on Modal.

$0.010 / 秒
MMAudio-v2 generates high-quality, synchronized audio from video or text inputs. Seamlessly integrate it with AI video models to create fully-voiced, expressive video content.
$0.040 / 秒
Generate realistic lipsync animations from audio using advanced algorithms for high-quality synchronization.

$0.100 / 秒
Replace faces in videos with stunning realism. Our AI ensures accurate expression transfer, lighting consistency, and smooth frame-by-frame blending.
$0.300 / 秒
AI Video Effects applies advanced visual transformations, color grading, and cinematic filters to create stunning videos from images.
$0.030 / 秒
The AI Video Upscaler is a powerful tool designed to enhance the resolution and quality of videos. Whether you're working with low-resolution videos that need a boost or aiming to improve the clarity of existing footage, this upscaler leverages advanced machine learning models to deliver high-quality, upscaled videos.

$0.025 / 秒
Transform and resize your videos effortlessly with remix video tool.
$0.500 / 秒
Convert long-form videos into engaging short clips using AI clipping.
$0.050 / 秒
Combine multiple short video clips (5s, 10s, etc.) into a single seamless full-length video. Upload your clips in order and choose the final output aspect ratio. 'Auto' preserves the aspect ratio of your first clip.
$0.050 / 秒
Automatically crop and reframe a specific video segment to your chosen aspect ratio using AI subject tracking.

$0.630 / 秒
Edit and modify a previously generated motion graphics animation using a text instruction.
$0.300 / 秒
AI Video Effects applies advanced visual transformations, color grading, and cinematic filters to create stunning videos from images.
$0.300 / 秒
Motion Controls adds dynamic camera movements, speed ramps, and zoom effects to bring your images to life as smooth, engaging videos.
$0.300 / 秒
VFX delivers high-impact visual effects like explosions, particles, and cinematic overlays to transform static images into action-packed videos.

$0.020 / 次生成
Transform blurry or pixelated images into high-definition visuals. Our AI Image Upscaler uses deep learning to reconstruct details and bring your visuals to life.

$0.020 / 次生成
Advanced facial recognition and blending algorithms enable precise face swaps while preserving skin tone, lighting, and facial geometry.

$0.010 / 次生成
Instantly remove image backgrounds with pixel-perfect precision. Ideal for product photos, profile pictures, and creative projects.

$0.060 / 次生成
Instantly generate studio-quality product images with AI. Upload your item photo and get clean, stylized shots perfect for e-commerce, ads, and catalogs.

$0.050 / 次生成
Create professional-grade product photos using AI. Upload your item image and describe it with a prompt, and get studio-style, lifestyle, or creative backgrounds in seconds

$0.015 / 1K tokens
Krea 2 Turbo generates aesthetic-first images natively at 1K to 2K resolution in seconds, supporting both text-to-image and image-to-image workflows.

$0.050 / 次生成
Bring your imagination to life with art inspired by the enchanting world of Studio Ghibli. This AI model generates dreamy, hand-drawn visuals with soft colors, whimsical characters, and painterly backgrounds

$0.030 / 1K tokens
Create stunning anime-style artwork instantly with our AI Anime Generator. Customize characters, scenes, and styles effortlessly in seconds!

$0.050 / 次生成
Easily remove unwanted objects, people, or text from any image using AI. Just select the area you want to erase, and the model will intelligently fill the space with realistic background matching the surrounding environment. No Photoshop skills needed.
$0.020 / 1K tokens
Meta Muse Image Text-to-Image generates images directly from a text prompt, covering subject, scene, composition, lighting, mood, and style in a single pass.

$0.010 / 次生成
Smooth skin, reduce blemishes, and enhance complexion with natural-looking results. Perfect for portraits, selfies, and professional photo retouching.
$0.020 / 次生成
Meta Muse Image Edit applies a text instruction to up to 10 reference images at once, for instruction-driven edits, composites, and style transforms.

$0.002 / 次生成
The SDXL LoRA image model enhances Stable Diffusion XL with specialized fine-tuning, letting you generate images in unique styles, characters, or themes. By applying LoRA weights, you can create visuals that match a specific aesthetic, celebrity look, anime style, or custom-trained subject.

$0.030 / 次生成
Expand the edges of any image with AI. This model continues your original photo or artwork beyond its borders while matching style, lighting, and content.

$0.030 / 次生成
AI Image Effects applies advanced visual transformations, color grading, and cinematic filters to create stunning images from a image.

$0.020 / 1K tokens
Croma Image is an advanced text-to-image generation model designed for high-quality, creative, and versatile visuals. It can produce anything from photorealistic portraits and products to imaginative concept art, fantasy illustrations, and cinematic scenes.

$0.020 / 1K tokens
Pony XL is a high-quality image generation model based on Stable Diffusion XL architecture. It specializes in character art, hybrid styles, and producing detailed, polished visuals even with simpler prompts.

$0.000 / 次生成
Add custom watermark to images with adjustable position, opacity, and size. Free local processing using PIL.

$0.028 / 1K tokens
AI TikTok Carousel Generator — create viral TikTok carousel posts from a single text prompt. Choose a proven storytelling format (Problem-Solution, Listicle, Tutorial, Before & After), set your slide count (3-10), and get stunning AI-generated images at 1080x1920 portrait resolution, ready to upload to TikTok.

$0.010 / 次生成
Professional AI portrait styles including hair, makeup, style, and fashion transformations.

$0.100 / 次生成
Instantly change outfits in images using AI. Visualize different clothing styles without the need for physical trials—perfect for fashion, e-commerce, and virtual try-ons.

$0.010 / 次生成
Automatically add lifelike colors to black-and-white images. Our AI brings history to life with natural tones, accurate shading, and context-aware colorization.

$0.004 / 1K tokens
SDXL is a high-quality, large Stable Diffusion model for creating photorealistic and stylized images from text. It excels at fine detail, realistic lighting, and complex scenes.

$0.020 / 1K tokens
Neta Lumina is a powerful anime-style text-to-image model developed by Neta.art Lab. It’s built on Lumina-Image-2.0, fine-tuned with over 13 million high-quality anime images. It offers strong understanding of multilingual prompts, excellent detail fidelity, support for Danbooru tags, and leaning into niche styles like furry, Guofeng, pets, scenic backgrounds, etc.

$0.020 / 次生成
SeedVR2 is a one-step diffusion-transformer model designed for image restoration, super-resolution, deblurring, and artifact removal. It enhances low-quality or compressed images into clean, sharp, high-resolution results while preserving natural colors and fine details.

$0.040 / 次生成
Change face expressions and camera angles using Qwen Image Edit with predefined angle and expression LoRA weights.

$0.000 / 1K tokens
DeepSeek V4 Pro is an advanced flagship multimodal reasoning model designed for complex coding, mathematical, and multi-step reasoning tasks.

$0.050 / 1K tokens
Analyze domain backlink profile, referring domains, domain authority rank, and link sources.

$0.003 / 1K tokens
Fetch Google My Business verified profile details, categories, contact information, ratings, and coordinates.

$0.001 / 1K tokens
Track target domain search ranking position for a keyword against live Google organic search results.

$0.003 / 1K tokens
Search Google Maps and Local Finder for local business rankings, map pack positions, star ratings, coordinates, and contact details.

$0.005 / 1K tokens
Search global business directory listings by keyword, category, and location with verified contact details, ratings, and coordinates.

$0.005 / 1K tokens
Fetch Google Ads monthly search volume, competition level, low/high CPC bid ranges, and search trends.

$0.010 / 1K tokens
Discover Google Ads keyword ideas and suggested expansions from seed keywords.

$0.010 / 1K tokens
Get DataForSEO Labs comprehensive keyword metrics, search intent, and difficulty for multiple keywords.

$0.010 / 1K tokens
Discover DataForSEO Labs related search keywords, difficulty, search volume, and SERP info.

$0.010 / 1K tokens
Fetch top organic ranking pages and ranking keyword counts for a target domain.

$0.010 / 1K tokens
Fetch complete YouTube video metadata, engagement stats, channel details, and tags (3x SERP).

$0.020 / 1K tokens
Analyze aggregated LLM mentions, brand sentiment, and cross-competitor share of voice across ChatGPT and Google AI.

$0.003 / 1K tokens
Fetch Google My Business Questions and Answers live from Google search.

$0.003 / 1K tokens
Fetch Google My Business posts, announcements, and updates.

$0.003 / 1K tokens
Fetch YouTube organic search rankings and video results for any query.

$0.003 / 1K tokens
Fetch user comments, replies, and like statistics for any YouTube video (billed per 20 comments).

$0.010 / 1K tokens
Convert text into natural-sounding speech using mmAudio-v2. Ideal for voiceovers, virtual assistants, and content narration with lifelike clarity and tone.

$0.033 / 1K tokens
Search and analyze brand mentions, citations, share of voice, and top cited pages across AI search engines (ChatGPT and Google AI Overviews).

$0.020 / 1K tokens
Fetch top backlinked pages for a domain with incoming link metrics and HTTP status.

$0.025 / 1K tokens
Fetch historical backlink growth, referring domains, and link velocity trends over time.

$0.017 / 1K tokens
Run on-page Google Lighthouse audit evaluating Core Web Vitals, Performance, Accessibility, Best Practices, and SEO.
$0.100 / 1K tokens
Generate viral short-form video scripts for social media based on a topic and niche.

$0.008 / 1K tokens
Fetch Google My Business customer reviews, ratings, review timestamps, and owner responses for a local business.

$0.100 / 1K tokens
Generate expressive, multilingual text-to-dialogue content using the ElevenLabs Text To Dialogue V3 model.

$0.010 / 1K tokens
Any LLM is a versatile large language model for text generation, comprehension, and diverse NLP tasks such as chat and summarization. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.025 / 1K tokens
Any LLM is a versatile large language model for text generation, comprehension, and diverse NLP tasks such as chat and summarization. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.050 / 1K tokens
Convert text to natural-sounding speech using the ElevenLabs TTS Turbo 2.5 model, with adjustable stability, similarity, and speed.

$0.000 / 1K tokens
Moonshot Kimi K3 is a flagship 2.8T Mixture-of-Experts (MoE) LLM with a 1M token context window, designed for long-context reasoning, coding, and complex agent workflows. Token-based pricing: $0.70/M input tokens and $2.80/M output tokens. Two endpoints: standard async (/kimi-k3) and live streaming (/kimi-k3/stream) via SSE.

$0.000 / 1K tokens
DeepSeek V4 Flash is an ultra-fast multimodal reasoning model optimized for low-latency text and image understanding tasks.

$0.010 / 1K tokens
Extract complete closed captions, subtitles, and transcript text from any YouTube video (3x SERP).

$0.050 / 1K tokens
Analyze domain organic traffic, ranking keywords distribution, top performing pages, and SERP competitors.

$0.017 / 1K tokens
Discover search volume, CPC, keyword difficulty, intent, and high-value keyword ideas for SEO campaigns.

$0.001 / 1K tokens
Classify any text snippet across OpenAI's standard moderation categories (sexual, hate, harassment, self-harm, violence, and more). Returns a boolean flag plus per-category booleans and confidence scores — a drop-in safety gate for chat inputs, generated text, and free-form prompts.

$0.001 / 1K tokens
Fetch a location- and device-specific Google organic search-result snapshot with normalized rankings, URLs, titles, and SERP features.

$0.025 / 1K tokens
Execute live prompt queries across frontier LLMs (ChatGPT GPT-5, Claude Sonnet 4.5, Gemini 2.5 Pro, Perplexity Sonar) with real-time web search and citation source extraction.

$0.001 / 1K tokens
Track Google organic search ranking positions for up to 50 keywords against a target domain in batch.

$0.020 / 次生成
Detect and extract text fragments and their positions from an image using local OCR.

$0.010 / 次生成
Upload and publish a video to a connected YouTube account.

$0.020 / 次生成
Upload and publish a video to a connected TikTok account.

$0.020 / 次生成
Publish a video or image to a connected Instagram Business account.
$0.010 / 次生成
Fetch the latest Reels metadata and metrics for an Instagram creator.

$0.020 / 次生成
Publish a video or image Pin to a connected Pinterest board.
$0.010 / 次生成
Fetch the latest posts and performance metrics for a Twitter/X user.

$0.010 / 次生成
Detect unsafe or policy-violating content in any image. Supply an image URL (and optional context text) and receive structured flags for adult, violent, hateful, or otherwise restricted content — ideal for pre-screening user uploads before they reach generation pipelines.
$0.010 / 次生成
Retrieve profile details, stats, and metadata for a TikTok user.
$0.010 / 次生成
Fetch the latest video posts and view metrics for a TikTok creator.
$0.010 / 次生成
Fetch the latest Shorts from a YouTube channel by ID.
$0.010 / 次生成
Fetch the latest Reels metadata and metrics for a Facebook page.

$0.020 / 次生成
Publish a video or image to a connected Facebook Page.

$0.020 / 次生成
Publish a video or image to a connected LinkedIn profile or page.

$0.020 / 次生成
Post a video or image to a connected X (Twitter) account.

$0.020 / 次生成
Publish a video or image to a connected Threads account.

$0.010 / 次生成
Set or replace the thumbnail image of a video on a connected YouTube account.

$0.010 / 次生成
Update the title, description, tags, category, privacy, or made-for-kids status of a video on a connected YouTube account.
$0.010 / 次生成
Download videos from YouTube in your chosen resolution or audio format.

$0.020 / 次生成
Scan a video for unsafe or policy-violating content. Supports MP4, MOV, and WebM URLs and returns structured safety classifications across harassment, hate, sexual, sexual-minors, and violence categories — useful for moderating user-generated or AI-generated video before publishing.

$0.015 / 次生成
Krea 2 Turbo LoRA runs custom LoRA adapters and generates personalized images natively at 1K to 2K resolution in sub-second speeds, supporting both text-to-image and image-to-image.