HiAPI
OverviewModel MarketplaceAPI KeysUsage StatisticsCall LogsBillingReferralPlaygroundStorageChangelogContact UsSettings
Display unit
N
Powered by hiapi
Settings

Welcome

⌘K
Provider(0)
Task(0)
SDK
Speed
Logic

Ready to start building?

Sign up for 200 Credits.

Developer avatarDeveloper avatarDeveloper avatar
10k+

Over 10,000+ developers have joined

Contact Us

AI Model Marketplace: Text, Image, Video & Audio APIs

The HiAPI model marketplace brings together text LLMs, image, video, music, and speech models. Available models include one-key access, usage-based pricing, an online Playground, and request examples; upcoming models include capability and launch previews. Filter by provider or task type.

  • 851 Background Remover API — image generation, usage-based pricing; see live pricing. Turn product, portrait and pet images into transparent PNGs. Explore five real before-and-after examples, input parameters and an API quickstart.
  • claude-opus-4-8 API — text generation, usage-based pricing; see live pricing. Claude Opus 4.8 supports text generation, streaming responses, and tool calling through an OpenAI-compatible API. This route may include additional upstream context. For the cleanest native API experience, choose the Claude Opus 4.8 AWS model.
  • claude-opus-4-8@aws API — text generation, usage-based pricing; see live pricing. Claude Opus 4.8 is designed for complex tasks and high-quality text generation. It supports text and image input, streaming responses, tool calling, and configurable reasoning effort. The model is available through an OpenAI-compatible API and is suitable for code generation, technical analysis, long-form processing, structured output, and automated API workflows.
  • Claude Sonnet 4.6 API — text generation, usage-based pricing; see live pricing. Claude Sonnet 4.6 supports text conversations, image understanding, adaptive thinking, streaming and tool calling.
  • DeepSeek V4 Flash API — text generation, usage-based pricing; see live pricing. DeepSeek V4 Flash is an open-weight language model for high-throughput generation, reasoning, coding, and agent workflows, with a 1M-token context window, thinking and non-thinking modes, JSON output, and tool calling.
  • DeepSeek V4 Flash Vision Exp API — text generation, usage-based pricing; see live pricing. DeepSeek V4 Flash Vision Exp accepts text and image input for multimodal understanding and agent workflows. It supports the Responses API and remains compatible with Chat Completions.
  • DeepSeek V4 Pro API — text generation, usage-based pricing; see live pricing. DeepSeek V4 Pro (0813) targets complex reasoning, coding, and agent workflows. It natively supports the Responses API and remains compatible with Chat Completions.
  • DeepSeek V4.1 Flash API — text generation, usage-based pricing; see live pricing. DeepSeek V4.1 Flash GA supports Responses and Chat Completions with text and image understanding.
  • ElevenLabs Text to Dialogue API — music generation, usage-based pricing; see live pricing. Create multi-speaker dialogue with ElevenLabs Text to Dialogue. Hear a real audio drama, map voices to lines, review pricing, and copy API code.
  • Eleven Music v2 API — music generation, coming soon. ElevenLabs' next-generation music model for prompt-led songs, instrumentals, long-form composition, reference-guided sound, and section-level editing.
  • FLUX 1.1 Pro API — image generation, usage-based pricing; see live pricing. Generate polished photography, architecture, food, nature, and commercial visuals with custom sizing and prompt upsampling controls.
  • FLUX.2 Image to Image API — image generation, usage-based pricing; see live pricing. Edit and compose images with FLUX.2 Image to Image using up to eight references, auto ratio, 1K/2K output, a matched watercolor example, and API quickstart code.
  • FLUX.2 [klein] 4B Image to Image API — image generation, usage-based pricing; see live pricing. A compact low-cost editing route for commercial refinements, background changes, recoloring, and rapid reference-based iteration.
  • FLUX.2 [klein] 4B Text to Image API — image generation, usage-based pricing; see live pricing. Try FLUX.2 [klein] 4B Text to Image for low-cost drafts and commercial visuals with eleven ratios, five megapixel tiers, seed control, and API code.
  • FLUX.2 [klein] 9B Image to Image API — image generation, usage-based pricing; see live pricing. Edit with FLUX.2 [klein] 9B Image to Image using one reference, real cleanup and text-replacement examples, five ratios, four to eight inference steps, and API code.
  • FLUX.2 [klein] 9B Text to Image API — image generation, usage-based pricing; see live pricing. Try FLUX.2 [klein] 9B Text to Image with real event, product, magazine, and neon examples, five aspect ratios, four to eight inference steps, and API code.
  • FLUX.2 Pro API — image generation, usage-based pricing; see live pricing. Black Forest Labs FLUX.2 Pro: high-fidelity text-to-image with strong prompt adherence and crisp in-image text.
  • FLUX.3 Video API — video generation, usage-based pricing; see live pricing. Try the FLUX.3 Video API with text prompts, 1 to 10 keyframes, source-video continuation, synchronized audio, Draft previews, and per-second pricing.
  • FLUX.1 [schnell] API — image generation, usage-based pricing; see live pricing. A fast, low-cost 1-4 step model for drafts, batch generation, and rapid prompt exploration.
  • GLM-5.3 API — text generation, usage-based pricing; see live pricing. Zhipu GLM-5.3 text model, available through an OpenAI-compatible Chat Completions endpoint.
  • GPT-5.6 Luna API — text generation, usage-based pricing; see live pricing. GPT-5.6 Luna targets cost-sensitive, high-throughput workloads such as bulk classification, extraction, rewriting, and lightweight coding. HiAPI currently exposes text, multi-turn, and function tool calling through the streaming Responses API.
  • GPT-5.6 Sol API — text generation, usage-based pricing; see live pricing. GPT-5.6 Sol is the flagship reasoning model in OpenAI's GPT-5.6 family for complex coding, professional analysis, and demanding agent workflows. HiAPI currently exposes text, multi-turn, and function tool calling through the streaming Responses API.
  • GPT-5.6 Terra API — text generation, usage-based pricing; see live pricing. GPT-5.6 Terra balances intelligence, speed, and cost for everyday coding, document work, and general agent workflows. HiAPI currently exposes text, multi-turn, and function tool calling through the streaming Responses API.
  • GPT-6 Astra API — text generation, usage-based pricing; see live pricing. GPT-6 Astra offers streaming text, five reasoning levels, function calls and JSON output through the Responses API.
  • GPT Image 2 API — image generation, usage-based pricing; see live pricing. GPT Image 2 API turns text or reference images into posters and product visuals, with background and style edits. Try it online or integrate it into your app.
  • gpt-image-2.5-flare API — image generation, usage-based pricing; see live pricing. Explore GPT Image 2.5 Flare for fast creative work, with photography, ads, comics, pixel art and recurring-character examples. Generate images online or integrate with HiAPI.
  • gpt-image-2.5-sunburst API — image generation, usage-based pricing; see live pricing. Explore GPT Image 2.5 Sunburst for detailed creation, multi-reference outfits, text localization, object removal and campaign revisions through HiAPI.
  • Grok Imagine 1.5 Image to Video API — video generation, usage-based pricing; see live pricing. Try Grok Imagine 1.5 Image to Video online. Animate a source image, review supported durations and resolutions, copy prompts, pricing, and POST /v1/tasks API code.
  • Grok Image 2.0 Image to Image API — image generation, usage-based pricing; see live pricing. Edit one public reference image for recoloring, restyling, scene replacement, and detail refinement while following its framing.
  • Grok Image 2.0 Text to Image API — image generation, usage-based pricing; see live pricing. Generate product, editorial, landscape, and character visuals in low or medium quality at 1K/2K across square to ultrawide formats.
  • Grok Imagine Image to Image API — image generation, usage-based pricing; see live pricing. Edit with Grok Imagine Image to Image using up to three references, a matched cabin-to-night example, fourteen ratios, flat 1k/2k pricing, and API code.
  • Grok Imagine Image to Video API — video generation, usage-based pricing; see live pricing. Try Grok Imagine Image to Video online. Animate a source image, review supported durations and resolutions, copy prompts, pricing, and POST /v1/tasks API code.
  • Grok Imagine Quality Image to Image API — image generation, usage-based pricing; see live pricing. Use Grok Imagine Quality Image to Image with up to three references, a matched night-scene edit, fourteen ratios, 1k/2k quality pricing, and API code.
  • Grok Imagine Quality Text to Image API — image generation, usage-based pricing; see live pricing. Create hero-grade cinematic images with Grok Imagine Quality Text to Image, thirteen ratios, 1k/2k tiers, real API examples, and quickstart code.
  • Grok Imagine Text to Image API — image generation, usage-based pricing; see live pricing. Generate cinematic landscapes and night scenes with Grok Imagine Text to Image, thirteen aspect ratios, flat 1k/2k pricing, real API examples, and quickstart code.
  • Grok Imagine Text to Video API — video generation, usage-based pricing; see live pricing. Try Grok Imagine text-to-video online, watch public API results, compare 480p and 720p pricing, and integrate POST /v1/tasks.
  • Hailuo 2.3 Fast Image to Video API — video generation, usage-based pricing; see live pricing. Try Hailuo 2.3 Fast Image to Video online. Animate a source image, review supported durations and resolutions, copy prompts, pricing, and POST /v1/tasks API code.
  • Hailuo 2.3 Image to Video API — video generation, usage-based pricing; see live pricing. Try Hailuo 2.3 Image to Video online. Animate a source image, review supported durations and resolutions, copy prompts, pricing, and POST /v1/tasks API code.
  • Hailuo 2.3 Text to Video API — video generation, usage-based pricing; see live pricing. Try Hailuo 2.3 text-to-video online, watch real motion and physics results, compare fixed 6- and 10-second pricing, and call POST /v1/tasks.
  • HappyHorse 1.0 API — video generation, usage-based pricing; see live pricing. Generate 3-15 second text-to-video clips at 720p or 1080p across five common aspect ratios.
  • HappyHorse 1.1 Image to Video API — video generation, usage-based pricing; see live pricing. Try HappyHorse 1.1 Image to Video online. Animate a source image, review supported durations and resolutions, copy prompts, pricing, and POST /v1/tasks API code.
  • HappyHorse 1.1 Reference to Video API — video generation, usage-based pricing; see live pricing. Try HappyHorse 1.1 Reference to Video reference-to-video online. Use reference images to guide subjects, style, and scenes, then review pricing, prompts, and POST /v1/tasks API code.
  • HappyHorse 1.1 Text to Video API — video generation, usage-based pricing; see live pricing. Generate 3-15 second 720p/1080p clips across nine aspect ratios for social, cinematic, and campaign video.
  • HeyGen API — video generation, coming soon. Create avatar-led explainers, complete videos from one prompt, and localized versions of existing footage with HeyGen's video API family.
  • heygen-avatar-v API — video generation, usage-based pricing; see live pricing. HeyGen Avatar V digital human video: drive a preset avatar with text or audio, with lip-sync, backgrounds, transparent WebM, and captions.
  • Ideogram V4 API — image generation, usage-based pricing; see live pricing. Try Ideogram V4 for accurate in-image text, logos, posters, and signage with TURBO, BALANCED, and QUALITY tiers plus real API examples.
  • kimi-k3 API — text generation, usage-based pricing; see live pricing. Kimi K3 is a reasoning model from Moonshot AI. Use Chat Completions for text conversations with three reasoning effort levels and streaming output.
  • Kling 3.0 Omni Image to Video API — video generation, usage-based pricing; see live pricing. Try Kling 3.0 Omni Image to Video online. Animate a source image, review supported durations and resolutions, copy prompts, pricing, and POST /v1/tasks API code.
  • Kling 3.0 Omni Text to Video API — video generation, usage-based pricing; see live pricing. Create 3-15 second videos at 720p, 1080p, or 4K with square, landscape, vertical, and optional sound output.
  • Kling 3.0 Turbo Image to Video API — video generation, usage-based pricing; see live pricing. Try Kling 3.0 Turbo Image to Video online. Animate a source image, review supported durations and resolutions, copy prompts, pricing, and POST /v1/tasks API code.
  • Kling 3.0 Turbo Text to Video API — video generation, usage-based pricing; see live pricing. Try Kling 3.0 Turbo text-to-video in the HiAPI Playground. Compare real 720p results, copy tested prompts, review 3–15 second parameters and pricing, then call POST /v1/tasks.
  • Kling 4.0 API — video generation, coming soon. Kling 4.0 is Kuaishou’s next-generation flagship AI video model for text-to-video and image-to-video—built for higher visual quality, steadier character consistency, controllable camera motion, and stronger audio-visual performance across ads, short drama, and social.
  • Lyria 3 Pro API — music generation, usage-based pricing; see live pricing. Try Google Lyria 3 Pro for AI music generation. Hear a playable sample, review prompt and optional image guidance, pricing, and API code.
  • Lyria 3.5 API — music generation, usage-based pricing; see live pricing. Lyria 3.5 turns prompts or lyrics into complete songs for songwriting, video soundtracks, and music ideas. Try it online and view pricing and API examples.
  • MAI-Image-2.6 API — image generation, coming soon. MAI-Image-2.6 is Microsoft’s next-generation text-to-image and image editing model—built for higher visual quality, controllable edits, multi-reference workflows, and web grounding across ads, e-commerce product shots, and design.
  • MiniMax H3 API — video generation, usage-based pricing; see live pricing. MiniMax H3 combines native 2K text-to-video, first and last frame control, and multimodal reference video under one model ID. Compare the workflows, check live pricing, and validate a result in Playground before integrating it.
  • MiniMax H3 Max API — video generation, usage-based pricing; see live pricing. MiniMax H3 Max: one model for text-to-video and optional first/last-frame guidance, with 480P / 768P and 5–15 second output.
  • MiniMax Music 1.5 API — music generation, usage-based pricing; see live pricing. Generate complete Chinese or English songs with MiniMax Music 1.5. Test real tracks, structured lyrics, audio settings, pricing, and API code.
  • MiniMax Music 2.6 API — music generation, usage-based pricing; see live pricing. Generate full songs with MiniMax Music 2.6 using custom lyrics, automatic lyrics, or instrumental mode. Hear real tracks and copy API code.
  • MiniMax Music 3 API — music generation, usage-based pricing; see live pricing. Generate songs with MiniMax Music 3 using a music prompt and lyrics. Review duration-based pricing, parameters, workflow, and API code.
  • Nano Banana API — image generation, usage-based pricing; see live pricing. A general-purpose image generator for illustration, photography, product concepts, and social visuals.
  • Nano Banana 2 API — image generation, usage-based pricing; see live pricing. Create or remix images with optional references, broad aspect ratios, and higher-resolution output.
  • Nano Banana 2 Lite API — image generation, usage-based pricing; see live pricing. A fixed-1K entry route for high-volume drafts, fast iteration, and optional multi-reference remixing.
  • Nano Banana Pro API — image generation, usage-based pricing; see live pricing. Generate high-detail commercial images or remix optional references with broad ratios and up to 4K output.
  • Qwen-Audio 3.0 TTS Flash API — music generation, usage-based pricing; see live pricing. Try Qwen-Audio 3.0 TTS Flash for low latency text-to-speech. Hear a real multilingual sample, compare voices and controls, and copy API code.
  • Qwen-Audio 3.0 TTS Plus API — music generation, usage-based pricing; see live pricing. Try Qwen-Audio 3.0 TTS Plus for high quality text-to-speech. Hear a real multilingual sample, compare voices and controls, and copy API code.
  • Qwen Image 2.0 API — image generation, usage-based pricing; see live pricing. A cost-efficient Chinese-friendly model with negative prompts, prompt extension, optional seed, and five 2K pixel-size presets.
  • Qwen Image 2.0 Pro API — image generation, usage-based pricing; see live pricing. Try Qwen Image 2.0 Pro with real Chinese typography, portrait, product, and concept examples, five 2K-class sizes, prompt extension, and API code.
  • Qwen Image 3.0 Image to Image API — image generation, usage-based pricing; see live pricing. Use Qwen Image 3.0 Image to Image with up to three references, a verified product recolor example, eight ratios, 1K/2K output, and API code.
  • Qwen Image 3.0 Pro Image to Image API — image generation, usage-based pricing; see live pricing. Edit images with Qwen Image 3.0 Pro using up to three references, exact cover-text replacement, 1K/2K pricing, eight ratios, and API quickstart code.
  • Qwen Image 3.0 Pro API — image generation, usage-based pricing; see live pricing. Use the Pro route for denser typography, architecture, science editorials, luxury products, and final campaign layouts.
  • Qwen Image 3.0 API — image generation, usage-based pricing; see live pricing. Create typography-led posters, editorial scenes, and product visuals with Chinese and English text, prompt extension, and 1K/2K output.
  • Recraft Remove Background API — image generation, usage-based pricing; see live pricing. Remove backgrounds with the Recraft API: one image URL in, transparent PNG out with hair and fur-level edges preserved. See a real cutout example, flat per-image pricing, and quickstart code.
  • Seedance 2.0 API — video generation, usage-based pricing; see live pricing. A multimodal video route supporting text-only generation, endpoint frames, image/video/audio references, 4-15 seconds, and up to 4K output.
  • Seedance 2.0 Fast API — video generation, usage-based pricing; see live pricing. A faster multimodal Seedance route with text, endpoint frames, optional references, generated audio, and 480p/720p output.
  • Seedance 2.0 Mini API — video generation, usage-based pricing; see live pricing. A lower-cost multimodal route for 4-15 second drafts with endpoint frames, optional references, generated audio, and 480p/720p output.
  • Seedance 2.5 Image to Video API — video generation, usage-based pricing; see live pricing. Try Seedance 2.5 image-to-video with first-frame and first/last-frame control. Compare real 720p results, copy tested prompts, review 4–30 second pricing, and call POST /v1/tasks.
  • Seedance 2.5 Reference to Video API — video generation, usage-based pricing; see live pricing. Try Seedance 2.5 reference-to-video with video, image, and audio references. Learn @video1 mappings, review 4–30 second pricing, and call POST /v1/tasks.
  • Seedance 2.5 Text to Video API — video generation, usage-based pricing; see live pricing. Try Seedance 2.5 text-to-video online, review a real API result, compare 720p and 1080p pricing, and integrate POST /v1/tasks.
  • Seedream 4.5 Image to Image API — image generation, usage-based pricing; see live pricing. ByteDance Seedream 4.5 image editing: unified generation-editing architecture with up to 14 reference images for edits and composites, consistent subjects, 2K/4K output.
  • Seedream 4.5 Text to Image API — image generation, usage-based pricing; see live pricing. Try Seedream 4.5 Text to Image with real landscape, food, sci-fi, and Chinese-aesthetic examples, eight ratios, 2K/4K output, and API code.
  • Seedream 5.0 Lite Image to Image API — image generation, usage-based pricing; see live pricing. Edit with Seedream 5.0 Lite Image to Image using up to 14 references, real recolor, text replacement, and composite examples, 2K or 4K output, and API code.
  • Seedream 5.0 Lite Text to Image API — image generation, usage-based pricing; see live pricing. Try Seedream 5.0 Lite Text to Image with real poster, infographic, banner, and photography examples, eight aspect ratios, 2K or 4K output, and API code.
  • Seedream 5.0 Pro Image to Image API — image generation, usage-based pricing; see live pricing. Edit with Seedream 5.0 Pro Image to Image using one to ten references, real recolor, weather, and multi-image examples, 1K or 2K pricing, and API code.
  • Seedream 5.0 Pro Text to Image API — image generation, usage-based pricing; see live pricing. Try Seedream 5.0 Pro Text to Image with real product, portrait, poster, and cinematic examples, 1K or 2K pricing, eight aspect ratios, and API code.
  • Veo 3.1 Fast Image to Video API — video generation, usage-based pricing; see live pricing. Try Veo 3.1 Fast Image to Video online. Animate a source image, review supported durations and resolutions, copy prompts, pricing, and POST /v1/tasks API code.
  • Veo 3.1 Fast Text to Video API — video generation, usage-based pricing; see live pricing. A faster Veo route for 4, 6, or 8 second text-to-video with 720p-4K output, seeds, negative prompts, and generated audio.
  • Veo 3.1 Image to Video API — video generation, usage-based pricing; see live pricing. Try Veo 3.1 image-to-video online with one source image, 4/6/8-second output, 720p to 4K resolution, native audio, pricing, prompts, and API code.
  • veo-3.1-lite/image-to-video API — video generation, usage-based pricing; see live pricing. Veo 3.1 Lite image-to-video with native audio, 720p or 1080p output, and 4, 6, or 8 second clips.
  • veo-3.1-lite/text-to-video API — video generation, usage-based pricing; see live pricing. Veo 3.1 Lite text-to-video with native audio, 720p or 1080p output, and 4, 6, or 8 second clips.
  • Veo 3.1 Text to Video API — video generation, usage-based pricing; see live pricing. Premium 4, 6, or 8 second text-to-video with 720p-4K output, negative prompts, seeds, and synchronized generated audio.
  • Wan 2.7 Image Text to Image API — image generation, usage-based pricing; see live pricing. Try Wan 2.7 Image Text to Image with Chinese prompts, 1K/2K/4K output, extreme wide and tall ratios, optional thinking mode, and API code.
  • Wan 2.7 Image to Video API — video generation, usage-based pricing; see live pricing. Try Wan 2.7 Image to Video online. Animate a source image, review supported durations and resolutions, copy prompts, pricing, and POST /v1/tasks API code.
  • Wan 2.7 Text to Video API — video generation, usage-based pricing; see live pricing. Try Wan 2.7 text-to-video online, review a cinematic preview, compare 720P and 1080P per-second pricing, and integrate POST /v1/tasks.
  • Wan 3.0 Video API — video generation, usage-based pricing; see live pricing. An all-purpose 2-30 second video route supporting text, first/last frames, reference media, files or webpages, 480P-1080P output, and optional synchronized audio.
  • Z-Image API — image generation, usage-based pricing; see live pricing. Try Z-Image for low-cost Chinese-friendly portraits, signage, city scenes, and food photography with five ratios, real examples, and API code.

Image to Image & Editing APIs

  • 851 Background Remover API (851 Labs)
  • FLUX.2 Image to Image API (Black Forest Labs)
  • FLUX.2 [klein] 4B Image to Image API (Black Forest Labs)
  • FLUX.2 [klein] 9B Image to Image API (Black Forest Labs)
  • gpt-image-2.5-flare API (OpenAI)
  • gpt-image-2.5-sunburst API (OpenAI)
  • Grok Image 2.0 Image to Image API (xAI)
  • Grok Imagine Image to Image API (Grok)
  • Grok Imagine Quality Image to Image API (Grok)
  • Qwen Image 3.0 Image to Image API (Alibaba)
  • Qwen Image 3.0 Pro Image to Image API (Alibaba)
  • Recraft Remove Background API (Recraft)
  • Seedream 4.5 Image to Image API (ByteDance)
  • Seedream 5.0 Lite Image to Image API (ByteDance)
  • Seedream 5.0 Pro Image to Image API (ByteDance)

Chat Completions APIs

  • claude-opus-4-8 API (Anthropic)
  • claude-opus-4-8@aws API (Anthropic)
  • Claude Sonnet 4.6 API (Anthropic)
  • DeepSeek V4 Flash API (DeepSeek)
  • DeepSeek V4 Flash Vision Exp API (DeepSeek)
  • DeepSeek V4 Pro API (DeepSeek)
  • DeepSeek V4.1 Flash API (DeepSeek)
  • GLM-5.3 API (Zhipu AI)
  • kimi-k3 API (Moonshot AI)

Responses APIs

  • DeepSeek V4 Flash Vision Exp API (DeepSeek)
  • DeepSeek V4 Pro API (DeepSeek)
  • DeepSeek V4.1 Flash API (DeepSeek)
  • GPT-5.6 Luna API (OpenAI)
  • GPT-5.6 Sol API (OpenAI)
  • GPT-5.6 Terra API (OpenAI)
  • GPT-6 Astra API (OpenAI)

text-to-dialogue

  • ElevenLabs Text to Dialogue API (ElevenLabs)

Text to Music APIs

  • Eleven Music v2 API (ElevenLabs)
  • Lyria 3 Pro API (Google)
  • Lyria 3.5 API (Google)
  • MiniMax Music 1.5 API (MiniMax)
  • MiniMax Music 2.6 API (MiniMax)
  • MiniMax Music 3 API (MiniMax)

Text to Image APIs

  • FLUX 1.1 Pro API (Black Forest Labs)
  • FLUX.2 [klein] 4B Text to Image API (Black Forest Labs)
  • FLUX.2 [klein] 9B Text to Image API (Black Forest Labs)
  • FLUX.2 Pro API (Black Forest Labs)
  • FLUX.1 [schnell] API (Black Forest Labs)
  • GPT Image 2 API (OpenAI)
  • Grok Image 2.0 Text to Image API (xAI)
  • Grok Imagine Quality Text to Image API (Grok)
  • Grok Imagine Text to Image API (Grok)
  • Ideogram V4 API (Ideogram)
  • MAI-Image-2.6 API (Microsoft)
  • Nano Banana API (Google)
  • Nano Banana 2 API (Google)
  • Nano Banana 2 Lite API (Google)
  • Nano Banana Pro API (Google)
  • Qwen Image 2.0 API (Alibaba)
  • Qwen Image 2.0 Pro API (Alibaba)
  • Qwen Image 3.0 Pro API (Alibaba)
  • Qwen Image 3.0 API (Alibaba)
  • Seedream 4.5 Text to Image API (ByteDance)
  • Seedream 5.0 Lite Text to Image API (ByteDance)
  • Seedream 5.0 Pro Text to Image API (ByteDance)
  • Wan 2.7 Image Text to Image API (Alibaba)
  • Z-Image API (Alibaba)

Text to Video APIs

  • FLUX.3 Video API (Black Forest Labs)
  • Grok Imagine Text to Video API (xAI)
  • Hailuo 2.3 Text to Video API (MiniMax)
  • HappyHorse 1.0 API (MiniMax)
  • HappyHorse 1.1 Text to Video API (MiniMax)
  • HeyGen API (HeyGen)
  • heygen-avatar-v API (HeyGen)
  • Kling 3.0 Omni Text to Video API (Kuaishou)
  • Kling 3.0 Turbo Text to Video API (Kuaishou)
  • Kling 4.0 API (Kuaishou)
  • MiniMax H3 API (MiniMax)
  • MiniMax H3 Max API (MiniMax)
  • Seedance 2.0 API (ByteDance)
  • Seedance 2.0 Fast API (ByteDance)
  • Seedance 2.0 Mini API (ByteDance)
  • Seedance 2.5 Text to Video API (ByteDance)
  • Veo 3.1 Fast Text to Video API (Google)
  • veo-3.1-lite/text-to-video API (Google)
  • Veo 3.1 Text to Video API (Google)
  • Wan 2.7 Text to Video API (Alibaba)
  • Wan 3.0 Video API (Alibaba)

Image to Video APIs

  • FLUX.3 Video API (Black Forest Labs)
  • Grok Imagine 1.5 Image to Video API (xAI)
  • Grok Imagine Image to Video API (xAI)
  • Hailuo 2.3 Fast Image to Video API (MiniMax)
  • Hailuo 2.3 Image to Video API (MiniMax)
  • HappyHorse 1.1 Image to Video API (HappyHorse)
  • Kling 3.0 Omni Image to Video API (Kuaishou)
  • Kling 3.0 Turbo Image to Video API (Kuaishou)
  • MiniMax H3 API (MiniMax)
  • MiniMax H3 Max API (MiniMax)
  • Seedance 2.0 API (ByteDance)
  • Seedance 2.0 Fast API (ByteDance)
  • Seedance 2.0 Mini API (ByteDance)
  • Seedance 2.5 Image to Video API (ByteDance)
  • Veo 3.1 Fast Image to Video API (Google)
  • Veo 3.1 Image to Video API (Google)
  • veo-3.1-lite/image-to-video API (Google)
  • Wan 2.7 Image to Video API (Alibaba)
  • Wan 3.0 Video API (Alibaba)

Video to Video APIs

  • FLUX.3 Video API (Black Forest Labs)

streaming

  • GPT-5.6 Luna API (OpenAI)
  • GPT-5.6 Sol API (OpenAI)
  • GPT-5.6 Terra API (OpenAI)

reasoning

  • GPT-5.6 Luna API (OpenAI)
  • GPT-5.6 Sol API (OpenAI)
  • GPT-5.6 Terra API (OpenAI)

tool-calls

  • GPT-5.6 Luna API (OpenAI)
  • GPT-5.6 Sol API (OpenAI)
  • GPT-5.6 Terra API (OpenAI)

Reference to Video APIs

  • HappyHorse 1.1 Reference to Video API (HappyHorse)
  • MiniMax H3 API (MiniMax)
  • Seedance 2.0 API (ByteDance)
  • Seedance 2.0 Fast API (ByteDance)
  • Seedance 2.0 Mini API (ByteDance)
  • Seedance 2.5 Reference to Video API (ByteDance)
  • Wan 3.0 Video API (Alibaba)

audio-to-video

  • heygen-avatar-v API (HeyGen)

lip-sync

  • heygen-avatar-v API (HeyGen)

avatar-video

  • heygen-avatar-v API (HeyGen)

Text to Speech APIs

  • Qwen-Audio 3.0 TTS Flash API (Alibaba)
  • Qwen-Audio 3.0 TTS Plus API (Alibaba)

851 Labs Model APIs

  • 851 Background Remover API — image generation

Anthropic Model APIs

  • claude-opus-4-8 API — text generation
  • claude-opus-4-8@aws API — text generation
  • Claude Sonnet 4.6 API — text generation

DeepSeek Model APIs

  • DeepSeek V4 Flash API — text generation
  • DeepSeek V4 Flash Vision Exp API — text generation
  • DeepSeek V4 Pro API — text generation
  • DeepSeek V4.1 Flash API — text generation

ElevenLabs Model APIs

  • ElevenLabs Text to Dialogue API — music generation
  • Eleven Music v2 API — music generation

Black Forest Labs Model APIs

  • FLUX 1.1 Pro API — image generation
  • FLUX.2 Image to Image API — image generation
  • FLUX.2 [klein] 4B Image to Image API — image generation
  • FLUX.2 [klein] 4B Text to Image API — image generation
  • FLUX.2 [klein] 9B Image to Image API — image generation
  • FLUX.2 [klein] 9B Text to Image API — image generation
  • FLUX.2 Pro API — image generation
  • FLUX.3 Video API — video generation
  • FLUX.1 [schnell] API — image generation

Zhipu AI Model APIs

  • GLM-5.3 API — text generation

OpenAI Model APIs

  • GPT-5.6 Luna API — text generation
  • GPT-5.6 Sol API — text generation
  • GPT-5.6 Terra API — text generation
  • GPT-6 Astra API — text generation
  • GPT Image 2 API — image generation
  • gpt-image-2.5-flare API — image generation
  • gpt-image-2.5-sunburst API — image generation

xAI Model APIs

  • Grok Imagine 1.5 Image to Video API — video generation
  • Grok Image 2.0 Image to Image API — image generation
  • Grok Image 2.0 Text to Image API — image generation
  • Grok Imagine Image to Video API — video generation
  • Grok Imagine Text to Video API — video generation

Grok Model APIs

  • Grok Imagine Image to Image API — image generation
  • Grok Imagine Quality Image to Image API — image generation
  • Grok Imagine Quality Text to Image API — image generation
  • Grok Imagine Text to Image API — image generation

MiniMax Model APIs

  • Hailuo 2.3 Fast Image to Video API — video generation
  • Hailuo 2.3 Image to Video API — video generation
  • Hailuo 2.3 Text to Video API — video generation
  • HappyHorse 1.0 API — video generation
  • HappyHorse 1.1 Text to Video API — video generation
  • MiniMax H3 API — video generation
  • MiniMax H3 Max API — video generation
  • MiniMax Music 1.5 API — music generation
  • MiniMax Music 2.6 API — music generation
  • MiniMax Music 3 API — music generation

HappyHorse Model APIs

  • HappyHorse 1.1 Image to Video API — video generation
  • HappyHorse 1.1 Reference to Video API — video generation

HeyGen Model APIs

  • HeyGen API — video generation
  • heygen-avatar-v API — video generation

Ideogram Model APIs

  • Ideogram V4 API — image generation

Moonshot AI Model APIs

  • kimi-k3 API — text generation

Kuaishou Model APIs

  • Kling 3.0 Omni Image to Video API — video generation
  • Kling 3.0 Omni Text to Video API — video generation
  • Kling 3.0 Turbo Image to Video API — video generation
  • Kling 3.0 Turbo Text to Video API — video generation
  • Kling 4.0 API — video generation

Google Model APIs

  • Lyria 3 Pro API — music generation
  • Lyria 3.5 API — music generation
  • Nano Banana API — image generation
  • Nano Banana 2 API — image generation
  • Nano Banana 2 Lite API — image generation
  • Nano Banana Pro API — image generation
  • Veo 3.1 Fast Image to Video API — video generation
  • Veo 3.1 Fast Text to Video API — video generation
  • Veo 3.1 Image to Video API — video generation
  • veo-3.1-lite/image-to-video API — video generation
  • veo-3.1-lite/text-to-video API — video generation
  • Veo 3.1 Text to Video API — video generation

Microsoft Model APIs

  • MAI-Image-2.6 API — image generation

Alibaba Model APIs

  • Qwen-Audio 3.0 TTS Flash API — music generation
  • Qwen-Audio 3.0 TTS Plus API — music generation
  • Qwen Image 2.0 API — image generation
  • Qwen Image 2.0 Pro API — image generation
  • Qwen Image 3.0 Image to Image API — image generation
  • Qwen Image 3.0 Pro Image to Image API — image generation
  • Qwen Image 3.0 Pro API — image generation
  • Qwen Image 3.0 API — image generation
  • Wan 2.7 Image Text to Image API — image generation
  • Wan 2.7 Image to Video API — video generation
  • Wan 2.7 Text to Video API — video generation
  • Wan 3.0 Video API — video generation
  • Z-Image API — image generation

Recraft Model APIs

  • Recraft Remove Background API — image generation

ByteDance Model APIs

  • Seedance 2.0 API — video generation
  • Seedance 2.0 Fast API — video generation
  • Seedance 2.0 Mini API — video generation
  • Seedance 2.5 Image to Video API — video generation
  • Seedance 2.5 Reference to Video API — video generation
  • Seedance 2.5 Text to Video API — video generation
  • Seedream 4.5 Image to Image API — image generation
  • Seedream 4.5 Text to Image API — image generation
  • Seedream 5.0 Lite Image to Image API — image generation
  • Seedream 5.0 Lite Text to Image API — image generation
  • Seedream 5.0 Pro Image to Image API — image generation
  • Seedream 5.0 Pro Text to Image API — image generation