HiAPI
OverviewModels MarketAPI KeysUsage StatisticsCall LogsBillingReferralPlaygroundStorageChangelogContact UsSettings
Display unit
N
Powered by hiapi
Settings

Welcome

MAI-Image-2.6 is Microsoft’s next-generation text-to-image and image editing model—built for higher visual quality, controllable edits, multi-reference workflows, and web grounding across ads, e-commerce product shots, and design.

Provider: Microsoft

Category: image generation

Status: In progress

Back to Models

MAI-Image-2.6

Coming SoonComing soon
by MicrosoftImage

MAI-Image-2.6 is Microsoft’s next-generation text-to-image and image editing model—built for higher visual quality, controllable edits, multi-reference workflows, and web grounding across ads, e-commerce product shots, and design.

Integration in progress

MAI-Image-2.6 is coming to HiAPI

What is MAI-Image-2.6: Microsoft text-to-image and image editing

MAI-Image-2.6 is Microsoft’s next-generation text-to-image and image editing model for creators who want higher visual quality, more controllable local edits, and flexible composition and style control. It covers both greenfield generation from text and iterative polish on existing frames—so product references, poster copy, and scene concepts move closer to design-ready directions.

MAI-Image-2.6 doubles down on detail and quality, controllable editing, multi-reference workflows, web grounding, and dynamic aspect ratios. Ads, commerce, and design teams can shape hero tests, product-image iteration, poster layouts, and location mood boards around these strengths.

Whether you start from a written brief or already have a product shot or poster to revise—background, objects, or in-image text—MAI-Image-2.6 offers clear text-to-image and editing workflows. Lock placement goals, reference roles, and acceptance criteria early so generation and edits land closer to review-ready, campaign-ready AI images.

Text-to-image vs image editing: two creative paths

Text-to-image fits briefs written from scratch; image editing revises product shots or posters—background, lettering, and local details. Multi-reference and web grounding can sit in the same design workflow.

Contact Us
Your jobProductMain inputOutput
Text-to-image from a briefText to imageSubject, setting, light, aspect ratio, and purposeA text-to-image direction ready to refine
Image editing on an existing designImage editingSource image plus clear keep vs. change notesA revision that keeps the original structure
Multi-reference subject, style, and sceneMulti-referenceReferences with clear subject, style, and scene rolesOne composition informed by each reference
Web grounding for real-world contextWeb groundingPlace or topic plus viewpoint, season, and weatherA scene illustration backed by context

Text to image

Text-to-image from a brief

Input
Subject, setting, light, aspect ratio, and purpose
Output
A text-to-image direction ready to refine

Image editing

Image editing on an existing design

Input
Source image plus clear keep vs. change notes
Output
A revision that keeps the original structure

Multi-reference

Multi-reference subject, style, and scene

Input
References with clear subject, style, and scene roles
Output
One composition informed by each reference

Web grounding

Web grounding for real-world context

Input
Place or topic plus viewpoint, season, and weather
Output
A scene illustration backed by context

MAI-Image-2.6 creative directions and example frames

The flagship focuses on detailed generation and iterative edits; Flash in the same family targets faster batch exploration. Both support multi-reference, web grounding, and flexible compositions.

A yellow and blue sneaker on a desert rock from Microsoft’s MAI-Image-2.6 page
Notice the contrast between the textured sole, rock surface, and strong daylight.

MAI-Image-2.6

Detailed text-to-image and product-grade edits

The flagship path joins text-led creation with iterative editing: start from a product reference, refine materials, backgrounds, or objects, and keep silhouette and spatial relationships stable across rounds. Sole texture, metal reflections, and packaging type are the details worth zooming. Flash in the same family helps explore directions faster before flagship polish.

Best for: Commerce heroes, brand mood boards, and detailed posters

A Madrid Metrópolis Building scene published by Microsoft
The building anchors the composition while street and sky establish depth.

Location and context

Location scenes with web-grounded context

Web grounding can bring relevant online context into image creation. For landmarks or destinations, specify viewpoint, season, and weather so the information serves a deliberate scene—not a pile of unrelated facts.

Best for: Destination concepts, architectural mood boards, and editorial illustration

From idea to MAI-Image-2.6 finished-frame workflow

  1. 1

    Align finished-frame goals with model strengths

    Map MAI-Image-2.6 text-to-image, image editing, quality-controlled edits, multi-reference, and web grounding to your ad, commerce, or design goals.

  2. 2

    Choose text-to-image or image editing

    Describe subject, environment, and light for a new scene. If the design exists, start from the image and name only the changes—plus matching references.

  3. 3

    Lock reusable visual language

    Organize prompts, framing notes, and reference roles into copyable blocks—reusable AI image assets for ads and commerce.

  4. 4

    Iterate detail-by-detail to review-ready frames

    Nail key composition first, then revise objects, color, lettering, or background. Use appearance locks and aspect-ratio templates until ad heroes or commerce shots hit review-ready quality.

Creative assets worth locking before you generate

Placement briefs, reference roles, and reusable prompt blocks become high-quality production inputs in MAI-Image-2.6 workflows.

01

Name placement and audience

State whether you need a product hero, event poster, or scene concept, then note channel aspect ratio and intended feeling.

02

Gather references and assign roles

Choose subject, color, and style references and assign each a role. For revisions, list the elements that must remain.

03

Decide how you will accept a frame

Pick three priorities such as product silhouette, readable lettering, and negative space. Focus each revision on one priority.

04

Build reusable prompt templates

Turn appearance locks, lighting phrasing, aspect-ratio specs, and keep-lists into templates so series ads and commerce heroes iterate by swapping variables.

MAI-Image-2.6 capabilities for design-ready AI images

Text-to-image + image editing dual path

Use text-to-image when ideating from scratch; lead with image editing when you already have a product shot or poster. Both paths cover concept tests and controllable polish—core creative modes in MAI-Image-2.6.

Quality detail and controllable local edits

A top capability for ads and commerce: replace or remove objects and update lettering on a higher-quality base. Name the edit target, then state what composition, light, or position must stay so iterative edits remain controllable.

Multi-reference with clear roles

People, products, styles, and settings can come from different assets. Label which reference supplies subject, palette, or environment so series visuals and brand consistency stay on track.

Dynamic aspect ratios and web grounding

Compose within the model card’s ~2,359,296-pixel budget (about 1536×1536). Plan subject and negative space separately for banner vs. portrait; location scenes can use web grounding for environmental cues in ads and design pitches.

Ads, commerce, design: where MAI-Image-2.6 shines

E-commerce product and campaign heroes

Commerce teams can pair product references with campaign keywords, explore studio, outdoor, and story-led settings, then lock a composition for the PDP hero.

Ad poster and packaging layouts

Put headline, supporting copy, and brand name in the brief, then compare thumbnail reading order so text-to-image and controllable edits balance legibility and hierarchy.

Series visuals with consistent style

Content teams can lock palette, lighting, and background language while swapping subjects—using multi-reference to keep a shared series cue.

Location scenes and design mood boards

Creative teams can describe place, time, viewpoint, and weather—optionally with web grounding—to produce discussable destination or architecture frames for early boards and pitches.

MAI-Image-2.6 text-to-image & editing prompt examples

A product hero brief

Create a product campaign concept for a matte cobalt-blue travel bottle. Place it on pale limestone in a quiet studio. Use soft light from the upper left, a close three-quarter view, and generous negative space on the right for a headline. Keep the bottle silhouette clean and its surface texture visible.

The brief connects product, materials, lighting, viewpoint, and headline space—ideal for a first text-to-image pass before polish.

An existing poster revision

Edit the supplied poster. Replace the headline with “WEEKEND IN BLOOM”. Keep the main illustration, background color, and headline position. Use compact, clearly separated letters with strong contrast. Remove the small decorative mark in the lower-right corner and leave that area open.

Separate changes from elements to preserve so each image-editing round is easy to accept or reject.

Explore MAI-Image-2.6

Get ready for text-to-image and image editing workflows

Build around MAI-Image-2.6’s quality, controllable edits, multi-reference, and web grounding—lock placement briefs, references, and prompt templates so ad heroes, commerce shots, and design scenes are ready to generate.

Sign up for 200 Credits

Use your trial credits on image models that are already available.

Explore image modelsSign up free for credits
MAI-Image-2.6 AI image generation: text-to-image and editing illustration
Integration in progress

MAI-Image-2.6 is coming to HiAPI

Related image models to explore alongside

GPT Image 2 model preview

GPT Image 2

Live text-to-image for posters, illustrations, and product concepts.

Nano Banana 2 model preview

Nano Banana 2

Explore new compositions and edits from text or reference images.

MAI-Image-2.6 FAQ

What is MAI-Image-2.6?

MAI-Image-2.6 is Microsoft’s next-generation text-to-image and image editing model—higher visual quality, controllable edits, multi-reference workflows, and web grounding for ad heroes, e-commerce product shots, posters, and design.

What is MAI-Image-2.6 good for?

E-commerce product and campaign heroes, ad posters and packaging layouts, series visual extensions, and location scenes or design mood boards that benefit from web grounding.

Text-to-image or image editing?

Use text-to-image when starting from scratch; use editing when you already have a product shot or poster to revise. Many teams generate a direction first, then iterate with edits.

Can I generate with MAI-Image-2.6 today?

This page is a model introduction and workflow prep guide—generation availability follows the page status. You can review capabilities and use cases now, and prepare placement briefs, reference roles, and acceptance criteria ahead of time.

When are multiple references useful?

They help when subject, style, and setting come from different assets. Explain which part of the desired image each reference should inform.

How do MAI-Image-2.6 and Flash differ?

The flagship focuses on quality and detailed control; Flash targets latency-sensitive, higher-throughput exploration. Compare them with the same brief.

Do dynamic aspect ratios cap the long side at 1536?

The model card limits total pixels (~2,359,296). One side may exceed 1536 within that total; available formats depend on the product surface.

How should I describe a product-image revision?

List shape, colors, and markings to preserve, then describe background, placement, and lighting changes. State when another reference supplies only style or setting.