Adobe · Filed Mar 31, 2026 · Published Aug 6, 2026 · verified — real USPTO data

Adobe Patents a Tool That Turns 3D Models Into AI-Generated Images With a Text Description

Adobe is patenting a way to combine a 3D modeling window and a text prompt box in a single interface, so you can pose an object, describe what you want it to look like, and get an AI-generated image that matches both at once.

Adobe Patent: 3D Model + Text Prompt Image Generation — figure from US 2026/0228970 A1
Figure from the official USPTO publication.
See all 10 drawings from this filing ↓
Publication number US 2026/0228970 A1
Applicant ADOBE INC.
Filing date Mar 31, 2026
Publication date Aug 6, 2026
Inventors Matheus Gadelha, Tomasz Opasinski, Kevin Blackburn-Matzen, Mathieu Kevin Pascal Gaillard, Giorgio Gori, Radomir Mech
CPC classification 345/419
Grant likelihood Medium
Examiner CENTRAL, DOCKET (Art Unit OPAP)
Status Docketed New Case - Ready for Examination (Apr 29, 2026)
Parent application is a Continuation of 18451267 (filed 2023-08-17)
Document 20 claims

How Adobe's 3D-plus-text image generator works

Imagine you're designing a product shot. You have a 3D model of a sneaker, and you want to see it sitting on a wooden shelf, lit like a magazine ad, in a peach colorway. Normally you'd either render it in expensive 3D software or try to describe it from scratch to an AI image generator and hope for the best.

Adobe's patent describes a single screen that handles both steps together. On one side, you adjust the 3D model: rotate it, change its pose, or move it around. On the other side, you type a plain-English description of the scene and the look you want. The system then generates an image that respects both the shape you set up and the style you described.

The key idea is that the system reads the 3D geometry of your model to understand depth and form, then uses your text description to apply textures, lighting, and atmosphere on top. You stay in control of the structure, and the AI fills in the visual style.

How depth maps bridge the 3D model and the AI image

The patent describes a combined interface with four distinct panels working together:

  • A 3D modeling panel that displays a 3D object you can rotate and edit in real time
  • A 3D editing panel where you input changes to that model (move a limb, adjust proportions, change orientation)
  • A text prompt panel where you type a description of the scene and the surface appearance you want
  • A preview panel that shows the AI-generated output image, updating as you change either the model or the text

The underlying process starts by generating a depth map from the 3D model. A depth map is essentially a grayscale image that records how far each part of the object is from the camera, giving the AI a clear picture of the object's shape and volume without needing full rendering.

An image generation model (an AI trained to produce photorealistic or stylized images) then takes that depth map alongside the text prompt and produces an output image. The depth map acts as a structural guide, so the AI can't ignore the shape you set, while the text drives the color, texture, material, and environmental details.

The result is an image where structure comes from your 3D model and style comes from your words, with both updating interactively as you make changes.

What this means for designers using AI image tools

For designers and artists, the frustrating gap between 3D modeling tools and AI image generators has always been control. AI image generators are fast but unpredictable about shape; 3D renders are precise but slow and require expertise to make look good. Adobe's approach tries to close that gap by making the 3D model a hard constraint the AI has to respect.

If this shows up in something like Adobe Firefly or a future version of Dimension, it could let someone with basic 3D skills produce production-quality images without a full rendering pipeline. That's a meaningful shift for small studios, freelancers, and product teams who need convincing visuals fast.

Editorial take

This is a logical next step for Adobe Firefly, and it solves a real problem that anyone who has fought with AI image generators over object placement will recognize immediately. The depth-map approach is not a new idea in research circles, but putting it inside a live, interactive combined interface is where Adobe could actually make it practical. Worth watching.

There are more where this came from

We read every patent application Big Tech publishes and send you the ones worth knowing. Plain English, free, every week.

The drawings

10 drawing sheets from US 2026/0228970 A1 · click any drawing to enlarge

Patent filing page

Source. Full patent text and figures from the official USPTO publication PDF.