Tutorial Guide

How to Convert an Image to a 3D Model with Meshy (9 Steps)

Convert any image to a 3D model in 2026 with this step-by-step guide. Get a textured, export-ready model in ~90 seconds — no 3D skills, free to try.

AI image-to-3D does exactly what the name says: you hand it a single photo, and it gives you back a finished 3D model — already textured and ready to use. No 3D modeling skills, no special camera, no studio shoot. This guide walks you through the whole workflow with Meshy: how it works, how to prep your input, every generation setting, export formats, and where teams put it into production.

Quick Facts

Generation time~90 sec (Preview) · ~3 min (Refine)
Export formatsGLB · FBX · OBJ · USDZ · STL
Input formatsPNG · JPG · WebP · ≥1024px recommended
Auto-texturesPBR: base color, normal, metallic, roughness
Free tier200 credits · no credit card
APIREST · Python & Node SDK

What Is AI Image-to-3D?

AI image-to-3D does exactly what the name says: you hand it a single photo, and it gives you back a finished 3D model — already textured and ready to use.

But how does it build a 3D object from one flat photo? The AI uses the same kind of technology behind image generators like Midjourney (a diffusion model) to read your photo's shape, materials, and depth. Then comes the clever part, called multi-view synthesis: it imagines what your object looks like from all the angles the photo never captured, and reconstructs the sides, back, and underside from there. The result is a complete 3D model with realistic textures, ready in as little as 90 seconds.

A few years ago, this kind of output looked rough and experimental. As of 2026, it's good enough for real production work — and Meshy leads the way. Here's what that looks like once you actually use it:

  • Clean mesh, ready to use — the model's underlying geometry (its topology) comes out tidy in either quad or triangle form, so it drops straight into Unity, Unreal, or Blender with no manual cleanup
  • Realistic textures, done for you — Meshy auto-generates the color and surface detail (the PBR maps), so your model looks right under any lighting
  • Change the look without starting over — don't like the texture? Retexture keeps the same shape and just repaints it, saving time and credits
  • Export to almost anything — GLB, FBX, OBJ, USDZ, and STL, covering web, games, AR, and 3D printing
  • Built for automation too — every feature is also available through an API if you want to run it at scale

How to Get the Best Image-to-3D Results

Your input photo is the single biggest factor in how good your 3D model turns out — get the photo right, and most other settings fall into place.

Step 1: Preparing Your Input Image

Five things matter most:

  • Shoot from the front or a ¾ angle — gives the AI enough depth cues to rebuild the sides and back
  • Use a plain or white background — keeps the focus on your subject, not its surroundings
  • Fill the frame — your subject should take up most of the photo; small subjects lose detail
  • Light it evenly — harsh shadows get baked into the texture as fake surface detail
  • One object at a time — the model reconstructs a single subject, not a whole scene

Technical requirements: PNG, JPG, or WebP · at least 1024px on the short side (2048px+ for Refine mode) · transparent PNG works great.

Step 2: Generation Tips That Actually Work

Photo ready? These settings and habits make the biggest difference:

  • Turn on Image Enhancement — it pre-processes your photo before generation. Keep it on for real-world photos; turn it off for clean renders or illustrations.
  • Add a text prompt too — the image handles shape, the prompt handles surface detail. Describe the material, e.g. "worn leather, brown, matte."
  • Retexture instead of regenerating — geometry looks right but the texture's off? Hit Retexture and tweak the prompt. Same mesh, fewer credits.
  • Know when to switch to Text-to-3D — no reference photo, or designing something that doesn't exist yet? Text-to-3D gives you more creative control.
  • Handling tricky materials — reflective surfaces (chrome, glass): shoot against a dark matte background. Translucent objects: a light dusting of chalk spray helps. Hair or fur: skip the photo and use Text-to-3D with a detailed prompt.

Step-by-Step: Convert an Image to a 3D Model with Meshy

Meshy Image to 3D step-by-step workflow

Step 1: Upload your image

Go to Meshy Image to 3D and click Upload Image. Supported formats: PNG, JPG, or WebP. For best results, use ≥ 1024px on the shortest side — 2048px or higher is recommended if you plan to use Refine mode.

Got more than one image? Use Batch Image to 3D to generate multiple models in a single run — same settings applied across the whole batch.

Step 2: Choose your generation mode

After uploading, select how you want the 3D mesh to be structured:

Standard — the default. Generates a high-poly mesh with full detail — best for rendering, visualization, or when you plan to optimize the topology yourself in tools like Blender.

Low Poly (Beta) — generates a game-ready, optimized mesh with clean topology and reduced polygon count straight out of the box. Ideal for:

  • Mobile or indie game assets
  • Direct import into game engines (Unity, Unreal)
  • Workflows where you need a lean mesh without manual cleanup

Note: Low Poly is currently in Beta — it works well for simple/medium-complexity objects, but may produce artifacts on highly complex shapes. The team is actively improving it.

Recommendation: Start with Standard if you're unsure. Switch to Low Poly when you need game-ready output.

Step 3: Select your AI model

Below the generation mode, choose your AI Model — this controls which algorithm processes your image.

Meshy 6 is the default and recommended for almost all use cases: cleanest geometry, best texture fidelity, and most accurate reconstruction from a single photo. Older models (Meshy 5, Meshy 4, etc.) are available for compatibility with existing pipelines — but unless you have a specific reason, stay on Meshy 6.

Step 4: (Optional) Image Enhancement, Multi-View & Pose

Image Enhancement is a toggle below the model selector. When on, Meshy pre-processes your photo before generation — useful for real-world photos with noise, harsh lighting, or a busy background. Turn it off for clean CGI renders or product shots that are already optimized.

After uploading, click Generate Multi-view and Meshy predicts three extra angles — the sides, back, and ¾ views — from your single photo. The 3D generation then uses all four images together, giving the model far better coverage of surfaces your original shot never captured.

It's worth the extra ~30–60 seconds for most real-world objects — especially complex subjects like furniture, characters, or vehicles, or anything with a back or underside the photo doesn't show clearly. You can skip it for simple shapes that read fully from one angle (flat icons, coins, pendants), or when the single-image result is already good enough.

If your image contains a character, enable the Pose toggle to force the output into a standardized rigging pose — A-Pose (arms at ~45°, recommended for most game and animation pipelines) or T-Pose (arms horizontal, standard for some auto-riggers). Skip this for non-character objects, or if you just need a display model and don't plan to rig or animate.

Step 5: Choose the proper license

After your model is generated, you can set its License to control how others can use it if you choose to share it publicly. Options typically include personal/non-commercial use and full commercial rights. Note: Enterprise plan users will not see the Change License button — this is expected behavior, not a bug. If your model is for internal use only and you're not publishing it, you can leave this at the default.

Step 6: Generate & review your white model

Click Generate. Meshy first builds a white model — geometry only, no texture — as your draft (~1–2 minutes; Low Poly is similar or a bit faster). You'll see a single progress bar; the draft, remesh, and texture passes all run in sequence behind it.

When the white model appears, rotate it in the 3D viewer and check the silhouette, proportions, and any obvious artifacts. If the shape is off, swap in a different photo now — before you spend credits on texturing. If it looks right, move on.

You can also download the white model as GLB at this stage — handy when you only need geometry for 3D printing, or plan to texture it manually in Blender.

Meshy white model preview in the 3D viewer

Step 7: Generate texture

Click Generate Texture. Meshy runs UV unwrapping, multi-view diffusion, back-projection, and super-resolution — all automatically.

Estimated time: 1–6 minutes (queue length affects this).

Optional before generating: add a text prompt to guide the surface material — e.g. "worn leather, brown, matte." The image handles shape; the prompt handles surface detail.

Generating PBR textures on a Meshy model

Step 8: Retexture & export

Happy with the geometry but not the texture? Don't start over — click Retexture on the finished model, tweak your text prompt, and Meshy re-runs only the texture pass: same mesh, lower credit cost, faster turnaround. It's the quickest way to dial in material quality.

Once you're satisfied, click Download and pick your export format.

Exporting a Meshy model in GLB, FBX, OBJ, USDZ, or STL

Step 9: Bring it into the real world

That's it — you've turned a single photo into a finished 3D asset. From here, send the STL to a 3D printer (via Cura or PrusaSlicer) to hold a physical copy in your hands, or drop the GLB or USDZ into a VR/AR viewer to walk around your model at real-world scale.

Viewing the finished 3D model in AR

What Export Formats Does Meshy Support for Image to 3D?

Meshy exports image-to-3D models in five formats: GLB, FBX, OBJ, USDZ, and STL. Pick based on where the model is going — not personal preference.

FormatBest forNotes
GLBWeb embeds, AR, SketchfabSingle binary file — textures included. Best default for sharing
FBXUnity, Unreal Engine, MayaStandard for game pipelines; preserves rig if auto-rigging was enabled
OBJBlender, general 3D softwareExports with a separate .mtl texture file; widely compatible
USDZiOS AR (Quick Look, Reality Composer)Apple's native AR format; drag into Xcode or share directly to iPhone
STL3D printingGeometry only — no textures. Most slicers (Cura, PrusaSlicer) accept it directly

Not sure? GLB is the safest default — it works everywhere and keeps textures bundled. For a deeper comparison, see our guide to 3D file formats.

Use Cases of Image-to-3D Model Generation

AI image-to-3D isn't a toy — here's how teams across industries are putting it into production today.

Game & XR Assets (Unity / Unreal)

Photograph a prop, export it as FBX, and drop it straight into your engine. Face count is controllable at export, so you can hit the polygon budget for each LOD tier — and the same asset works in VR and AR scenes too. See the Unity workflow guide for a full import walkthrough.

E-Commerce & Product Visualization

Shoot your product on a plain background, generate a textured model, and embed it on the product page with model-viewer (GLB). Shoppers rotate and inspect the item right in the browser — and the same file powers an AR "view in your room" preview — with no 3D studio involved.

3D Printing & Physical Products

Upload a clear photo, export as STL, and load it into Cura or PrusaSlicer. Meshy outputs clean, manifold geometry — just check the scale before slicing. Perfect for custom miniatures, replacement parts, prototypes, and board-game pieces. For print-ready settings, see Photo to Printable Model.

Marketing, Ads & Social AR

Turn a single product shot into a 3D model you can spin in hero animations, motion-graphics ads, and AR effects on Instagram, Snapchat, or TikTok — refresh creative across every format without re-shooting.

Education, Museums & Digital Archiving

Photograph an artifact, specimen, or classroom object and turn it into an interactive model students can rotate and zoom — no scanning rig or photogrammetry setup. Ideal for digitizing collections, building virtual exhibits, and preserving cultural heritage.

Why Meshy for Image to 3D?

  • End-to-end in one tool — upload, generate, retexture, and export without switching apps
  • PBR textures out of the box — base color, normal, metallic, and roughness auto-generated, no manual baking
  • Flexible topology — choose quad or triangle mesh at export, with face-count control for any pipeline
  • API-first — every workflow available via REST API; see the API quickstart

Frequently Asked Questions

Is Meshy free for image to 3D?

Yes. Meshy offers a free tier with 200 credits on sign-up — enough to generate several models and explore the full workflow. Paid plans (Pro, Studio, Enterprise) unlock higher monthly limits, faster queues, and commercial licensing.

What AI technology powers Meshy?

Meshy is built on cutting-edge generative AI models. You can explore the broader AI/ML ecosystem on Hugging Face, where many foundational models used in 3D generation research are published.

How long does generation take?

Preview mode generates in ~30 seconds. Refine mode takes 2–4 minutes and produces higher-fidelity geometry and textures. Start with Preview to validate shape, then switch to Refine for your final asset.

Do generated models come with textures?

Yes — when PBR is enabled, Meshy auto-generates four texture maps: base color, normal, metallic, and roughness. These work out of the box in Unity, Unreal, Blender, and any PBR-capable renderer.

Can I use Meshy models commercially?

Yes, on paid plans (Pro and above). Free-tier models are for personal and non-commercial use. See the Terms of Service for full licensing details.

Can I edit the model after generation?

Yes. Export as FBX or OBJ and open in Blender or Maya for retopology, UV editing, or rigging. To update textures without rebuilding the mesh, use Meshy's built-in Retexture feature.

What's the difference between image-to-3D and text-to-3D?

Image-to-3D reconstructs from a reference photo — best when you need to match a specific real-world object. Text-to-3D generates from a written description — better for original characters, fantasy props, or anything that doesn't exist yet.

Which objects are hardest to convert?

Highly reflective surfaces (chrome, glass), transparent objects, and hair/fur. For reflective items, shoot against a dark matte background. For hair and fur, a detailed Text-to-3D prompt usually outperforms a photo.

Next Steps & Related Tutorials

Ready to go deeper? These guides continue where this one leaves off:

  • Text to 3D Model Tutorial — No reference photo? Start from a written description instead. Covers prompt engineering tips for better geometry and materials.
  • Photo to 3D Printable Model (STL) — Optimizing your output specifically for 3D printing: manifold checks, scale, and slicer-ready export.
  • 3D Model for Unity Workflow — Import your Meshy model into game engines with correct scale, coordinate system, format selection, and material assignments.
  • Character Auto-Rigging Workflow — Turn a character model into an animation-ready asset using Meshy's built-in auto-rig, then export FBX for downstream use.
  • PBR Texturing with Meshy — A deep dive into PBR maps: what base color, normal, metallic, and roughness each do — and how to use Retexture effectively.
  • Export to Blender Workflow — Bring your Meshy model into Blender via the official plugin for retopology, UV cleanup, and production-quality rendering.
  • API Quickstart: Image to 3D — Automate the entire pipeline via REST API. Includes authentication, request examples, and webhook setup.

Related Guides

3D, On Command

Contact Sales