Image to Prompt Generator

Upload a reference image. Get a prompt you can edit, copy, and reuse.

Upload image
Upload an image, or choose from your assets
Click to upload or choose an image
PNG / JPG / WEBP
Upload or choose image
Output format
Reference image
Generated prompt

A realistic vertical smartphone selfie of a young adult woman in her early 20s, shot indoors at night in a modern office or studio room. She is framed in a tight close-up from the upper chest upward, looking directly into the front camera with calm, slightly distant eyes and a soft closed-mouth expression. Her head is slightly tilted forward, and one hand is placed under her chin in a gentle “flower pose,” palm supporting the jawline. The other arm extends toward the camera, creating subtle selfie perspective distortion and a casual, intimate feeling. She has long, glossy black hair with a soft center part, smooth straight lengths falling over both shoulders, with a few natural flyaway strands catching the overhead light. Her makeup follows a refined contemporary East Asian / K-pop-inspired beauty style: clean fair neutral-cool skin, softly blurred base makeup, defined but delicate brows, long curled lashes, subtle winged eyeliner, faint lower-eye shimmer, grey-brown contact-lens-like eyes, soft pink blush, and glossy rose-pink lips. Preserve the original face feeling: oval face, large reflective eyes, slim nose bridge, small softly shaped lips, delicate chin. Visible skin includes the full face, neck area, both forearms, part of the upper arms due to the short sleeves, and the hand touching the chin. No waist, legs, hips, back, or chest exposure is visible. The skin tone is fair with a neutral-cool cast under cool indoor lighting. The skin should look smooth, hydrated, and softly satin to the touch, like lightly powdered but still supple skin, with gentle highlights on the nose bridge, cheekbones, forehead, and the back of the hand. She wears a fitted black short-sleeve T-shirt and silver accessories: thin hoop earrings, a chunky silver chain necklace, metallic rings, and a silver bracelet or cuff near the lower edge of the frame. The background is a real nighttime office interior: large dark windows with faint city lights outside, white ceiling panels, cool fluorescent or LED ceiling reflections, a desk with papers on the left, and a light wood cabinet or podium on the right. The lighting is cool and physical: overhead rectangular office lights create bright streaks on the top of her hair, while soft front-facing smartphone/room light fills her face evenly. The background remains visible but slightly less emphasized. The image should feel like it was taken late at night after work, quiet and slightly tired but composed, as if she had paused for one polished selfie before leaving. Keep the framing slightly imperfect and human: mild high-angle selfie perspective, close lens distance, natural facial asymmetry, tiny hair flyaways, real office details, slight smartphone smoothing, and subtle digital noise. Avoid making it look like a studio fashion shoot, an AI-perfect portrait, or an overly airbrushed commercial beauty image.

What does the tool read from your image?

It looks beyond the main subject and captures the style, lighting, composition, camera feel, and small visual details that shape the final image.

Subject and scene

Finds the main subject, setting, objects, action, and visible relationships.

Style and medium

Names the visual direction, such as photo, cinematic, 3D, anime, editorial, product shot, or illustration.

Lighting and color

Captures light direction, contrast, palette, shadows, and the overall mood.

Composition and camera

Turns framing, angle, lens feel, depth of field, perspective, and crop into prompt language.

Details and mood

Adds texture, clothing, materials, environment details, emotion, and finish.

Ready-to-edit prompt

Gives you a prompt you can use as a base for Midjourney, GPT image, Nano Banana, and other generators.

Promptsref advantageMore than a one-line caption

The tool reads the image like a prompt brief, then gives you both a detailed JSON prompt and a shorter natural-language version.

01

Structured JSON prompt

The JSON version breaks the image into subject, pose, outfit, lighting, camera, background, mood, constraints, and negative prompts. That gives the model clearer instructions for rebuilding the look.

02

Short prompt included

You also get a natural-language prompt. It is shorter, easier to copy, and useful when you want to generate quickly without editing a long JSON block.

03

Output in your language

Choose the output language before analysis. The prompt keeps the visual details, but reads naturally in the language you use for work.

Reference portrait paired with a shortened structured JSON prompt

How to use the image to prompt generator

01 · Reference image

Archival-style reference photograph of an elderly shopkeeper with his dog

02 · Turn image to prompt

A vertical, documentary-style street photograph with a strong early-20th-century atmosphere: an elderly white-bearded man stands in the narrow doorway of an old dark shop or pub, accompanied by a small scruffy cream-colored terrier-like dog. The man is centered but slightly off-axis, leaning casually against the doorframe with one leg crossed over the other and his weight settled into the threshold. He wears a weathered brown tweed flat cap, a dark charcoal-gray overcoat hanging open, a textured gray herringbone waistcoat, a muted rust-brown tie, loose high-waisted gray trousers, and worn dark leather shoes. His left hand rests inside a trouser pocket while his right hand raises a short smoking pipe to his mouth. His expression is quiet, stern, and contemplative, with narrowed eyes directed toward the camera and a face marked by deep age lines, a full white mustache, and a neatly heavy white beard. Smoke curls softly from the pipe toward the shadowed doorway. The dog sits in the lower-left foreground, looking upward toward the man with an alert, affectionate expression. Its fur is shaggy, wiry, and uneven, pale ivory with slightly darker beige patches around the ears and muzzle; the fur should appear tactile, dusty, and naturally tousled rather than groomed. The man’s face and hands show dry, weathered skin with visible wrinkles, creases, and realistic texture, softly touched by warm reflected light. No other body skin is visible. Frame the scene as a close full-body portrait from a slightly low, eye-level-to-chest-height viewpoint, as if taken by a photographer standing just outside the doorway with a vintage 35mm camera and a normal-to-slightly-wide lens around 40mm. Use a vertical 4:5 composition. Keep the man dominant in the center, the dog anchoring the lower-left corner, and the doorway creating strong dark vertical framing on both sides. Allow slight asymmetry, imperfect centering, subtle lens distortion, a little foreground obstruction, and the feeling of a found archival photograph rather than a polished studio portrait. The old shopfront is cramped and heavily worn: dark brown and black wooden doorframes, chipped paint, stained surfaces, faded hand-painted signs, shelves of bottles, stacked papers, tins, and small miscellaneous goods visible along the edges. The interior behind the man is almost black, creating a deep recess that makes his face, clothing, pipe smoke, and the dog stand out. Use warm, low-contrast daylight mixed with amber reflections from aged wood and nearby shop surfaces. A soft directional light enters from the front-left, grazing the man’s beard, cap, coat shoulders, waistcoat texture, and the dog’s fur while leaving the doorway behind him in dense shadow. Add faint cool ambient fill from the street, subtle rim light on the coat edge, and naturally uneven exposure across the frame. The light should behave physically: highlights catch on worn fabric fibers, the pipe, dusty bottles, and rough wood; shadows remain brown-black and slightly lifted rather than pure digital black. Use shallow-to-moderate depth of field, keeping the man’s face, pipe, waistcoat, and dog recognizable while allowing the shop edges and background objects to soften gradually. Create a tactile analog image with muted sepia, tobacco brown, olive gray, charcoal, faded cream, and subdued rust tones. Include fine film grain, slight halation around bright edges, soft focus falloff, minor dust specks, faint scratches, gentle vignette, restrained color fading, and the imperfect tonal softness of an aged color photograph or hand-tinted archival print. Avoid a modern digital-clean appearance, fashion-editorial polish, exaggerated sharpness, glamour retouching, plastic skin, contemporary clothing, excessive smoke, dramatic cinematic lighting, or a pristine background. The image should feel like a fleeting moment that happened decades ago and was discovered later: a tired shopkeeper pausing in the doorway, his loyal dog waiting beside him, while the street continues just outside the frame. Caption energy: “He had been standing there all afternoon, pipe glowing quietly, the dog watching the world go by.”

03 · Generated image

Generated archival-style portrait of an elderly shopkeeper with a terrier
📱

Upload or choose the image you want to analyze

Wait a few seconds for the prompt

😊

Copy the prompt or edit it before using it

When to use image to prompt

01

Break down the style of a reference image.

02

Turn a portrait, product shot, poster, or illustration into a reusable prompt.

03

Move visual inspiration into GPT image, Nano Banana, Midjourney, or another model.

04

Learn how good image prompts describe subject, style, lighting, and composition.

05

Create prompt variations without losing the original visual direction.

Reference street photo paired with a short natural-language prompt

Use the prompt with your image model

Keep the visual description, then add the aspect ratio, style syntax, or parameters your model needs.

GPT image 2.0

Use detailed natural-language prompts for subject, pose, camera, light, and finish.

Nano banana

Works well for reference recreation, visual remixing, and image-aware prompts.

Midjourney

Useful for style matching, cinematic references, SREF exploration, and concept images.

Other AI image models

Use the extracted prompt as a base, then add the syntax your generator expects.

Reference anime artwork paired with a Chinese-language prompt

The Image-to-Prompt tool converts images into descriptive prompts. It uses artificial intelligence to analyze image content and generate corresponding prompts, which can be used for creating similar images or other text-to-image generation tasks. This tool helps users better understand and describe image content, facilitating AI image generation and creative work.

Your uploaded image is securely stored so you can access the generated prompt and your history. It remains private to your account and is visible only to you. It is not displayed in the public gallery or shared with other users through Image to Prompt.

Image-to-Prompt plays a crucial role in AI image generation by helping users extract key information from existing images and generate accurate prompts. These prompts can serve as new descriptions for creating similar or improved images, helping users better understand how AI "sees" images and improve their prompt-writing skills.

Image-to-Prompt output typically includes descriptions of the main subject, setting, artistic style, color scheme, lighting conditions, and overall mood or atmosphere. It may also describe specific subjects, actions, poses, composition, perspective, coloring, style, or detailed features present in the image, aiming to provide a comprehensive textual representation.

Using Image-to-Prompt is simple: just upload an image to our tool. Our AI will analyze the image and generate detailed prompts. You can then use this prompt as a starting point for creating new images with AI image generators or refine it further based on your specific needs.

Our Image-to-Prompt generator uses advanced AI models to provide highly accurate image descriptions. However, like all AI technologies, it may occasionally misinterpret certain elements. We continuously work to improve its accuracy and encourage users to review and adjust generated prompts as needed for their specific purposes.

Our Image-to-Prompt generator is designed to handle various types of images, including photographs, digital art, and illustrations. However, it performs best with clear, high-resolution images. Some highly abstract or complex images may result in more generalized descriptions.

Using Image-to-Prompt helps you gain insight into how AI interprets visual information. This helps you understand what elements to include in your own prompts when generating images and serves as an excellent tool for learning the vocabulary and structure of effective image prompts.

No, the Image-to-Prompt generator does not allow any NSFW content. Uploading NSFW content will be blocked by our system.