I Stopped Writing AI Prompts and Started Uploading Screenshots Instead

· Source: Artificial Intelligence on Medium · Field: Technology & Digital — Artificial Intelligence & Machine Learning · Depth: Intermediate, quick

Summary

Using screenshots as direct visual inputs significantly enhances AI image generation, allowing users to achieve precise aesthetic outcomes more efficiently than with traditional text prompts. The author found that uploading a grainy YouTube screenshot, rather than describing its mood, composition, and palette, yielded a "startlingly close" movie poster in one attempt. This contrasts sharply with eight attempts using pure text prompts that still looked "off." The underlying mechanism is that AI models directly extend existing visual information, such as "ominous red lighting," instead of attempting to imagine a mood from abstract adjectives, thereby preserving composition, palette, and grain.

Key takeaway

For creative technologists or designers struggling with AI image generation, you should prioritize visual inputs like screenshots over lengthy text descriptions to achieve precise aesthetic results. This approach bypasses the AI's tendency to generate generic outputs from abstract adjectives, directly conveying mood, composition, and palette. By feeding the AI an image, you can significantly reduce iteration time and improve output fidelity, moving from multiple attempts to a single, more accurate generation.

Key insights

Uploading images directly to AI image generators bypasses the limitations of text prompts for conveying visual mood and composition.

Principles

Method

Upload a cropped screenshot or reference image to an AI image generator, then provide minimal text instructions for modifications like text swaps or aspect ratio adjustments.

In practice

Topics

Best for: AI Student, Creative Technologist

Related on AIssential

Open in AIssential →

Editorial summary, takeaway, and curation by AIssential. Original article published by Artificial Intelligence on Medium.