BC

BityClips

Prompt pack

All prompts

5 prompts · YouTube thumbnail

AI image prompts for YouTube thumbnails

You want thumbnail images that read at phone size and match the title's promise.

A thumbnail is judged at roughly 320 pixels wide on a phone, in under a second, next to eleven competitors. That constraint — not artistic quality — is what these prompts optimise for: one focal subject, extreme contrast, and negative space reserved for overlay text you add later in an editor rather than asking the image model to render.

Faceless channels that cannot use a reaction faceBatch-producing thumbnail variants for A/B testingCreators without a designer

The prompts

1. Single-subject thumbnail image

Midjourney, Leonardo AI, or DALL-E

The core thumbnail generation prompt

A YouTube thumbnail image: [SUBJECT], dramatic single-source lighting from [DIRECTION], deep shadows, highly saturated [COLOR] accent against a dark desaturated background, shallow depth of field, subject positioned on the [left/right] third leaving the opposite third as clean empty space for text overlay, no text, no watermark, no logo, sharp focus, high contrast, 16:9.

In practice: A YouTube thumbnail image: a single cracked shipping container isolated on a dock, dramatic single-source lighting from the left, deep shadows, highly saturated orange accent against a dark desaturated background, subject on the left third leaving the right third clean, no text, 16:9.

2. Concept generation before image generation

Claude or ChatGPT

Deciding what to render in the first place

My video title is: [TITLE]. The core tension is: [WHAT SURPRISES THE VIEWER].

Give me 6 thumbnail concepts. For each: the single focal image, the 2-4 words of overlay text, the emotional read in one word, and why it earns a click from someone who does not know my channel.

No concept may need more than one focal object. Reject anything that requires reading to understand.

In practice: Run this first — most bad thumbnails are bad concepts rendered well.

3. Before/after split thumbnail

Midjourney or Leonardo AI

Transformation, comparison, and versus videos

A YouTube thumbnail image: split composition, left half shows [BEFORE STATE] in cold desaturated blue-grey tones, right half shows [AFTER STATE] in warm saturated tones, hard vertical divide down the centre, dramatic lighting on both halves, no text, no arrows, 16:9, high contrast, sharp.

In practice: Add the arrow and the text in Canva or Photoshop afterwards — image models render both badly.

4. Thumbnail critique against competitors

Any vision-capable model (Claude, ChatGPT, Gemini)

Testing a thumbnail before you publish

Here is my thumbnail: [ATTACH IMAGE]. My title is [TITLE].

Critique it for: readability at 320px wide on a phone, whether the focal subject is instantly identifiable, whether the contrast survives a bright-room screen, whether the overlay text is redundant with the title, and whether the image matches the promise of the title.

Give me the single highest-impact change, not a list.

In practice: Ask for one change. A list of nine tweaks never gets applied.

5. Variant set for A/B testing

Claude or ChatGPT, then Midjourney

Producing testable alternatives, not random redesigns

Take this thumbnail concept: [DESCRIBE CONCEPT]. Produce 4 variants that each change exactly one variable: (1) accent colour, (2) subject scale — close crop versus wide, (3) emotional read — curiosity versus urgency, (4) background — busy versus empty.

Describe each as a complete image prompt I can paste into an image model.

In practice: One-variable-at-a-time is the only way a thumbnail test tells you anything.

Variables to fill in

  • [SUBJECT] — one object or one figure, never a scene
  • [COLOR] — pick an accent that is not in your niche's default palette
  • [TITLE] — the thumbnail must not repeat it word for word

What actually improves output

  • Reserve a third of the frame as empty space and add text in an editor. Image models still garble typography.
  • Test at 320px. If you cannot tell what it is at that size, the render quality is irrelevant.
  • Generate the concept in a text model and the image in an image model — asking one tool to do both produces mush.

Mistakes that ruin the result

  • Asking the image model to render the overlay text.
  • Two or three focal subjects competing for attention.
  • A thumbnail that repeats the title instead of adding to it.

Tools that finish the job

Where these prompts go next

A prompt produces the script, the shot list or the markup. These are the tools that turn that output into a finished video.

Midjourney

Leading AI image generation tool expanding into video with high aesthetic quality output.

Basic from $10/mo; Standard from $30/mo; Pro from $60/mo

View tool profile →

Leonardo AI

AI creative suite with video generation and motion tools.

Free; paid from $12/mo

View tool profile →

Predis.ai

AI-powered social media video creator that generates posts from product URLs and keywords.

Free (5 posts/mo); Lite $29/mo; Starter $59/mo; Agency $129/mo

View tool profile →

More prompt packs