Tutorial ai image generation ·27 min ·Recorded Jul 2026

How to Make AI Product Photography That Doesn't Look Fake (Nano Banana Pro & GPT Image 2)

Jamey Gannon presents a workshop for e-commerce teams on generating AI product photography at scale using multimodal models (Nano Banana Pro, GPT Image 2) inside Flora and Higgsfield. She walks through four techniques: creating product shots from just a label design, converting iPhone photos into studio-quality shots, generating stylized editorial imagery via style transfer, and producing AI UGC with product insertion. Core principles include using multimodal (not diffusion) models for precision, working in discrete steps, the SWORD prompting framework, and reusable "prompt boosters" for maintaining product consistency.

What's discussed, in order

3 named frameworks

01 Starting Conditions / Inputs
The three categories of inputs that shape an AI image generation.
presenter's own · ~06:22Play
02 SWORD Framework
A mnemonic for structuring detailed prompts for AI image generation.
presenter's own · ~09:17Play
03 Prompt Booster
A reusable pre-written paragraph of directives appended to prompts to enforce text, logo, geometry, and color consistency.
presenter's own · ~11:17Play

What's actually believed — in their own words

Job titles are really starting to blur.

Jamey Gannon · 2026 · observation 00:47 #

With label-based mockup generation, you can create product images before the product is in production, enabling paid ad testing and pre-sales.

Jamey Gannon · 2026 · opinion 02:16 #

The demonstrated workflow is tool-agnostic; any tool with access to GPT Image 2 or Nano Banana Pro will work.

Jamey Gannon · 2026 · observation 02:45 #

Multimodal AI models can "think and understand what you're asking for," unlike diffusion models.

Jamey Gannon · 2026 · observation 05:52 #

Diffusion models like Midjourney and Flux are optimized for aesthetics rather than subject/text consistency.

Jamey Gannon · 2026 · observation 06:00 #

AI is not yet good enough to handle complex product photography tasks in a single prompt.

Jamey Gannon · 2026 · opinion 07:22 #

Asking a model to do many things at once lowers prompt adherence.

Jamey Gannon · 2026 · observation 07:55 #

Saying "product photography shot" is a shortcut that leverages the model's training to produce centered, high-quality, photorealistic framing.

Jamey Gannon · 2026 · observation 10:16 #

Her "prompt booster" paragraph makes generations approximately 50% more accurate for text/logo/product consistency.

Jamey Gannon · 2026 · observation 11:32 #

Higgsfield Soul (not Soul 2.0) is the model to use for realistic UGC generation.

Jamey Gannon · 2026 · recommendation 15:48 #

Higgsfield's back-end prompt rewriter is unusually helpful for UGC use cases (contrary to her usual dislike of prompt rewriters).

Jamey Gannon · 2026 · opinion 16:23 #

AI understands abstract concepts like "luxury" and applies associated aesthetics (clothing, interior design, location) automatically.

Jamey Gannon · 2026 · observation 17:21 #

Prompting for a visually similar stand-in object (e.g., a candle instead of a supplement jar) reduces work required during product insertion.

Jamey Gannon · 2026 · observation 18:20 #

Product consistency across scale generations remains a weak spot; training a mini-model on your product is an emerging solution.

Jamey Gannon · 2026 · observation 24:14 #

For high-quality outputs, Photoshop is still often necessary for fine edits (logos, small details); this is not "doing it wrong.

Jamey Gannon · 2026 · opinion 25:53 #

The do's and don'ts pulled from the session

Do this
  • Jamey Gannon: Create a blank product mockup first (by removing branding from a similar competitor's product), then combine it with your cropped label in a separate step. 01:58 #
  • Jamey Gannon: Crop your label to show only the front-facing portion before feeding it to the model. 03:20 #
  • Jamey Gannon: Use multimodal models (Nano Banana Pro, GPT Image 2) whenever precision, text, or subject consistency is required. 05:41 #
  • Jamey Gannon: Break complex generation tasks into discrete sequential steps. 07:35 #
  • Jamey Gannon: Use the SWORD framework (Subject, Where, Orientation, Rendering, Directives) to structure prompts. 09:17 #
  • Jamey Gannon: Use "product photography shot" as a shorthand phrase to invoke centered, high-quality framing conventions. 10:16 #
  • Jamey Gannon: Append a reusable "prompt booster" paragraph to every generation involving text, logos, or product consistency. 11:17 #
  • Jamey Gannon: For editorial/stylized shots, upload your product image plus a style reference image and prompt the model to combine them. 11:50 #
  • Jamey Gannon: Try multiple models (Nano Banana Pro and GPT Image 2) on the same prompt to compare outputs. 13:20 #
  • Jamey Gannon: For UGC, use Higgsfield Soul with style presets (0.5 selfie, iPhone, general) rather than Soul 2.0. 15:48 #
  • Jamey Gannon: Include "choice words" like "luxury" to invoke rich, pre-associated aesthetic contexts without long descriptions. 17:21 #
  • Jamey Gannon: Pre-plant a stand-in object with similar visual weight to your product (e.g., a matte black candle for a jar) in the initial scene generation. 18:20 #
  • Jamey Gannon: Explicitly instruct the model to "adjust the angle and lighting to fit seamlessly with the scene" when inserting products. 18:58 #
  • Jamey Gannon: Fix AI slop by replacing objects with things your target audience would recognize (e.g., swapping a generic book for a David Goggins title). 19:32 #
  • Jamey Gannon: Use Photoshop for fine details when needed; don't rely on AI for every edit. 25:53 #
Don't do this
  • Jamey Gannon: Don't use diffusion models (Midjourney, Flux) when text or subject consistency matters. 06:05 #
  • Jamey Gannon: Don't attempt to one-shot complex mockup generations; prompt adherence drops significantly. 08:18 #
  • Jamey Gannon: Don't rely on a lazy "insert this product" prompt — models may superimpose the product like a sticker without matching lighting or perspective. 19:05 #
  • Jamey Gannon: Don't select Higgsfield Soul 2.0 for realistic UGC — use regular Soul. 15:48 #
  • Jamey Gannon: Don't expect AI to read your mind on unusual product forms (e.g., optical-illusion patterned toilet paper holder) — provide many reference angles. 24:14 #

Numbers quoted in this talk

"This helps my generations be like 50% more accurate." — Jamey Gannon, 11:32 — referring to appending her prompt booster paragraph.
2026 · #
"After just about 23 minutes, we have four totally new from scratch product photos." — Jamey Gannon, 20:30 — self-reported workshop pace.
2026 · #

Everything referenced on-screen and by name

People mentioned (excluding speakers listed above)

  • David Goggins — Author — cited — His book "Can't Hurt Me" used as a book-cover swap example.
  • Mario Testino — Photographer — cited — Referenced in a "vibes" prompt example on the Starting Conditions slide.

Brands / companies referenced

  • Portal — Speaker's own sleep supplement company; primary product example.
  • Pomelli by Google — Google Labs AI experiment where speaker leads social.
  • Armra — Colostrum supplement brand whose jar was used as a base mockup.
  • Byredo — Perfume brand (Animalique) used as an editorial style reference.
  • Nike — Referenced as a logo that got "slopped up" in a UGC example.
  • Motion — Host organization for the Creative Strategy Summit 2026.

Tools / products referenced (excluding Motion)

  • Flora — Primary multimodal AI canvas tool used throughout the demo.
  • Higgsfield — AI platform used for UGC generation.
  • Higgsfield Soul — Specific model recommended for realistic UGC (not Soul 2.0).
  • Nano Banana Pro — Multimodal image model.
  • GPT Image 2 — Multimodal image model.
  • Midjourney — Diffusion model, cited as aesthetics-optimized.
  • Flux — Diffusion model.
  • Weavy (Figma) — Multimodal AI tool mentioned in passing.
  • Krea — AI tool mentioned in passing.
  • Magnific — AI tool mentioned in passing.
  • Pixfield — AI tool mentioned in passing.
  • Spline — 3D software mentioned as legacy alternative.
  • Adobe Illustrator — Source of the label PNG export.
  • Adobe Photoshop — Recommended for fine-detail post-processing.
  • CapCut — Analogy for tool learning curves.
  • The AI Creative Director — Speaker's own course (next cohort July).
  • brand-sprints.com/links — Speaker's resource hub.

External frameworks / concepts cited

10 ads referenced

Show all 10 ads with extraction details
Ad #1 — PURE POUCHES
PURE POUCHES ·image ·00:06
Duration shown in this video
2 seconds
Hook (first 3 sec)
A top-down shot of a white circular tin of "PURE POUCHES" resting on a bed of fresh mint leaves.
Product / pitch
Nicotine-free pouches with a mint flavor.
Key on-screen text
PURE POUCHES
Key spoken lines
None used
Visual style
high-fi
CTA / offer (if shown)
None used
Narrative arc
None observable
Why shown in this video
To show an example of AI-generated product photography for e-commerce teams.
Speaker's take
None used
Ad #2 — Woman with product UGC
unknown brand ·UGC ·00:06
Duration shown in this video
2 seconds
Hook (first 3 sec)
A smiling blonde woman in a white blazer holds up a small, white product jar, with palm trees in the background.
Product / pitch
A cosmetic or supplement product in a small white jar.
Key on-screen text
Your Story, 17h
Key spoken lines
None used
Visual style
UGC
CTA / offer (if shown)
None used
Narrative arc
None observable
Why shown in this video
To show an example of AI-generated product photography for e-commerce teams.
Speaker's take
None used
Ad #3 — seek can
seek ·image ·00:06
Duration shown in this video
2 seconds
Hook (first 3 sec)
A close-up shot of a white "seek" beverage can submerged in a dark, bubbly liquid with ice cubes.
Product / pitch
A canned beverage.
Key on-screen text
seek
Key spoken lines
None used
Visual style
high-fi
CTA / offer (if shown)
None used
Narrative arc
None observable
Why shown in this video
To show an example of AI-generated product photography for e-commerce teams.
Speaker's take
None used
Ad #4 — Blue dropper bottle
unknown brand ·image ·00:06
Duration shown in this video
2 seconds
Hook (first 3 sec)
A blue glass dropper bottle is shown against a dark blue background.
Product / pitch
A serum or tincture in a dropper bottle.
Key on-screen text
None used
Key spoken lines
None used
Visual style
high-fi
CTA / offer (if shown)
None used
Narrative arc
None observable
Why shown in this video
To show an example of AI-generated product photography for e-commerce teams.
Speaker's take
None used
Ad #5 — PORTAL product shots
PORTAL ·image ·00:17
Duration shown in this video
14 seconds
Hook (first 3 sec)
A set of four images shows a black supplement jar with the word "PORTAL" in metallic purple text. The images include a standard product shot, a shot with a reflection, two jars floating, and a lifestyle shot of a man reading with the jar nearby.
Product / pitch
A supplement product in a black jar.
Key on-screen text
PORTAL, SUPERCHARGE YOUR SLEEP, MIDNIGHT PASSIONFLOWER CHAMOMILE
Key spoken lines
None used
Visual style
high-fi
CTA / offer (if shown)
None used
Narrative arc
None observable
Why shown in this video
To illustrate the different types of product photography that can be created with AI, from studio shots to lifestyle images.
Speaker's take
"So we're going to go over starting from a label design or just like a basic iPhone photo. We're going to do creating high quality studio photography, creating editorial and stylized photography, and then also creating AI UGC and inserting your product into it."
Ad #6 — ARMRA product shot
ARMRA ·image ·03:50
Duration shown in this video
2 seconds
Hook (first 3 sec)
A studio product shot of a black ARMRA supplement jar on a white background.
Product / pitch
A colostrum supplement for performance revival.
Key on-screen text
ARMRA, COLOSTRUM, Performance Revival
Key spoken lines
None used
Visual style
high-fi
CTA / offer (if shown)
None used
Narrative arc
None observable
Why shown in this video
To show how to use a competitor's product image as a reference to create a blank mockup for your own product.
Speaker's take
"I actually didn't have a picture of the plain black jar... but I remembered that Armra has an extremely similar, probably the exact same packaging that we do. So I just took Armra's."
Ad #7 — Diffusion model examples
unknown brand ·image ·05:36
Duration shown in this video
2 seconds
Hook (first 3 sec)
A split-screen shows three highly aesthetic, artistic images: a face with rainbow light, a woman in a mossy sphere, and an office chair.
Product / pitch
Not applicable. These are style examples.
Key on-screen text
Diffusion
Key spoken lines
None used
Visual style
high-fi, artistic, abstract
CTA / offer (if shown)
None used
Narrative arc
None observable
Why shown in this video
To contrast the output of diffusion models (optimized for aesthetics) with multimodal models.
Speaker's take
"...this is different from diffusion models like Midjourney or Flux that are more optimized for aesthetics."
Ad #8 — Multimodal model examples
unknown brand ·image ·05:36
Duration shown in this video
2 seconds
Hook (first 3 sec)
A split-screen shows three realistic images: a man in a hoodie, a woman's portrait, and a product shot with text.
Product / pitch
Not applicable. These are style examples.
Key on-screen text
Multimodal, Real Cellular Repair.
Key spoken lines
None used
Visual style
high-fi, photorealistic
CTA / offer (if shown)
None used
Narrative arc
None observable
Why shown in this video
To show that multimodal models are better for realistic imagery, product photography, and following precise instructions.
Speaker's take
"So both of these models are actually multimodal models, which means they can actually think and understand what you're asking for."
Ad #9 — BYREDO ANIMALIQUE perfume
BYREDO ·image ·11:44
Duration shown in this video
2 seconds
Hook (first 3 sec)
A high-quality studio shot of a BYREDO perfume bottle with a black cap, sitting on a reflective surface with a blue gradient background.
Product / pitch
A luxury perfume.
Key on-screen text
BYREDO, ANIMALIQUE, EAU DE PARFUM
Key spoken lines
None used
Visual style
high-fi, editorial
CTA / offer (if shown)
None used
Narrative arc
None observable
Why shown in this video
To use as a style reference for generating a new product image.
Speaker's take
"So all you're going to need to do is find an image with the style you love."
Ad #10 — UGC Instagram Stories
unknown brand ·UGC ·14:51
Duration shown in this video
2 seconds
Hook (first 3 sec)
A slide shows four different user-generated content style images formatted as Instagram Stories. They depict a woman holding a product, a man at the gym, a man holding a supplement bottle, and a woman in bed.
Product / pitch
Various products in lifestyle settings.
Key on-screen text
Your Story, 17h
Key spoken lines
None used
Visual style
UGC, lo-fi
CTA / offer (if shown)
None used
Narrative arc
None observable
Why shown in this video
To introduce the concept of creating AI-generated UGC.
Speaker's take
"Okay, let's move on to UGC."

26 slides, in order

Show all 26 slides with full slide content
Slide #1 — Creative Strategy Summit 2026 Intro
image+text ·00:00 ·Play
Title / header text
CREATIVE STRATEGY SUMMIT 2026
Body content
• Motion present
Embedded data (charts/tables)
None used
Embedded examples
• Motion logo • Abstract wireframe globe icon
Annotations / visual emphasis
None used
Reveal state
None used
Re-reference
None used
Speaker's framing
"All right. Hi everyone."
Slide #2 — Creative Strategy Summit 2026 Speakers
3x3 grid ·00:01 ·Play
Title / header text
CREATIVE STRATEGY SUMMIT 2026
Body content
• Motion presents • NEW WORLD. NEW PLAYBOOK • Watch the full event
Embedded data (charts/tables)
None used
Embedded examples
• A 3x3 grid of 8 speaker headshots.
Annotations / visual emphasis
A mouse cursor hovers over and selects one of the headshots.
Reveal state
None used
Re-reference
None used
Speaker's framing
"Um, today I'm going to show you a very, very simplified system..."
Slide #3 — AI Workshop Title
image+text ·00:03 ·Play
Title / header text
AI WORKSHOP:
Body content
• USING AI TO GENERATE QUALITY PRODUCT PHOTOS AT SCALE • WITH JAMEY GANNON • Watch the full event
Embedded data (charts/tables)
None used
Embedded examples
• Headshot of Jamey Gannon.
Annotations / visual emphasis
None used
Reveal state
None used
Re-reference
None used
Speaker's framing
"...for making basically any AI product photography that you could possibly need for your brand."
Slide #4 — Presentation Title
image+text ·00:06 ·Play
Title / header text
AI Product Photography for E-Commerce Teams
Body content
• With Jamey Gannon
Embedded data (charts/tables)
None used
Embedded examples
• A small video feed of the speaker, Jamey Gannon, is in the top left corner. • Four product photos are displayed at the bottom: • A tin of "PURE POUCHES" on mint leaves. знамени- A woman holding a product in an Instagram story format. • A can of "seek" soda with ice. • A blue dropper bottle.
Annotations / visual emphasis
None used
Reveal state
None used
Re-reference
None used
Speaker's framing
"Uh, we only have 30 minutes today, so I'm going to run through sort of like my favorites and the most core things I think you may need."
Slide #5 — Agenda
2x2 grid ·00:17, revisited 20:05, 20:33 ·Play
Title / header text
None used
Body content
None used
Embedded data (charts/tables)
None used
Embedded examples
• A small video feed of the speaker, Jamey Gannon, is in the top left corner. • Four product photos of a black jar labeled "PORTAL": • Top-left: Standard product shot on a white background. • Top-right: Two jars floating against a white background. • Bottom-left: Product shot on a reflective surface with a blue gradient background. • Bottom-right: A man on a couch reading a book with the product on the table next to him.
Annotations / visual emphasis
None used
Reveal state
None used
Re-reference
At 20:05 and 20:33, the slide is shown again without changes.
Speaker's framing
"Um, so we're going to go over starting from a label design or just like a basic iPhone photo. We're going to do creating high quality studio photography, creating editorial and stylized photography, and then also creating AI UGC and inserting your product into it."
Slide #6 — Speaker Introduction
image+text ·00:32 ·Play
Title / header text
Jamey Gannon
Body content
@jameygannon (aka techbimbo)
Embedded data (charts/tables)
None used
Embedded examples
• A small video feed of the speaker, Jamey Gannon, is in the top left corner. • A large photo of Jamey Gannon winking.
Annotations / visual emphasis
None used
Reveal state
None used
Re-reference
None used
Speaker's framing
"But before we dive in, I got to want to share a little bit about who I am. So my name is Jamie. If you follow me on X, I'm kind of known there as Tech Bimbo."
Slide #7 — "I have like 4 jobs"
mixed ·00:39 ·Play
Title / header text
I have like 4 jobs
Body content
• Freelance Brand Designer and (AI) Creative Director • Head of Brand @ Portal (supplement company) • Teaching creatives how to be AI Creative Directors • Content Creator • Leading social @ Pomelli by Google
Embedded data (charts/tables)
None used
Embedded examples
• A central animated GIF cycles through various branding and design examples. • A product shot of a "PORTAL" supplement jar. • A screenshot of a course titled "Learn to Control AI like a Creative Director". • A screenshot of a social media post about "NEW Pomelli Features".
Annotations / visual emphasis
Text labels point to the different visual elements on the slide.
Reveal state
None used
Re-reference
None used
Speaker's framing
"Um, and I don't know about you guys, but I find it harder and harder to introduce myself like every week."
Slide #8 — Label Design to Mock-up Workflow
hierarchy diagram ·01:56, revisited 07:03 ·Play
Title / header text
None used
Body content
• 01: Label Design • 02: Label Design • 03: Mock-up creation
Embedded data (charts/tables)
None used
Embedded examples
• A small video feed of the speaker, Jamey Gannon, is in the top left corner. • Step 1 shows a full, flat product label for "PORTAL". • Step 2 shows a cropped version of the label. • Step 3 shows a three-part flow: an image of an "ARMRA" product, which is transformed into a plain black canister, which is then combined with the label to create the final "PORTAL" product mock-up.
Annotations / visual emphasis
Arrows connect the steps in the workflow.
Reveal state
None used
Re-reference
At 07:03, the slide is shown again.
Speaker's framing
"So, uh, I'm going to be showing you guys a couple techniques, basically is how I think of it. And the first technique I want to show you guys is creating a product shot from just a label."
Slide #9 — Removing Design from a Canister
screenshot-with-annotations ·03:51 ·Play
Title / header text
None used
Body content
• Prompt: Remove all design from this canister, make it a plain matte black canister.
Embedded data (charts/tables)
None used
Embedded examples
• A screenshot of the Flora AI interface. • Left image: A product shot of an "ARMRA" canister. • Right image: A plain, matte black canister generated by the AI.
Annotations / visual emphasis
An arrow connects the "ARMRA" image to the prompt and the resulting plain canister.
Reveal state
None used
Re-reference
None used
Speaker's framing
"So I just took Armra's, um, and I did a very simple prompt just telling the model to like remove all the text."
Slide #10 — Applying Label to Mock-up
screenshot-with-annotations ·04:02 ·Play
Title / header text
None used
Body content
• Prompt: Place this matte black label on the base of the plastic black jar. The label should be nearly the same height as the base of the jar, with a small margin on the top and the bottom. The large purple "PORTAL" logo is a metallic pinkish purple hue, all the other text is white. The color of the label should be the same as the jar, with a smooth matte textural difference, making it reflect light slightly different.
Embedded data (charts/tables)
None used
Embedded examples
• A screenshot of the Flora AI interface. • Top-left image: The cropped "PORTAL" label design. • Bottom-left image: The plain matte black canister. • Right image: The final generated product mock-up with the label applied.
Annotations / visual emphasis
Arrows connect the two input images to the prompt and the final output image.
Reveal state
None used
Re-reference
None used
Speaker's framing
"Then from here, I just created a new image block and I connected that mock-up that I made, and then I connected that cropped label design and then I gave it this prompt."
Slide #11 — Final Product Image
image-only ·05:12 ·Play
Title / header text
None used
Body content
None used
Embedded data (charts/tables)
None used
Embedded examples
• A close-up of the final AI-generated product shot for "PORTAL".
Annotations / visual emphasis
None used
Reveal state
None used
Re-reference
None used
Speaker's framing
"And then here we have our final result."
Slide #12 — Why Nano Banana Pro/GPT 2?
2-column comparison ·05:36 ·Play
Title / header text
Why Nano Banana Pro/GPT 2?
Body content
Diffusion
(Column with 3 images)
Multimodal
(Column with 3 images)
Embedded data (charts/tables)
None used
Embedded examples
Diffusion examples
Abstract face with rainbow light, woman in a mossy sphere, a white office chair.
Multimodal examples
Man in a hoodie, a woman making a surprised face, a product shot for "Real Cellular Repair".
Annotations / visual emphasis
None used
Reveal state
None used
Re-reference
None used
Speaker's framing
"So first, why are we using Nano Banana Pro and or GPT-2? And why are they so uniquely good at product photography and realistic imagery like this?"
Slide #13 — Starting Conditions / Inputs
3-column list ·06:22 ·Play
Title / header text
Starting Conditions / Inputs
Body content
• **Model Selection** • Diffusion (with 2 example images) • Multimodal (with 2 example images) • **Reference Material** • Subject (with 4 example images) • Scene (with 4 example images) • Style (sref/style ref) (with 4 example images) • **Prompt**
Vibes
Woman in a sun-kissed meadow. In the style of Mario Testino.
Specific
A long, detailed paragraph describing a photo shoot.
Embedded data (charts/tables)
None used
Embedded examples
Various small thumbnail images are used to illustrate each category under "Model Selection" and "Reference Material".
Annotations / visual emphasis
None used
Reveal state
None used
Re-reference
None used
Speaker's framing
"Now, I consider multimodal models one of many starting conditions you have for your generations."
Slide #14 — Work in steps
image+text ·07:32 ·Play
Title / header text
04
Body content
Work in steps
Embedded data (charts/tables)
None used
Embedded examples
• A photo of a woman in a white dress walking up stairs, her blonde hair in motion.
Annotations / visual emphasis
None used
Reveal state
None used
Re-reference
None used
Speaker's framing
"...you should always work in steps."
Slide #15 — One-Step Workflow Example
hierarchy diagram ·08:14 ·Play
Title / header text
None used
Body content
None used
Embedded data (charts/tables)
None used
Embedded examples
• A screenshot of the Flora AI interface showing a simplified, less effective workflow. • Top input: The full, uncropped "PORTAL" label. • Bottom input: The "ARMRA" product image. • Output: A poorly generated "PORTAL" product mock-up where the label is distorted and incorrect.
Annotations / visual emphasis
Arrows connect the two inputs to the single output.
Reveal state
None used
Re-reference
None used
Speaker's framing
"So for example, I want to show you what it looks like if we do the exact same task, but just jump straight into it."
Slide #16 — iPhone to Studio Shot Workflow
screenshot-with-annotations ·08:49 ·Play
Title / header text
None used
Body content
• Prompt: Create an ultra-high-quality product photography shot of the matte black supplement jar with metallic blue text against a bright white background. Shot on a Canon 5D Mark IV 50mm lens, f/2.8, front lit with a low contrast light. Crisp details. Ensure the "DREAM" text on the bottom stays in the same serif font. All text, logos, typography, iconography, and graphic elements must remain exactly as they appear in the reference image... [rest of prompt is visible but not fully transcribed by speaker]
Embedded data (charts/tables)
None used
Embedded examples
• A screenshot of the Flora AI interface. • Left image: A casual iPhone photo of the "PORTAL" product on a table. • Right image: A high-quality, studio-style product shot generated by the AI.
Annotations / visual emphasis
An arrow connects the iPhone photo to the prompt and the final studio shot.
Reveal state
None used
Re-reference
None used
Speaker's framing
"I want to show you guys a very similar task and output, but we're going to approach it in a different way, and I think that this one will be like much more common as well for you guys."
Slide #17 — SWORD Framework
bullet list ·09:17 ·Play
Title / header text
S W O R D
Body content
• **S**ubject: Who or what is in the image. • **W**here: Setting, scene, background, placement. • **O**rientation: Pose, angle, framing, product position, body direction. • **R**endering: Camera, lighting, color, style, texture, realism. • **D**irectives: Rules, constraints, common-sense checks
Embedded data (charts/tables)
None used
Embedded examples
None used
Annotations / visual emphasis
None used
Reveal state
None used
Re-reference
None used
Speaker's framing
"So when I'm prompting for specific things, like the last thing we just did, I am unconsciously and sometimes consciously doing what I call the SWORD framework."
Slide #18 — Prompt Breakdown with SWORD
text-only ·09:21 ·Play
Title / header text
None used
Body content
• Create an ultra-high-quality product photography shot of the matte black supplement jar with metallic text. Photographed in a white photo-studio environment with strong directional natural light from one side. Shot on a Canon 5D Mark IV 50mm lens. Crisp details. • All text, logos, typography, iconography, and graphic elements must remain exactly as they appear in the reference image, with no reinterpretation, stylization, resizing, reflowing, or distortion of any kind. Do not alter spacing, alignment, scale, or relative dimensions of any elements. The product's geometry, height-to-width ratio, curvature, and panel relationships must remain identical to the reference. Do not change any of the fonts. Do not change any of the colors.
Embedded data (charts/tables)
None used
Embedded examples
None used
Annotations / visual emphasis
State 1 (09:21)
"matte black supplement jar with metallic text" is highlighted in black.
State 2 (09:28)
"Photographed in a white photo-studio environment" is highlighted in black.
State 3 (09:32)
"product photography shot" is highlighted in black.
State 4 (09:36)
"strong directional natural light from one side. Shot on a Canon 5D Mark IV 50mm lens. Crisp details." is highlighted in black.
State 5 (09:44)
The entire second paragraph starting with "All text, logos..." is highlighted in black.
Reveal state
The slide progressively highlights different parts of the prompt text.
Re-reference
None used
Speaker's framing
"So let's like break down this prompt a little bit more."
Slide #19 — Style Transfer Workflow
hierarchy diagram ·11:51 ·Play
Title / header text
None used
Body content
• Prompt: Create an image of the product in the first photo in the same style as the second reference photo. • [Prompt Booster Text]: All text, logos, typography, iconography, and graphic elements must remain exactly as they appear in the reference image...
Embedded data (charts/tables)
None used
Embedded examples
• A screenshot of the Flora AI interface. • Top-left input: A photo of a Byredo "Animalique" perfume bottle. • Bottom-left input: A product shot of the "PORTAL" jar with strong shadows. • Top-right output: The "PORTAL" jar shot in the style of the Byredo photo (blue gradient, reflective surface). • Bottom-right output: A second variation of the same concept.
Annotations / visual emphasis
Arrows connect the two input images to the prompt and then to the two output images.
Reveal state
None used
Re-reference
None used
Speaker's framing
"So all you're going to need to do is find an image with the style you love."
Slide #20 — UGC Examples
1x4 grid ·14:51 ·Play
Title / header text
None used
Body content
None used
Embedded data (charts/tables)
None used
Embedded examples
• Four images formatted to look like Instagram stories, showing different people in various settings (outdoors, gym, home).
Annotations / visual emphasis
None used
Reveal state
None used
Re-reference
None used
Speaker's framing
"So for this section, we're going to be using a combination of Pika's Soul and then again, either Nano Banana Pro or GPT Image 2."
Slide #21 — Higginsfield Platform UI
screenshot ·15:34 ·Play
Title / header text
None used
Body content
A screenshot of the Higginsfield AI platform, showing various models and tools like "Create Image", "Higginsfield Soul 2.0", "GPT Image 2", "Nano Banana Pro", etc.
Embedded data (charts/tables)
None used
Embedded examples
The screenshot shows the full user interface of the Higginsfield platform.
Annotations / visual emphasis
The speaker's cursor moves around the screen, highlighting different options.
Reveal state
None used
Re-reference
None used
Speaker's framing
"So Higginsfield is a really big platform. Um, so I'll show you exactly what you need to do."
Slide #22 — UGC Inpainting Example: Reading on Couch
image+text ·16:23 ·Play
Title / header text
None used
Body content
• [Simple Prompt]: A man reading his book on his couch in an apartment. In front of him is a matte black candle on a coffee table. • [Detailed Prompt]: A candid, spontaneous snapshot of a man casually reading a book while seated comfortably on a soft fabric couch... [rest of prompt is visible]
Embedded data (charts/tables)
None used
Embedded examples
• An AI-generated image of a man reading a book on a couch with a black candle on the table.
Annotations / visual emphasis
An arrow points from the simple prompt to the detailed prompt below the image.
Reveal state
None used
Re-reference
None used
Speaker's framing
"And then another great thing about Higginsfield is they have this incredible prompt rewriter in the back end..."
Slide #23 — UGC Inpainting Example: Woman in Apartment
image+text ·16:58 ·Play
Title / header text
None used
Body content
• [Simple Prompt]: A woman drinking a glass of light blue liquid in a luxury apartment in the evening. She is wearing pajamas • [Detailed Prompt]: Bathed in the soft, diffused glow of evening city lights filtering through floor-to-ceiling windows, a woman reclines gracefully on a plush velvet chaise lounge... [rest of prompt is visible]
Embedded data (charts/tables)
None used
Embedded examples
• An AI-generated image of a woman in silk pajamas on a chaise lounge in a high-rise apartment at night.
Annotations / visual emphasis
An arrow points from the simple prompt to the detailed prompt below the image.
Reveal state
None used
Re-reference
None used
Speaker's framing
"And a way that you can make this really work for you is by having choice words in your prompt, like 'luxury' for example."
Slide #24 — UGC Inpainting Workflow
hierarchy diagram ·18:05 ·Play
Title / header text
None used
Body content
• Prompt: Replace the black candle in the first image with the matte black supplement jar with metallic text in the second image. Maintain the exact same size. Adjust the angle and the lighting on the jar to fit in seamlessly with the scene.
Embedded data (charts/tables)
None used
Embedded examples
• A screenshot of the Flora AI interface. • Top-left input: The studio shot of the "PORTAL" product. • Bottom-left input: The AI-generated image of the man reading with a black candle. • Right output: The image of the man reading, but the candle has been replaced with the "PORTAL" product.
Annotations / visual emphasis
Arrows connect the two input images to the prompt and the final output image.
Reveal state
None used
Re-reference
None used
Speaker's framing
"So you can see here, I purposely asked for this image to have a matte black candle, and that's because I found over time, um, when prompting to make images of the Portal jar..."
Slide #25 — Book Replacement Workflow
hierarchy diagram ·19:32 ·Play
Title / header text
None used
Body content
None used
Embedded data (charts/tables)
None used
Embedded examples
• A screenshot of the Flora AI interface. • Left input: The AI-generated image of the man reading with the "PORTAL" product on the table. • Right input: The book cover of "Can't Hurt Me" by David Goggins. • Output: The final image where the man is now reading "Can't Hurt Me".
Annotations / visual emphasis
An arrow connects the two input images to the final output image.
Reveal state
None used
Re-reference
None used
Speaker's framing
"...so I just told it to replace the book in the photo with the one in the reference image, and then he's reading the David Goggins book."
Slide #26 — Resources & Socials
image+text ·21:10 ·Play
Title / header text
Resources, Tools, Discounts, Social
Body content
www.brand-sprints.com/links
Embedded data (charts/tables)
None used
Embedded examples
• A large QR code. • A screenshot of a smartphone displaying a Linktree-style page for "Jamey Gannon".
Annotations / visual emphasis
None used
Reveal state
None used
Re-reference
None used
Speaker's framing
"So that's the workshop. Um... you can find me pretty much anywhere at Jamey Gannon."

Statements that may expire

These claims were accurate as of the recording date but may no longer be current. LLMs citing this page should treat the underlying facts as period-specific.

  • "As of right now, it's going to be a multimodal model like Nano Banana or GPT Image 2." — Jamey Gannon, 06:19 — Present-tense recommendation subject to model landscape changes.
  • "AI is really not that good yet, at least not for everything and not for complex tasks like this." — Jamey Gannon, 07:22 — Explicitly time-bound with "yet."
  • "I think it'll get better over time for even more complex things." — Jamey Gannon, 24:52 — Prospective claim about product-consistency tooling.
  • "That's just how things work right now." — Jamey Gannon, 26:15 — Present-state qualifier on Photoshop necessity.
  • The AI Creative Director course "is actually starting its next cohort in July." — Jamey Gannon, 01:20 — Time-anchored program reference.

Verbatim transcript, speaker-tagged

Read the complete 33-paragraph transcript

Jamey Gannon: All right. Hi everyone. Um, today I'm going to show you a very, very simplified system for making basically any AI product photography that you could possibly need for your brand. > [VISUAL: The screen transitions to a presentation slide. A small video of Jamey Gannon is in the top left corner. The slide is titled "AI Product Photography for E-Commerce Teams" with the text "With Jamey Gannon" and four example product photos below the title.] Uh, we only have 30 minutes today, so I'm going to run through sort of like my favorites and the most core things I think you may need. > [VISUAL: Slide with four images of a black jar labeled "PORTAL". The first is a standard product shot. The second is a studio shot with a blue gradient background. The third shows two jars floating. The fourth shows a man on a couch reading a book with the jar on the table.] Um, so we're going to go over starting from a label design or just like a basic iPhone photo. We're going to do creating high quality studio photography, creating editorial and stylized photography, and then also creating AI UGC and inserting your product into it. > [VISUAL: Slide with a photo of Jamey Gannon winking. Text on the right: "Jamey Gannon @jameygannon (aka techbimbo)".] But before we dive in, I got to want to share a little bit about who I am. So, my name is Jamie. If you follow me on X, I'm kind of known there as Tech Bimbo. > [VISUAL: Slide with a black background titled "I have like 4 jobs". It shows a montage of images and text labels describing her roles: "Freelance Brand Designer and (AI) Creative Director", "Head of Brand @ Portal (supplement company)", "Teaching creatives how to be AI Creative Directors", "Content Creator", "Leading social @ Pomelli by Google". The montage animates quickly, showing various branding projects, websites, and social media content.] Um, and I don't know about you guys, but I find it harder and harder to introduce myself like every week. Um, job titles are really starting to blur. But anyway, if you've never come across my content before, I am a brand designer and creative director who uses AI every single day and I basically make a ton of content talking about it as well. So, my background is primarily in early stage companies across tech and D2C, but I now sort of split my time between those actual branding services, like this GIF that you see on the screen, and then helping companies and individuals bring generative AI into their creative workflows. So, doing workshops like this, um, talks, and then I also have a course called The AI Creative Director, which is actually starting its next cohort in July. Um, and then I have some side jobs if you can call them that. Um, I actually have a supplement company with my boyfriend called Portal. It's a sleep supplement. And then I also, uh, lead social at a Google Labs experiment called Pomelli, which is another really cool AI tool, um, for business and e-com stuff. Um, but yeah, basically I've gotten to work with loads and loads of people and teams across like every single intersection of design, business, tech, AI. Um, and I've actually even worked with like some of the speakers, um, today as well on this very stuff for their brands. Um, yeah, so that's who I am. Uh, let's dive into it.

Slide with a black background showing a 3-step process. "01: Label Design" shows a flat, unwrapped product label for "PORTAL". "02: Label Design" shows the same label but cropped to just the front-facing portion. "03: Mock-up creation" shows an image of a competitor's product ("ARMRA"), then an arrow to a blank black jar, then an arrow to the final mockup with the "PORTAL" label applied.] So, uh, I'm going to be showing you guys a couple techniques, basically is how I think of it. And the first technique I want to show you guys is creating a product shot from just a label. And I love this use case for a few reasons. Number one, it's great if you're just trying to iterate on a new design or like a new SKU. And two, with this technique, you can actually start making images of your product before it's even in production. And I think this opens doors for so many things like paid ad testing, doing pre-sales, and also just moving faster in this market, which I know is probably top of mind for a lot of you.

So, to break this down, firstly, this software that you're seeing, anytime you're seeing like a dark mode software, um, in this talk today, it's going to be Flora. Uh, and Flora is just a multimodal tool. They're very popular for their node system. Um, but this workflow is like really tool agnostic. Um, there's so many incredible AI tools to use these days. Uh, I use almost all of them. Um, you have like Weavvy from Figma, you have Pixfield, Krea, Magnific. Uh, for the product design or for the product photography use case that we're talking about today, the models are primarily going to be using our GPT Image 2 or Nano Banana Pro. So, any tool where you can access those models is going to be totally fine for these techniques.

Um, but yeah, the first thing you're going to want to do is get your label design. So this is just a PNG export of, uh, my Illustrator file, like my actual packaging design. And then I just place it in the top here. And then the next thing you're going to want you're going to want to do is crop it and you just want to show the part that's actually going to be in the image. Um, so make sure you take out like your supplement facts, your story facts, or whatever else you have on your packaging that you don't want on there. And then next you're going to need your actual mockup. So in this example here, I actually didn't have a picture of the plain black jar for whatever reason. I couldn't find it from my supplier, I couldn't find it online. > [VISUAL: Slide showing the process of removing a label. On the left is the "ARMRA" product. An arrow points to the right, where a plain black jar is shown. Below the plain jar is a text prompt: "Remove all design from this canister, make it a plain matte black canister."] Uh, but I remembered that Armra has an extremely similar, probably the exact same packaging that we do. Um, so I just took Armra's, um, and I did a very simple prompt just telling the model to like remove all the text. Um, so what I say here, I said, remove all the text from this canister and make it a plain black canister. And you can see, it just did that.

Slide showing the final step. The cropped "Supplement Label" and the "Matte Black Canister" are shown as inputs pointing to the final "PORTAL" product image. A long text prompt is displayed on the right.] Then from here, I just created a new image block and I connected that mockup that I made, and then I connected that cropped label design and then I gave it this prompt. Um, you guys can read this, I won't read it to you. But there's a couple things that I want to call out on this prompt here. So you can see it's quite detailed in terms of instructions. And this is something that you will learn from trial and error with your exact specs for your product. So, I've tried to make this image many times over the last few years with all these different AI tools. And I noticed that, you know, there's a few things it constantly messes up like the the texture of the label, the exact color of the label, um, the color and material of the text. Like sometimes AI will just make it not the way it looks in person. Um, but over time, I now kind of have this like master prompt that I use, uh, whenever I'm doing imagery of this particular product. And over time for your product as well, especially if you have like really unique form factors, um, or a super custom packaging, like if you're a perfume or something, you'll need to figure out like what does the model need to know every single time, because otherwise it's just going to like default to kind of like what it knows. > [VISUAL: A close-up of the final generated product image of the "PORTAL" jar.] Um, and often times it won't get this right. So that's one important note. Um, and then here we have our final result. Um, I wish I should have put a jar of portal here, but this is exactly what it looks like. Um, it's matte, it has the metallic text. Um, and yeah, I think this was a pretty good result. Now, what we just did may seem simple, but we actually made some intelligent and considered decisions when running through this process and I want to break them down for us.
Slide titled "Why Nano Banana Pro/GPT 2?". It's split into two sections. "Diffusion" on the left shows three artistic, abstract images. "Multimodal" on the right shows three realistic photos of people and a product.] So first, why are we using Nano Banana Pro and or GPT2? And why are they so uniquely good at product photography and realistic imagery like this? Um, so both of these models are actually multimodal models, which means they can actually think and understand what you're asking for. And this is different from diffusion models like Midjourney or Flux that are more optimized for aesthetics. > [VISUAL: Slide titled "Starting Conditions / Inputs". It has three columns: "Model Selection" (with examples for Diffusion and Multimodal), "Reference Material" (with examples for Subject, Scene, and Style), and "Prompt" (with examples for Vibes and Specific).] And basically, if you're using a multimodal model, or if you aren't using a multimodal model, you should just not expect it to keep your subject or your text, for example, exactly the same. So whenever you see something that involves extreme precision or instructions, like mockups or text, or even these images I made of myself here, uh, as of right now, it's going to be a multimodal model like Nano Banana or GPT Image 2. Um, now I consider multimodal models one of many starting conditions you have for your generations. And the first starting condition is model selection, which we just went over. The next is your reference material. So models and tools handle this differently. Um, for most of them, it's image uploads. Sometimes it's mood boards or codes if you're using Midjourney, but they're used to do three things: to establish your subject, like an object or a character, to establish your scene or establish your style, like illustrative versus photorealistic.
The slide from 1:56 reappears, showing the 3-step process for creating the product mockup.] And then lastly, the one that we all know is your prompt. And depending on your other starting conditions and sort of what your goal is, um, you can do more vibey short prompts or longer, more instructional prompts. Um, next, you might be wondering, Jamie, if AI is so good, like everyone has told us, I don't know about today, but like I feel like the narrative is like AI is incredible and it's taking over everything. Um, you know, why did we have to work in so many steps like this? Why couldn't we just like one shot it or just like do one prompt? Um, well, the truth is, AI is really not that good yet, at least not for everything and not for complex tasks like this. Um, so this I want to kind of talk about some of my principles I actually teach in my course. I have five main ones. > [VISUAL: Slide with the number "04" in the top left. The main text is "Work in steps". To the right is a photo of a woman in a white dress running up stairs, her blonde hair flying.] Um, and one of them is that you should always work in steps. Now, if you've ever been in like a coaching call with me, or been in like a session where I'm like reviewing prompts or helping people work through why their prompts aren't working, you'll probably hear me say this 10 times. Um, but basically when you are dealing with complex tasks, it's always going to be easier to tackle the problem in steps, whether or, you know, when you ask the model or even yourself to do many things at once, your prompt adherence is going to be lower and you're just going to have more variables to work with. But when you work in steps, you're able to nail your elements one at a time and it'll just make your life and your generations a lot easier. Now, you can definitely break the rule with this depending on the model and the task and like how particular you are about your prompting, but I would say as a best practice, most of the time you should work in steps.
Slide showing a failed one-step attempt. The "Sleep Supplement Label" and the "Nutritional Supplement" (the ARMRA product) are shown as inputs pointing to a poorly generated "PORTAL" product image where the label is distorted and incorrect.] Um, so for example, I want to show you what it looks like if we do the exact same task, but just jump straight into it. Um, so I'm using the same image, the same label, the same model family, and practically the same prompt. Um, the only thing I did was say remove the text on the first image. Um, and you can see how much lower the prompt adherence is. So it's inventing details, it's missing colors, and like overall even the image just doesn't look as good. Like the aesthetic of the actual jar is kind of bad to me. Um, so, yeah. > [VISUAL: Slide showing a process. An iPhone photo of a "PORTAL" jar on a table is the input. The output is a high-quality studio shot of the same jar. A long, detailed prompt is shown on the right.] Okay. I want to show you guys a very similar task and output, but we're going to approach it in a different way, and I think that this one will be like much more common as well for you guys. Um, so we're still making a studio quality image with a white background, but this time we're going to start from an iPhone photo. So, as you can see here, this is my only input and it's just this image and this prompt in Nano Banana Pro. And this is the result we got.
Slide titled "SWORD" with an acronym breakdown: Subject, Where, Orientation, Rendering, Directives.] So, it's pretty clean, I would say. And let's break down this prompt. So, when I'm prompting for specific things, like the last thing we just did, I am unconsciously and sometimes consciously doing what I call the sword framework. Um, so S, which is the subject, which is who or what is in the image. We have W, the where, so the setting, the scene, the background, the placement. We have O, orientation. So it's like pose, framing, product position, body direction. We have rendering, which, um, is kind of like camera style, lighting, color, texture, um, is it realistic? Is it illustrative? And then D is for directives. So these are like the rules, the constraints, common sense checks. Um, you know, when you're saying like, don't change anything, those are the sort of instructions that you have to add. > [VISUAL: The prompt from the previous slide is shown again, with different parts highlighted as the speaker talks.] And I love this framework because it's extremely legible for even the least like creatively minded people on your team, like someone that doesn't necessarily think of themselves as an image maker or creative director. Um, and it's a really good reminder of what details AI needs to be given. Uh, you can't just like lazy prompt everything, especially when you're trying to be as specific as like product photography. Um, so let's like break down this prompt a little bit more. Um, so here we have our subject, which is matte black supplement jar with metallic text. Next, we have our where, which is a white photo studio environment. And then the orientation here is kind of sneaky. So, you may think saying product photography shot is kind of obvious, but it's actually sort of a shortcut that we're giving the model. So, the model has access to the internet and, you know, training, it knows what product photography shots are and it knows that they're typically very high quality, very photorealistic, very straight on, not typically obscured or abstract. So, instead of saying like centered in the frame or other attributes that you want in a product photography shot, saying product photography shot just does that for us. Um, our rendering aspect are these details here. So we're saying ultra high quality, strong directional strong directional natural light, and the camera type and the style. And then lastly, we have our direction. Um, so this paragraph here is actually what I call one of my prompt boosters that I throw on any generation where I'm dealing with text or logo or product consistency. And this basically just says like, don't change anything. Um, and for whatever reason, this helps my generations be like 50% more accurate. Um, and I think as you start to develop your own like custom workflows and like I said, finding unique things that your brand or style or product needs, having these just like little prompt boosters ready to go all the time, um, can be really helpful.
Slide showing a process. On the left is a photo of a Byredo perfume bottle and a photo of a "PORTAL" jar with dramatic lighting. Arrows point to two generated images on the right, both showing the "PORTAL" jar in the style of the Byredo photo.] Okay. Next technique I want to show you guys is how you can achieve really aesthetic stylized editorial images like this, but with your own products. So, all you're going to need to do is find an image with the style you love. It doesn't actually have to be a product photography image. This is just like an easy example. Um, it could be literally any photo that you want. Um, and then you're basically just going to tell AI to create an image of the product in the first photo in the same style as the second reference photo. So, make sure you hook these up or upload them in whatever order you write the prompt in. Um, and then of course, this big paragraph here is that prompt booster again. Um, so a few things I want to note. Uh, we're using one of the iterations of that really clean white background studio shot. So, in this case, mine has some shadows to add just a little bit of like flair. Um, and you can tell that we're kind of going back to that work in steps principle. So, instead of having to tell the model to, you know, remove the background and make it a high quality image, like take it from iPhone to studio to then the aesthetic studio, we're already we're just skipping a few of those steps by giving it like a plain background studio shot. Um, and then for this example on the screen, I tried it in both Nano Banana Pro and GPT Image 2. You can see that the results are both great. Like technically, I think the GPT Image 2 one on the bottom was a little bit more faithful to the style, but for better or worse, they're both good results. Um, and I encourage you, you know, if you are going through a process and maybe something's not working or you just want to see like what's going to give you a result you like better, feel free to like try multiple models and multiple prompts. > [VISUAL: A new image appears, showing two cylindrical containers, one brown and one green, with the brand name "DALGONA" on them.] Um, I have a typo here. Oh, a composition. A composition like this would be really great for, uh, you know, if you to extend it in like 9 by 16 for ads, you could put a bunch of text on it. Um, or even like a hero shot for like a landing page. Um, and I really love how easy it is to get this like really like physical type of imagery that previously would have needed like a serious understanding of like spline or another 3D software. And you can just get this done in like a couple seconds really.
Slide showing a grid of various product photos in different styles: a perfume bottle, a soda can, a sunscreen tube, a dropper bottle, and a can held by a person.] So, you can basically rinse and repeat these techniques with almost any reference or idea or object. Um, you know, we're just see here like stealing the style. Um, you can even like have it be floating again, it could be sat down. You could even take like this uh, this dropper here, maybe even swap it for this one. Um, but at the end of the day, all you're doing is uploading images and then telling it how to use those images. > [VISUAL: Slide showing four images in the style of Instagram stories. They are selfies of different people holding or interacting with products.] Okay. Let's move on to UGC. So, for this section, we're going to be using a combination of Pixed Soul and then again, either Nano Banana Pro or GPT Image 2. So, the reason I love Pixed in particular for this task is two reasons. Um, number one, their model Pixed Soul. It's just extremely realistic. I don't know how they trained it, uh, or if it's like a layer on top of another model, but it's really, really good. And then two, they also have tons of templates and essentially shortcuts that are particularly great for e-com use cases, uh, when it comes to like content and consistent characters, um, and just like social-esque stuff.
A screen recording of the Pika Labs interface. The user navigates through different models and presets.] Um, and we can't like get into everything today, but I do want to encourage you to like go and play on that platform because it's kind of bonkers. Um, so Pixed is a really big platform. Um, so I'll show you exactly what you need to do. Um, you're going to go to this like farthest top left menu here. You're going to hit image and then generate image or create image. And then you're going to want to go down to the bottom and make sure that you select Pixed Soul as the model, not Pixed Soul 2.0, it's different. So just regular Pixed Soul. Um, and then you can actually adjust these like style presets. Um, they have tons of different ones. Uh, for UGC, I'm typically using like 0.5 selfie or iPhone or general. Um, but they also have like, you know, realistic, they have flight mode, which is like airport, like hotel mode, they have like Japandi, uh, like Tik Tok aesthetics basically, which depending on your audience can be really helpful shortcuts. > [VISUAL: Slide showing a process. On the left is a photo of a man reading a book on a couch with a black candle on the table. An arrow points to the right, where the candle has been replaced by the "PORTAL" product. A text prompt is shown above the final image.] And then another great thing about Pixed is they have this incredible prompt rewriter in the back end, which I typically don't love, but for this use case, I think it's really, really helpful. Um, so for example, to get this image, my prompt was something really simple. A man reading a book on his couch in an apartment. In front of him is a matte black candle on a coffee table. And then Pixed will then also using that style selector, rewrite the prompt in the back end to be much more detailed and have things that feel super natural and things that you actually want. Um, so it describes the texture of the candle, it describes the quality of the lighting in the room, it describes his outfit, which we never mentioned, um, all this stuff.
Slide showing a process. On the left is a photo of a woman in pajamas on a chaise lounge in a high-rise apartment at night. An arrow points to the right, where the prompt rewriter has generated a very detailed descriptive paragraph based on a simple initial prompt.] And a way that you can make this really work for you is by having choice words in your prompt, like luxury, for example. Um, so with this image, again, we did a very simple prompt, but because we used the word luxury, we saved a lot of time in actually having to describe what luxury might look like. So, again, AI understands the world, it understands what luxury is. It knows it's going to be a certain type of woman, a certain style of clothes, a certain style of interior design. We can see that she's obviously in a high rise in a city with lots of windows and nice drapes. And this sort of prompting actually works for every model, not just Pixed Soul. > [VISUAL: Slide showing three images in the style of Instagram stories: a man on a treadmill, a man sleeping, and a woman holding a product.] Um, so luxury is like one that's really easy to use. And just getting specific with your prompts in that way can make you like a good lazy prompter, I'll say. Um, and yeah, Pixed Soul is great at pretty much everything you can see here. Um, a lot of it looks like terrifyingly realistic. Um, but you'll notice if you look closely and you're not just like viewing these as like super quick ads, there's definitely a bit of slop in them. And then of course, our product is not inserted yet.
Slide showing a process. On the left are two images: the "PORTAL" product and the photo of the man reading with the candle. Arrows point to the final image on the right, where the candle is replaced by the "PORTAL" product. A text prompt is shown.] So, I want to show you this as an example. And if you'll notice, I actually did a bit of pre-work in my original prompt. So, you can see here, I purposely asked for this image to have a matte black candle. And that's because I found over time, um, when prompting to make images of the portal jar, if I said like supplement jar, it would sometimes give me like something very small, or it would give me like a protein jar or like odd shapes. And then it was just like another step for AI to work through, um, when I actually went to go do the insertion. Um, but one day when I was looking for inspiration for product photography, I realized like, oh, candles are like so similar to like the actual visual weight of Portal. So now I started doing this as a shorthand. And then when I go and do the replacement, um, it like just works a lot better. > [VISUAL: Slide showing a process. On the left is an iPhone photo of the "PORTAL" jar. The output on the right is a high-quality studio shot. The same jar. A long, detailed prompt is shown.] Um, another thing I have in this prompt here, um, the the prompt is like pretty self-explanatory. Um, but I also have like adjust the angle and the lighting to make it fit seamlessly with the scene. Um, sometimes the models can be a little bit dumb. If you say like insert this product into this image, it might just like superimpose it like a sticker. Um, but if you tell it to like actually, you know, make it look realistic, it can tend to do it better. And that's another benefit of having an object that looks similar enough to your product to begin with. It's not having to think so hard like what the lighting is going to be across the shadow and the shadow it's going to cast and things like that.
Slide showing a process. On the left is the photo of the man reading with the "PORTAL" jar. An arrow points to the right, where the book cover has been replaced with the cover of "Can't Hurt Me" by David Goggins. An image of the book cover is shown as an input.] Um, if you have like a supplement bottle that's like just like a plain white bottle, you can probably just do plain white bottle, but again, experiment and you may need to get creative, um, to get like the best results in this case. Um, and then when it comes to fixing slop, you can do like obvious things, like just telling it to like fix the shirt or change the color of something. Um, but you can also fix slop in a way that like really gets at your target audience, especially if you're making this UGC for like, uh, more evergreen content on your website. Um, so because he's reading a book in this image, I thought, hey, why not make the book something my target audience might like. Um, so I just like grabbed the cover of David David Goggins' book, um, on Google and just told it to replace the book in the photo with the one in the reference image. And then he's reading the David Goggins book. > [VISUAL: A slide with a black background showing a grid of images. Text overlays the images, making them hard to read. The speaker seems to be describing a different slide.] Um, oh no, there's text covering this. Um, and you can basically use the same flow again, whether it's like clothes or people or locations. Um, but again, while I'm doing this, I'm doing this very smart. So, my poor text. Let me see if I can remove this.
The slide from 20:05 is still on screen, with the text overlay issue. The speaker continues to describe a different process.] Uh, anyway, it's a, it's a guy in like a shirt with a Nike logo, but the Nike logo was like slopped up. Um, so I just told it to, uh, replace the shirt that he's wearing with this one, and then it looks more realistic. > [VISUAL: The slide from 00:17 reappears, showing the four "PORTAL" product photos.] Um, so after just about 23 minutes, we have four totally new from scratch product photos.

Jamey Gannon: And realistically, you could make like dozens of these in the same amount of time that we've been talking. Um, especially when you get into like certain automations and different tools and utilizing like the nodes and workflows. There's a lot of crazy stuff out there. Um, and more is definitely coming down the pike. But for this workshop, I wanted to focus on what's actually happening in the back ends of the models and get you like the basics. So as you start to use these more templated tools, um, you know what's going on and you know how to fix it. And yeah. So that's the workshop. > [VISUAL: Slide with a large QR code on the left and an image of a smartphone displaying a social media profile for "Jamey Gannon" on the right. Text on the slide: "Resources, Tools, Discounts, Social" and "www.brand-sprints.com/links". Music starts playing.] Um, I guess I have five minutes to do a Q&A. Why is the music playing? Hold on.

The music stops. The QR code slide remains on screen. A second video feed of a man, "Evan", appears in the bottom left.] Uh, you can find me pretty much anywhere at Jamey Gannon. Um, if you also head to like my, my link tree thing, uh, brand-sprints.com/links, you can find loads and loads of resources, whether it's my course, templates, um, I have like 25% off discounts to Flora and pretty much all the AI tools. Um, and of course, you can keep up with me on social or maybe we can even work together. Um, yeah, I guess I have five minutes to do a Q&A.

Evan: Amazing, Jamie.

Jamey Gannon: Um, I'm just reading some of these questions out. How do you deal with the negative feedback from social when they figure out it's AI? As a creator, I don't care. As a brand owner, that's something that you need to to think about and have a comm strategy for.

Evan: Yeah.

Jamey Gannon: Um, the name of the platform is Flora. That was the one I was using for the most part. And then the AI UGC one is Pixed Soul. Um, where would I recommend to start for beginners using doing AI image generation? Um, I'd follow me on X. Um, and then just start doing stuff. Like try and find a simple task, um, maybe like one of the examples I did today. Um, and just start small. I think AI is just like any other tool. So, you know, you didn't learn Google Ads or Illustrator, um, or CapCut in a day. You know, it took you time. Uh, you made a short video first, one logo, uh, you know, you did your first campaign. Um, so yeah, a couple hours a week, just start getting into it and then slowly you'll start to see more use cases for it.

A question from "Jose Garcia" appears in a purple banner at the bottom of the screen: "Some of the biggest issues when generating AI assets at scale is that, when the product is a little bit tricky, the agent will not do a very good job at maintaining the product consistency over time. Is there a way to prevent that?"] Evan: Jamie, I'll pull one final one up from and put it on the screen here. From Jose. So it seems like he's he's using it quite a bit already, but some of the biggest issues when generating AI assets at scale is when the product's a little bit tricky, the agent will not do a very good job at maintaining the product consistency over time. Is there a way to prevent it?

Jamey Gannon: Yeah, I think there's specific tools that are starting to get better at this where you basically train a mini model on your product. Um, but as good as AI is, like it definitely isn't magic. Um, there was like one example I did, I did like a three-hour version of this workshop and someone, one of the students, of course, had a like optical illusion checker print toilet paper holder. And I was like, dude, I don't know. Um, and then we kind of we kind of worked through it like telling the model it's a toilet paper holder, um, creating like a shot list, like there's something you can do where you have, uh, you ask Nano Banana or you take a photo of it, um, in like a bunch of different angles and then every single time you like are using that as a reference image, it has like all of those. Um, I think it'll get better over time for even more complex things. But if like, like even that image that she gave me of the checkerboard toilet paper holder, like I wasn't even totally sure what it is. So I'll say like, like yes, AI is very intelligent, but it's not going to like read your mind. Um, so you might just need to like have a very long prompt, have more reference images. Um, also, I Photoshop so much. Like I, um, one of another one of my principles is use the right tool for the right job. And we're still at a point where like sometimes to like edit a logo or add a fingernail or something that gets lost, especially really tiny stuff, you just got to have to Photoshop it. Um, now if you're if you want to try and generate hundreds of images in a day, totally different thing, but if you're trying to make like really high quality images, um, you know, going into Photoshop is there's nothing wrong with that. It doesn't mean that you're doing it wrong. That's just how things work right now.

Evan: Huge. And Jamie, any final words you want to leave with the audience today?

Jamey Gannon: Uh, persevere. This stuff is really hard. People want you to believe that it's not hard. People that sell courses like me, they want you to think it's very easy to keep generating stuff. The models want you to think it's very easy to keep generating stuff, but it is a tool and a skill just like any other. And it will take you time.

Evan: Amazing.

Jamey Gannon: And you'll get better at it. And then you'll be able to do it in 20 minutes.

Evan: Amazing. Thank you so much, Jamie. That was an incredible presentation. I think everyone here got a ton of value out of it. Um, and yeah, we'll see you on the next one.

Jamey Gannon: Yeah.

Evan: Thanks, Jamie.

Jamey Gannon: Bye.

Evan: Bye.