Piktochart · Case Study

AI Prompt Enhancer

Context

Piktochart could make the image.
Users couldn't ask for it.

Piktochart is an AI design platform for non-designers. 34M+ users worldwide have made 220M+ visuals: infographics, presentations, reports, social graphics, and AI images.

The generator was fine. The prompting wasn't. Most users aren't prompt writers. They'd type something rough, get a flat image, then go find a prompt tool somewhere else and paste the result back.

Case Study

An in-app image prompt enhancer that gets users a better image on the first try, without leaving Piktochart

RoleFeature OwnerProduct · Design · Prompt engineering
Team3 Super IC's1 Designer & PM (Me) • 2 Developers
Duration1.5 monthsIdea to ship
YearMar, 2026
The Problem

Users were voting with their feet

01

Generate in Piktochart

Type a prompt, hit generate.

02

Hit friction

Mediocre result, unclear why.

03

Leave Piktochart

Use another tool's prompt enhancer.

04

Paste back to generate

Return only for the image.

The signalThey liked us for the generating. They went elsewhere for the words. Every trip out the door was a chance not to come back.
Research & Insight

Not a feature request.
A retention leak.

Nobody asked for a prompt tool. They wanted one less step. They knew their prompts were weak, they just couldn't fix them alone, and guessing felt worse than switching tabs.

The realization

  • The pattern was repeatableGenerate → friction → leave → paste back
  • The fix was structuralBring the prompting in-house
  • The prize was the whole loopIdea, refine, generate, without leaving
The Goal

Keep them here.
Give them a better image.

For users

  • Skip the guesswork
  • One toggle, several clear directions
  • Better images on the first try
  • No reason to open another tab

For business

  • Keep the users we were losing
  • Make each generation worth more
  • Own the whole workflow, not just the output
  • Something competitors don't have
The Solution

One toggle.
Three modes.
Better image.

Off

Uses the original prompt without changes.

On-Auto

Improves the prompt and generates straight away.

On-Review For power users

For people who want more control. It reads what the prompt is going for and offers 10 rewrites. Any of them can be pushed further:

StyleFramingLighting & Atmosphere
AI Image Generator with the Prompt Enhancer dropdown: Off, On-Auto, On-Review
AI Image Generator · Off / On-Auto / On-Review
How It Works

From a vague idea to a usable prompt

On-Review
1
Detect intent

It reads what the prompt is going for, then writes 10 versions of it. Each one gets a short title and a few tags.

2
Pick a direction

Mood Focused, Action & Drama, Artistic & Minimal, and so on. The full text is one tap away.

3
Customize the prompt

Six styles to pick from, plus Framing and Lighting. Change any of them and the prompt rewrites itself.

4
Generate

One click. No prompt writing needed.

On-Auto

Skips all of this. It picks the closest match and generates straight away.

Review Enhanced Prompt: 10 enhanced prompts with titles and attribute tags
Review · pick a direction
Customize Prompt: Style, Framing, and Lighting and Atmosphere controls
Customize · Style / Framing / Lighting & Atmosphere
Under the hood

Two model calls, one strict contract

The model doesn't return prose. It returns an interface. Every field is something you can see, so I wrote the schema and designed the screen at the same time.

1
Enhance

User idea in. Ten suggestions out, each with a title, tags, and a generation-ready prompt.

2
Customize

The chosen prompt plus the user's Style, Framing and Lighting picks. One refined prompt out.

3
Version

Every edit links back to the one before it, so users can walk back through their changes.

Runs on Gemini
JSON field maps to interface
titleCard heading
visual_attributesTags under each title
styleSix-option picker static
dynamic_fieldsFraming and Lighting inferred
full_promptSent to the image model
Noun phrases only 3 to 6 words No verbs or articles Style never leaks into dynamic fields Vague input infers, never fails
Design Deep Dive

How do you judge an image
you can't see yet?

Two ways to solve it. I bet on the cheaper one.

Option A ×

Preview image per suggestion

Slower, pricier, more API calls. All for a problem users might not have.

Option B · shipped

Trust structured text

Short titles, tags for the visual bits, the full prompt on tap. Cheap and fast, and people got it straight away.

Before: the raw prompt a user types
Before · raw prompt
After: a scannable title with visual attribute tags
After · title + attribute tags
The surpriseNobody asked for previews. People scanned, picked, and generated. Structure did the work the images would have done.
Impact

They found it,
and they finished.

16K+

people tried it in the first month

96%

turned it on and finished a generation

1st try

better results, so fewer second attempts

Mixpanel: unique users and adoption
Mixpanel · adoption
Mixpanel: task completion funnel
Mixpanel · completion funnel
The real winFewer generations looks like disengagement. Here it was the opposite. People got what they wanted first time and stopped leaving.
Why It Worked

Three assumptions that turned out right

01

Text-only was enough

People read visual direction straight from text. No previews, no image cost, and it shipped faster than planned.

02

The model handled rough prompts

Tight instructions meant even a lazy prompt came back with ten usable directions. Nobody had to write well to get something good.

03

It kept users in

One toggle, ten suggestions, one click. People stopped opening other tabs. That was the whole point.

Launch

Promoted in public, on Piktochart's own channels

Piktochart promoted the feature on its own Instagram channel:

Learnings

What I'd carry into the next AI feature

a.

Write the instructions like a spec

The tighter I wrote the rules on structure, tone and wording, the more predictable the model got. Most of the design work was in the prompt.

b.

"One-time use" can be a win

Repeat use dropped because the first try worked. A number only means something next to the goal.

c.

Text-first UX is underrated

Structured text beat image previews on speed and cost, and people understood it just as well.

What's next

  • Design ReviewAI feedback on existing work
  • AI CopywriterText suggestions, not just images
  • TemplatizerTemplates matched to intent
  • Deeper customization UXUsers under-explored it
Piktochart · AI Prompt Enhancer

Thanks for reading.

Happy to go deeper on any of it: the decisions, the prompt work, or the numbers.

The Prompt Enhancer in action
Anil Kumar Giridhar Senior Product Designer · anilkg.design