Three Lowboy mobile screens on a plum background, including recipe recording, onboarding, and an empty dinner menu

CASE STUDY · 2026

Lowboy: Conducting AI

ROLE
Product design · AI direction · Prototyping
TIMELINE
Ongoing · 2026
TEAM
Two-person founding team

Lowboy became my test for a different way of working. I set the direction and made the calls. AI helped me explore flows, build systems, document decisions and keep the project moving with a friend.

AT A GLANCE

5

USER-FLOW VERSIONS

4

SPEC PROBLEMS FOUND

46+

SYSTEM COMPONENTS

WHAT WE MADE TOGETHER

We made the product and kept the thinking visible.

I wanted to know whether AI could help two people hold a product this complicated without losing the thread. The screen count was beside the point.

MY FINAL LOWBOY SCREENS

SETTING THE DIRECTION

I was the conductor. AI played different parts.

I chose the problem and set the constraints. Opus helped me reason through the system. Fable pushed the visual work further. My friend and I made the product calls.

Division of labour

I made the final calls.

I brought the kitchen context and references, then decided what felt right, what needed another pass and what did not belong.

AI gave me range and kept track of the mechanical work. It generated routes, built components and surfaced contradictions I might have missed.

How a change moved through the work

Ideate, document, build

Every change went through the same three acts, in that order. Skipping the middle one is how decisions get lost.

The loop:

Ideate: stay in conversation while changing direction is cheap.

Document: write decisions into Linear and FigJam before building.

Build: create, inspect, measure and revise the artifacts.

FIVE PASSES AT THE FLOW

Each model gave me a different map.

V1 and V2 came from Opus 4.8. Opus 5 produced V3. Fable produced V4 and V5, which were much easier to follow. I still had to decide what was actually better.

V1 TO V5 · USER FLOW ARCHIVE
What each model gave me

Three models, three different maps

I ran the same flow through three models and kept five versions. Each one was easier to follow, but I still had to decide whether it worked.

The passes:

V1 and V2 · Opus 4.8: the first usable structure.

V3 · Opus 5: more product logic held together.

V4 and V5 · Fable: a clear jump in visual organization.

AI’S SCREENS, MY DECISIONS

Fast output gave me something to argue with.

The AI files are untouched. Some ideas were useful; some were generic or simply wrong. I checked them against the kitchen rules, kept what held up and rebuilt the rest.

UNTOUCHED AI EXPLORATION · MOSTLY FABLE
MY AUTHORED LOWBOY DIRECTION
Keep, change, own

What I kept, what I changed, and what stayed mine

Fast output is only useful if you are willing to throw most of it away.

The three piles:

Keep: fast variations, useful starting points and states I had not considered.

Change: generic hierarchy, questionable logic and layers that looked finished before they were sound.

Own: the product direction, visual language and every final decision.

BUILDING THE SYSTEM

Opus built the first system. Fable and I tried to break it.

Opus 5 built the first Figma foundations and components. Fable audited the library. I checked every component, fixed the wonky layers and made the final calls.

COLOUR TOKENS + CONTRAST AUDIT
COST-PILL RULES + STATES

48+

COLOUR VARIABLES

12

TEXT STYLES

2

THEMES, FROM DAY ONE

AI AS PRODUCT MANAGER

My morning brief changes with my energy.

I start the Linear review myself. AI pulls together open tickets, what I finished yesterday and the blockers I recorded. On slow mornings I ask for easy wins first, then move into the heavier work once I have momentum.

The working loop

Capture, review, start, record

The same four steps every morning, whatever the energy level.

The loop:

Capture: meetings, ideas and decisions.

Review: open tickets and yesterday’s work.

Start: easy wins or the hardest item.

Record: progress, decisions and blockers.

WHAT ALMOST BROKE IT

AI can make more than a person can responsibly review.

That was the real risk. Screens looked finished before the logic was. Some components needed cleanup. A cleaner flow turned out to be worse. I had to slow the machine down and check the work.

Three ways it could have gone wrong

Polish and volume can both hide bad work.

Each of these cost me time before I learned to check for it.

What I watch for now:

Plausible ≠ correct: a polished screen can still hide a bad rule.

Volume ≠ progress: more output creates more review work.

One audit is not enough: metrics and screenshots catch different defects.

WHAT MADE IT WORK

Build it. Inspect it. Measure it. Keep it, fix it or throw it away.

The collaboration worked because AI was allowed to disagree, but never allowed to quietly become the decision-maker.

01

STAY IN PLANNING

Do not build while the important questions are still moving.

02

GIVE FULL CONTEXT

The detail that looks irrelevant usually changes the design.

03

LET IT PUSH BACK

A useful collaborator should be able to explain why an idea is weak.

04

VERIFY EVERYTHING

Check the picture and the numbers before calling it done.

86-SCREEN LOWFED COVERAGE BOARD · DETAIL

LOOKING BACK

I would not hand the product to AI.

The useful part is being able to explore more directions without becoming attached to the first one. AI can play a lot of instruments. I still have to know what I am listening for.

IF YOU GOT THIS FAR...

Thank you.