TOOLS Twenty of them, and four gaps

Research Design Build Camera Four named gaps

Which question each method answers

20 Methods, each one run on real work
06 Case studies they ran on
04 Gaps, named at the bottom of this page

Filter by discipline

Methods

  1. The default. Watch somebody attempt the real task, count where they stop. Remote over Zoom on the McDonald’s study, in person on ioMoVo and Toronto Cupcakes.

  2. Fast, cheap, and best used after a live session rather than before one. An audit tells you what violates a principle, not what stops a person.

  3. For reasoning at a scale interviews cannot reach. Useful for how often and how many. Unreliable for why, without a session behind it.

  4. Turns a recommendation into a measurement. Only worth setting up once you have named the threshold that counts as success.

  5. Google Analytics, to check whether a qualitative finding happens often enough to prioritise. Frequency, not cause.

  6. Recordings reviewed afterward so the same observation is counted the same way for every participant. It is what turns eight anecdotes into seven of eight.

  7. Writing the task so it does not contain its own answer. The hardest part of the whole method and the least discussed.

  8. Naming things the way people ask for them instead of the way the database stores them. Three of the six case studies turned out to be a naming problem that had been filed as a navigation problem.

  9. Clickable enough to test, not so polished that feedback becomes about the visuals. Mobile and web flows kept consistent across both.

  10. Type scale, spacing scale, one accent, states defined once. This site runs on one twelve-column grid, three measures, three breakpoints, and an eight-step spacing scale; there is exactly one grid declaration in the stylesheet and no page sets a width in a style attribute.

  11. Contrast checked against WCAG AA, full keyboard operability, visible focus, reduced motion respected. A requirement, not a pass at the end.

  12. Plain language in the interface, especially in errors. Most confusing UI is a labelling problem wearing a layout costume.

  13. Semantic markup, modern CSS, vanilla JavaScript. This site carries a little over six kilobytes of script, and no framework runs in the browser.

  14. Deployed on Netlify, with edge functions and serverless endpoints where a page needs something the browser cannot do alone.

  15. When an interaction depends on timing or momentum, a coded prototype answers the question and a Figma frame guesses at it.

  16. Specifications in the units engineering uses, with states and edge cases written down. Knowing what is expensive to build changes which design I argue for.

  17. Documentary, automotive, and brand work. Handheld and locked-off, available light first.

  18. Portrait, editorial, event, and product. Short sessions, because people stop being themselves after about forty minutes.

  19. Grade and cut in house. One revision round inside the agreed turnaround, delivered at the sizes the destination needs.

  20. Getting somebody comfortable enough to be themselves on camera is the same skill as getting a participant to think out loud in a session.

Limits

What I am still bad at

Large-n quant
I am confident with eight participants and a coded transcript. Statistical significance on a sample of ten thousand is a different discipline and I bring somebody in for it.
Motion design
I can specify and code an interaction. I do not build illustrated animation, and I stop before pretending otherwise.
Back end
Comfortable with static hosting, serverless endpoints, and APIs. Database architecture and scaling are not mine.
Scoping too wide
My instinct on a new brief is to research more than the decision requires. I now write the decision the study has to serve at the top of every plan, and cut back to it.