اجرای usability test
طراحی تست، جمعآوری داده و اصلاح UX.
قدم اول: Planning و objectives
Usability test بدون objective مشخص، فقط demo میشود و نتیجهاش «جالب بود» است نه actionable insight. objective written برای پروژه سارا: «آیا user با profile سارا میتواند بدون کمک moderator پروژه با deadline مشخص اضافه کند و در week view پیدا کند؟» secondary: «FAB و week navigation قابل discover هستند؟» هر objective باید yes/no یا metric قابل measure داشته باشد.
Metrics را upfront تعریف کنید: task completion rate (چند از ۵ نفر succeed)، time on task (median)، error count (misclick، wrong path، validation fail)، و optional SUS questionnaire بعد session. qualitative «why» از think-aloud میآید — quantitative از spreadsheet.
Method: moderated remote با Zoom/Google Meet + Figma prototype share معمولاً برای Nova Lab کافی است. ۴۵ دقیقه session: ۵ intro، ۲۵ tasks، ۵ post-task questions، ۱۰ interview. unmoderated (Maze) بعداً برای scale — start moderated برای depth.
Test plan document: objectives، recruit criteria، script، task list، success/fail rubric، consent form، data storage policy. یک moderator، یک note-taker ideal — اگر solo، record video و بعد review.
Recruit ۵ participant: match persona ± variance (freelancer، multi-project، mobile daily). student peers for learning OK با brief persona read beforehand. incentive gift card یا swap favor. schedule buffer ۱۵ دقیقه بین session برای notes.
قدم دوم: Test script و tasks
Script structure: welcome → purpose (test product not you) → think aloud → consent record → prototype access check → tasks → wrap-up. language calm، neutral. «اگر گیر کردید طبیعی است — ما میخواهیم بفهمیم product کجا گیر میدهد.»
Task 1 warm-up (easy): «app را باز کن و بگو dashboard چه اطلاعاتی نشان میدهد.» success: describe list/empty. Task 2 core: «پروژه جدید برای client فرضی با deadline پنجشنبه اضافه کن.» success: item on dashboard. Task 3: «همان deadline را در نمای هفته پیدا کن.» Task 4 optional: edit deadline یا delete — اگر scope دارید.
Wording neutral critical: بگویید «پروژه اضافه کن» نه «روی دکمه + بزن». leading wording validity test را destroy میکند. order tasks از easy به hard. same scenario data across users برای compare.
Post-task بعد هر task: «confidence ۱–۵» و «چی سخت بود؟» Post-study ۵ دقیقه open: «چیز دیگری که expected داشتی؟» «با چه ابزاری الان این کار را میکنی؟» — competitive insight.
Pilot: یک dry run با همکلاسی قبل real sessions. timer script، fix awkward phrasing. pilot participant in final report نیست ولی script refine میشود.
قدم سوم: Moderation techniques
Moderator neutral facilitator است، نه defender of design. وقتی user «این confusing است» میگوید، reply «ممنون، ادامه بده» نه «ولی ما فکر کردیم obvious است». emotional neutrality trust میسازد و data بهتر. قبل session با designer هماهنگ کنید که feedback personal نیست — product under test است.
Silence tool powerful است. بعد سؤال «چه کار میکنی؟» ۵–۱۰ ثانیه صبر کنید. urge to help را resist کنید. اگر ۲ دقیقه stuck: «اگر alone بودی چه try میکردی؟» — still not «click FAB».
Abort criteria: same task بیش از ۳–۴ دقیقه بدون progress → «بیاییم task بعدی»، mark fail، note reason. participant حق دارد skip. abort data هم valuable است — نشان میدهد blocker severe است.
Note-taker timestamp issues: «03:22 FAB scroll off screen»، «05:10 date picker — selected wrong month». اگر solo moderator، lightweight shorthand بین tasks بنویسید؛ video بعداً review. دو نفر = یک نفر eye contact participant، یک نفر notes.
Tech checklist: prototype link incognito test، screen record permission، backup link، participant phone hotspot if wifi bad. ۲ دقیقه اول session برای «صدای من را میشنوی؟ prototype load شد؟»
قدم چهارم: Analysis و synthesis
Within ۲۴ ساعت synthesis شروع کنید — memory fade میشود و detail از دست میرود. list raw observations per participant بدون judgment. سپس affinity cluster: «FAB visibility»، «date picker»، «terminology»، «week view density». pattern ۳/۵ users = serious؛ ۱/۵ = maybe outlier unless P0 blocker مثل crash یا complete inability to finish.
Rainbow spreadsheet: rows = tasks، columns = P1–P5، cells pass/fail/partial + time. one glance completion rate. color code fail cells. median time task 2 compared to target (مثلاً under 90 sec).
Severity = frequency × impact. FAB missed 3/5 → high frequency، blocks core task → high impact → P0 fix. typo label 2/5 → medium. prioritize با team در ۳۰ دقیقه meeting.
Video clips ۲۰–۴۰ ثانیه برای stakeholder empathy — face optional، screen + audio enough. clip «FAB miss» در report link میشود. quotes verbatim در findings: «فکر کردم باید از settings add کنم.»
Separate symptom از cause: «date wrong» ممکن است cause «picker UI» یا «mental model Friday = end of week». think-aloud transcript helps root cause. recommendation باید cause address کند.
قدم پنجم: Report و action items
Report structure Nova Lab: executive summary (۱ paragraph findings + top ۳ fixes)، methodology (who, when, prototype version)، findings با format Observation — Evidence — Recommendation، prioritized backlog، appendix (script, clips, spreadsheet).
Example finding: Observation — «۳ از ۵ participant FAB را در اولین ۳۰ ثانیه پیدا نکردند.» Evidence — timestamps video، quotes. Recommendation — «Extended FAB با label متن «پروژه»؛ first-visit tooltip one-time.» هر finding actionable.
Prioritize fixes: RICE یا simple effort/impact matrix. quick wins (label add) در sprint این هفته؛ redesign date picker next iteration. attach Figma link frames annotated before/after. dev و design در same doc comment.
Share report با stakeholder non-design: executive summary کافی؛ full doc for team. presentation ۱۵ دقیقه: ۱ clip، ۱ quote، ۱ metric before (completion ۲/۵ = 40%). honesty about small sample — «directional not statistical».
Action items با owner و due date: «Ali — FAB label — Wed»، «Sara persona doc update frustration FAB — Fri». بدون owner، report archive میشود.
قدم ششم: Iterate، re-test و continuous discovery
Implement top ۳ fixes در prototype v2 — نه کل redesign یکباره. validate fix regression: task که pass بود هنوز pass. unmoderated Maze با ۱۰ user optional برای quantitative check completion lift.
Re-test subset: همان task core با ۳ user جدید (نه same ۵ — learning effect). compare completion 40% → target 80%+. اگر lift نیامد، hypothesis wrong بود — deeper interview یک user.
Update persona و problem statement اگر finding جدید: سارا frustration «FAB پیدا نمیشود» add. living research cycle: launch → test → fix → test → minor release → repeat.
Continuous discovery: monthly mini-test ۳ user، ۲ task، ۱ ساعت synthesis — cheaper than big bang every ۶ month. metric dashboard بعد real launch: time to add project، retention week 1.
Celebrate و document wins: «completion 60% → 100% after FAB label» در portfolio Nova Lab case study. lessons learned: what we'd do differently next test (pilot earlier، shorter tasks، recruit closer to persona). UX maturity = habit test، not one hero study — تیمهای قوی هر sprint یا release کوچک یک touchpoint research دارند.
برای گزارش نهایی course، timeline visual بسازید: persona → wireframe v1 → test 1 → fix → prototype v2 → test 2. stakeholder خارج از UX با این narrative سریعتر value research را میفهمند و budget iteration بعدی را راحتتر approve میکنند.
سوالات متداول
Moderated یا unmoderated — کدام اول؟
Moderated برای why و deep insight — ۵ user qualitative کافی است early stage. Unmoderated (Maze, UserTesting) برای scale و pattern quantitative بعد fix اولیه. Nova Lab path: moderated → fix → optional unmoderated validate.
۵ participant واقعاً enough است؟
Nielsen research نشان میدهد ~۸۵٪ usability issues با ۵ user کشف میشود برای one iteration. بیشتر بدون fix بین batch diminishing returns دارد. better: fix → retest ۳–۵ user جدید.
Prototype test کنیم یا product live؟
Prototype زودتر، ارزانتر، قبل code. Live beta برای production constraints (performance, real data, edge cases). lifecycle: prototype test → build MVP → beta test → monitor analytics.
Designer میتواند moderator باشد؟
ممکن است ولی bias risk — defend design unconsciously. separate moderator ideal. اگر solo: script strict، pause before respond، document «I wanted to help at 4:32». participant recruitment خارج از immediate team.