+ Carefully separates evidence from items needing confirmation.
- Its conclusion on whether there is exactly one primary action is muted.
States plainly what cannot be judged from a description at all.
| Category | Design › Product design |
|---|---|
| Tags | ReviewingAnalyzingQuestion generation |
Review this screen for usability problems. ***From my description — say what you cannot judge.*** **Check:** 1. **What is the screen for?** ***If I cannot state it in one sentence from your reading of my description, that is the first finding*** 2. **The primary action** — ***is there exactly one, and is it obvious?*** **Two equally weighted actions means the user decides, and deciding is the cost** 3. **What the user has to know before they can act**, and whether the screen tells them 4. **Where they have to remember something from a previous screen** — ***recall is more expensive than recognition*** 5. **Irreversible actions** — ***is there a way back, and does the user know before they commit*** 6. **Where the wording assumes internal vocabulary** 7. **What happens to someone who is not the intended user** — *arrived by search, arrived confused, arrived on a phone at 2am* **Then:** - ***The one change that would help most***, and what it costs - **Findings I am confident about** versus ***findings I would need to see the screen to confirm.* **Keep them separate** - ***What I cannot judge from a description at all*** — **visual hierarchy, spacing, contrast, actual reading order.** **Say so rather than guessing** ⚠️ ***You cannot see the screen. Work only from what I described*** — **do not assume a layout, a component, or a platform convention I did not state.** Where my description is not enough, **say what you would need to see.**
Reviewing a screen in text tempts a model into judging hierarchy and spacing it cannot see. This draws that line and separates confident findings from ones needing the actual screen.
ChatGPT best respects the evidence and its limits. Gemini is detailed but overstates some assumptions and runs long; [C] was not provided.
+ Carefully separates evidence from items needing confirmation.
- Its conclusion on whether there is exactly one primary action is muted.
+ Covers every requested area with concrete detail.
- Treats some unstated UI behavior and missing guidance as facts.
| Criterion | ChatGPT | Gemini | Leader |
|---|---|---|---|
| Instruction following | 9 | 9 | Tie |
| Accuracy | 9 | 6 | ChatGPT +50% |
| Specificity | 9 | 9 | Tie |
| Structure | 9 | 9 | Tie |
| Right length | 9 | 6 | ChatGPT +50% |
Scored 1–10 by gpt-5.6-sol with model names hidden (2026-09-25). This is an AI review, not a measurement.
We gave three models the same input and copied their answers unedited. Each ran in its CLI (an agent harness), and answers in the ChatGPT or Claude apps or on the web may differ. Outputs are in Korean.
The screen: 회원가입 완료 후 뜨는 '프로필 설정' 화면. 상단에 프로필 사진 업로드용 원형 아이콘, 그 아래 닉네임 입력창, 관심사 태그 선택(20개 중 최대 5개), 하단에 '다음' 버튼과 '건너뛰기' 텍스트 링크가 있음. '다음' 버튼은 관심사를 1개 이상 선택해야 활성화됨. Who uses it and when: 20대 대학생 대상 스터디 매칭 앱. 회원가입 직후 온보딩 3단계 중 2단계이며 대부분 모바일 세로 모드로 접속. 앱을 처음 켠 사용자라 서비스 용어에 아직 익숙하지 않음.
“스터디 매칭에 사용할 프로필 정보와 관심사를 설정하고 다음 온보딩 단계로 이동하는 화면”이라고 한 문장으로 이해할 수 있습니다.
의도된 주요 행동은 다음으로 보입니다. 건너뛰기가 텍스트 링크라면 보조 행동으로 설계된 것으로 보입니다.
다만 관심사를 1개 이상 선택해야 다음이 활성화되므로, 사용자는 먼저 “왜 관심사를 선택해야 하는가?”를 이해해야 합니다. 설명 없이 버튼이 비활성화되어 있다면 무엇이 부족한지 즉시 알기 어렵습니다.
또한 건너뛰기가 존재하므로 사용자는 다음 중 하나를 결정해야 합니다.
두 행동의 시각적 우선순위가 실제로 충분히 구분되는지는 설명만으로 판단할 수 없습니다.
사용자는 최소한 다음을 알아야 합니다.
건너뛰기 후 이 화면으로 돌아올 수 있는지현재 설명에는 이 정보가 없습니다. 특히 다음이 관심사 1개 선택 전에는 비활성화되므로, 그 조건을 화면이 안내하는지가 중요합니다.
이 화면 자체에서 이전 화면의 정보를 기억해야 하는 부분은 많지 않습니다. 다만 사용자는 회원가입 때 정한 닉네임이나 계정 정보와 별도로 여기서 다시 닉네임을 정해야 하는지 혼란스러울 수 있습니다.
“닉네임”이 가입 과정에서 이미 입력된 값이라면 자동으로 보여주거나, 왜 다시 입력하는지 설명해야 합니다.
프로필 사진, 닉네임, 관심사 설정 자체가 되돌릴 수 없는지는 설명만으로 알 수 없습니다.
특히 건너뛰기를 누르면 설정을 나중에 할 수 있는지, 다음을 누르면 수정이 제한되는지 안내가 필요합니다. 되돌릴 수 없다면 커밋 전에 알려야 합니다. 수정 가능하다면 “나중에 프로필에서 변경할 수 있어요” 같은 문구가 불안을 줄입니다.
관심사 태그는 앱 내부 용어일 가능성이 있습니다. 처음 사용하는 사람에게는 다음이 모호할 수 있습니다.
“관심사를 선택하면 맞는 스터디를 추천해드려요”처럼 결과 중심으로 설명하는 편이 이해하기 쉽습니다.
다음이 비활성화되는 이유와 관심사 선택의 목적을 화면에 명시하는 것이 가장 효과적입니다.
예:
“관심사를 1개 이상 선택하면 나에게 맞는 스터디를 추천해드려요. 최대 5개까지 선택할 수 있어요.”
가능하다면 건너뛰기 옆에도 “나중에 프로필에서 설정할 수 있어요”를 덧붙이세요.
비용은 설명 문구가 추가되어 화면이 조금 더 복잡해지고, 앱의 추천 로직을 사용자에게 약속해야 한다는 점입니다. 하지만 현재의 비활성화 상태를 추측하게 만드는 비용보다 작습니다.
다음은 설명만으로 확정할 수 없습니다.
다음과 건너뛰기 중 무엇이 실제로 더 눈에 띄는지다음이 충분히 비활성 상태로 보이는지설명만으로는 실제 시각적 계층, 여백, 색상 대비, 정렬, 컴포넌트 배치, 화면 높이에 따른 스크롤 여부를 판단할 수 없습니다. 이를 확인하려면 화면 캡처나 실제 화면을 봐야 합니다.
| 7v7 Football Team Generator App | |
| Accessibility Auditor | |
| Accessibility Auditor Agent Role | |
| Accessibility Expert | |
| Accessibility Testing Superpower |