+ Lists meta-rules and correctly finds no order conflict.
- Mislabels live rules as dead and sets some priorities itself.
Catches conflicts of degree too, and refuses to choose when both sides are requirements.
| Category | Using AI › Context management |
|---|---|
| Tags | AnalyzingReviewingChecklist |
My instructions are contradicting each other. **Find the conflicts and tell me which to drop.** 1. **List every instruction I have given**, including ones implied by examples I provided 2. **Find the conflicts:** - **Direct** — two rules cannot both hold - **By degree** — "be thorough" and "keep it short" both hold until they do not. ***Say where the line falls*** - **Conditional** — they only conflict in certain cases. **Name those cases** - **Order** — the output cannot satisfy both orderings 3. **Per conflict: which instruction is load-bearing and which is a preference.** *A preference yields to a requirement* 4. **Find instructions that have become dead** — the task moved and they no longer apply to anything **Then:** - **The resolved instruction set**, as something I can paste - ***What I lose by each resolution.*** **A conflict resolved is something given up — say what** - **Conflicts I have to resolve myself** because both sides are requirements. ***Do not pick for me on those*** - **Which conflict was most likely causing the behavior I am seeing**
Adding instructions creates pairs like "be thorough" and "keep it short" that coexist until they do not. This classifies each conflict, separates requirements from preferences, and leaves the real trade-offs to you.
ChatGPT is the most complete and logically sound. Gemini is specific but overclaims and resolves required choices; [C] was not provided.
+ Lists meta-rules and correctly finds no order conflict.
- Mislabels live rules as dead and sets some priorities itself.
+ Connects the symptoms to rules 1, 2, 3, and 5.
- Assumes impossibility and generation order, then chooses for the user.
| Criterion | ChatGPT | Gemini | Leader |
|---|---|---|---|
| Instruction following | 9 | 6 | ChatGPT +50% |
| Accuracy | 8 | 4 | ChatGPT +100% |
| Specificity | 9 | 9 | Tie |
| Structure | 9 | 8 | ChatGPT +13% |
| Right length | 8 | 5 | ChatGPT +60% |
Scored 1–10 by gpt-5.6-sol with model names hidden (2026-09-25). This is an AI review, not a measurement.
We gave three models the same input and copied their answers unedited. Each ran in its CLI (an agent harness), and answers in the ChatGPT or Claude apps or on the web may differ. Outputs are in Korean.
My instructions: 1) 답변은 항상 3문장 이내로 짧게 써줘 2) 근거는 반드시 출처를 각주로 달아줘 3) 전문 용어는 쓰지 말고 초등학생도 이해할 수 있게 풀어서 써줘 4) 이전 답변들이랑 톤과 형식을 계속 유지해줘 5) 내가 따로 말 안 해도 관련된 배경지식은 알아서 덧붙여줘 What is going wrong: 출처를 달면 항상 5문장을 넘어가고, 용어를 쉽게 풀어 쓰면 문장이 길어져서 3문장 제한을 못 지켜요. 배경지식까지 덧붙이라고 하니 아예 출처를 생략하고 넘어가는 답변도 나와요.
| 충돌 | 유형 | 중요한 쪽 | 선호 쪽 | 버릴 것과 잃는 것 |
|---|---|---|---|---|
| “항상 3문장 이내” ↔ 이번처럼 모든 충돌과 손실을 자세히 설명 | 직접 | 이번 요청의 완전한 분석 | 3문장 제한 | 3문장 제한을 버리면 짧음과 한눈에 보는 편리함을 잃음 |
| “철저히 설명” ↔ “짧게 쓰기” | 정도 | 필요한 정보 전달 | 짧음 | 짧음을 우선하면 예외·손실 설명이 빠짐 |
| 각주 달기 ↔ 3문장 제한 | 조건 | 근거가 중요한 답변에서는 각주 | 짧음 | 각주를 생략하면 근거 확인이 어려움 |
| 쉬운 말 쓰기 ↔ 3문장 제한 | 정도 | 이해 가능성 | 짧음 | 쉬운 설명을 줄이면 독자가 뜻을 놓칠 수 있음 |
| 배경지식 덧붙이기 ↔ 3문장 제한 | 조건 | 배경지식이 판단에 꼭 필요할 때만 추가 | 항상 추가 | 배경지식을 줄이면 초보자에게 맥락이 부족함 |
| 이전 형식 유지 ↔ 이번 요청의 표·분류 형식 | 조건 | 이번 요청의 명시적 형식 | 이전 형식 | 이전 형식을 버리면 일관성이 줄어듦 |
| “모든 것을 말하라” ↔ “3문장 이내” | 직접 | 둘 다 요구사항이라면 사용자가 선택해야 함 | — | 대신 결정하면 사용자의 우선순위를 임의로 정하게 됨 |
정도 충돌의 선은 “짧게 써도 핵심 판단·조건·손실을 빠뜨리지 않는 지점”입니다. 이 답변처럼 요구된 항목이 여러 개면 3문장은 이미 선을 넘으므로, 상세 분석을 위해 제한을 풀어야 합니다.
순서 충돌은 실제로 없습니다. “각주를 달라”는 배치 위치를 정하지 않았고, “이전 형식 유지”도 구체적인 순서를 제시하지 않았습니다.
답변은 핵심을 먼저 말하고, 필요한 만큼만 자세히 쓴다.
요구된 항목을 빠뜨리지 않는 것이 짧게 쓰는 것보다 우선한다.
근거가 필요한 주장은 출처를 각주로 단다.
전문 용어 대신 쉬운 말을 쓰되, 정확성을 해치지 않는다.
배경지식은 이해나 판단에 필요한 경우에만 덧붙인다.
이전 답변의 형식은 현재 요청의 명시적 형식과 충돌하지 않을 때만 유지한다.
두 요구사항이 모두 필수라서 선택이 필요한 경우에는 사용자에게 결정하도록 한다.
“항상 3문장 이내”와 “모든 충돌·손실·조건을 빠짐없이 설명”이 모두 절대적인 요구라면 둘 다 지킬 수 없습니다. 어느 쪽을 우선할지는 제가 정하면 안 되며, 이번 답변에서는 현재의 상세 분석 요구를 우선한 것으로 해석했습니다.
가장 문제를 일으킨 충돌은 “항상 3문장 이내”와 “쉬운 말·각주·배경지식 추가”의 결합입니다. 특히 배경지식을 강제로 넣으면서도 짧게 쓰라고 하면, 모델이 문장을 줄이기 위해 각주나 설명을 빼는 현상이 생깁니다. [^2]
[^1]: 출처: 사용자가 이번 메시지에서 제시한 지시와 “내가 준 지시들” 목록.
[^2]: 출처: 사용자가 제시한 “뭐가 잘못되는지”에 대한 관찰.
| Chat Summary and Export Instructions | |
| Context Migration | |
| Continue and Recap Assistant | |
| Plan how to split long material for a model | |
| Stop the output from mixing languages |