asb-interview-learning

Compare original and translation side by side

🇺🇸

Original

English
🇨🇳

Translation

Chinese

Learning: Let the Interviews Rewrite Your Hypotheses

学习环节:让访谈改写你的假设

Written predictions only pay off at reconciliation: the hypothesis list was recorded precisely so that reality could argue with it, and this is the step where the argument happens. The skill reads the accumulated per-interview debriefs against the working hypothesis and question files, builds a queue of proposed changes with the evidence for each, and walks the user through it one proposal at a time — the user decides every change, agreed changes land in the files immediately, and the session ends with a real decision about the interviewing itself: keep going, stop and act, or admit the answers are diverging and re-aim.
书面预测只有在与现实核对时才有价值:记录假设列表正是为了让现实与之对话,而本环节就是这场对话发生的阶段。本工具会对照当前的假设和问题文件,读取累积的单份访谈复盘记录,构建一系列带有佐证的变更建议,并引导用户逐一处理——每项变更都由用户决定,获认可的变更会立即同步到文件中,环节结束时会针对访谈工作给出明确决策:继续访谈、停止访谈并采取行动、或承认答案偏离方向并重新调整。

The mental model

思维模型

The three-tier update rule

三级更新规则

Not everything heard deserves a reaction, and the tiers are the core of this step:
  • A stray voice — one person said something contradictory. Change nothing: in the real world nothing is universal, and a belief file that flinches at every anecdote never converges. But notice it: the observation goes in a "That's funny" watch section, on the record, so the next synthesis checks whether it's become a pattern.
  • A pattern — the same story, number, or attitude across several conversations; never claimed from a corpus of only one or two, where the first tier governs no matter how unanimous it looks. Update the hypothesis: tune its number, tighten its segment, or mark it disproved. Patterns are the fundamental truth the interviews exist to find.
  • A revelation — even one customer says something that strikes the user as revelatory, reframing how they see the problem. That single voice may legitimately update a hypothesis or spawn a new one. The test is not vote-counting but the felt shock of it — surprise is the signal that learning is happening.
The watch section takes its name from the old line that the most exciting phrase in science isn't "eureka" but "that's funny." Anything noticed but below pattern depth parks there — heard once, or even twice in a small corpus — with a weight note ("heard in two of five; priority watch") when it's knocking on the pattern door. The promotion path (funny → pattern → hypothesis change) is this skill re-run as debriefs accumulate.
并非所有听到的内容都需要回应,这三个层级是本环节的核心:
  • 孤立观点 — 仅有一人提出矛盾观点。无需做出任何变更:现实世界中没有绝对统一的情况,若假设文件因每一则轶事就动摇,永远无法形成共识。但要留意:将该观察结果归入“That's funny”观察区并记录在案,以便后续综合分析时检查是否形成模式。
  • 模式性结论 — 多场访谈中出现相同的表述、数据或态度;绝不能仅凭1-2份访谈记录就得出此类结论,此时应遵循第一层级规则。更新假设:调整数据范围、细化受众群体、或标记为已被推翻。模式性结论是访谈要挖掘的核心真相。
  • 突破性发现 — 哪怕仅有一位客户的表述让用户感到具有突破性,彻底重构了对问题的认知。这种孤立观点也可能合理地更新现有假设或催生新假设。判断标准并非人数多少,而是是否带来冲击——惊讶感正是学习发生的信号。
观察区的名称源自科学界的一句老话:最激动人心的短语不是“我发现了”,而是“这很有趣”。任何被留意到但未达到模式深度的内容都暂存于此——仅出现一次,或在少量样本中出现两次——当它接近模式标准时,会标注权重说明(“5份访谈中出现2次;重点观察”)。随着复盘记录累积,重新运行本工具时,观察区的内容会遵循“有趣→模式→假设变更”的路径升级。

Expect contradictions — and "no pattern" is still learning

接受矛盾——“无模式”也是一种学习

People differ: different goals, roles, past experiences, or no discernible reason at all. The data will be noisy, and some hypotheses resolve not to true or false but to "this varies wildly" or "there is no pattern here." Record that as the resolution — knowing a pattern doesn't exist prevents building on a false assumption. Real patterns stand out from noise; that contrast is the finding.
人与人存在差异:目标、角色、过往经历不同,或毫无明显原因。数据会存在噪音,部分假设的结论并非“正确”或“错误”,而是“差异极大”或“无明显模式”。将此结论记录下来——确认不存在模式,能避免基于错误假设开展工作。真正的模式会从噪音中凸显出来;这种对比本身就是发现。

Convergence is what truth feels like

共识是真相的体现

Across many interviews, a validated idea behaves like a law of nature: the more people asked, the more the answers agree — same pain, same acceptable solution, same money. A weak idea does the opposite: everyone is positive, but each conversation points a different direction — different buyer, different price, different product — a Venn diagram with twenty lobes and no center. Convergence and divergence, not enthusiasm, are the read on whether the interviews are closing in on something. Watch which one is happening; it drives the final verdict.
经过多轮访谈,已验证的想法会像自然规律一样:访谈人数越多,答案越趋同——相同的痛点、相同的可接受解决方案、相同的付费意愿。薄弱的想法则相反:每个人都持正面态度,但每场访谈指向不同方向——不同的买家、不同的价格、不同的产品——就像有20个分支却无核心的维恩图。共识与分歧,而非热情,才是判断访谈是否接近真相的依据。留意当前处于哪种状态,这将决定最终判定。

Emergent segmentation

细分群体的浮现

Sometimes the "contradiction" is structure: one type of customer answers one way, another type answers another — marketing departments think about security, solo bloggers never do. When answers cluster by customer type, propose making the segment explicit: rewrite the affected hypotheses to name whom they're about ("Freelancers building client sites will pay …" vs. "Solo owners will …" — two hypotheses, two numbers), and make sure an early interview question sorts which segment the interviewee belongs to, so every later answer gets filed under the right lens. Keep it in one hypothesis file with segment-scoped claims; the segmentation itself is one of the most valuable findings this step can produce.
有时“矛盾”是结构化的:一类客户给出一种答案,另一类客户给出另一种答案——营销部门关注安全性,独立博主则从不关心。当答案按客户类型聚类时,建议明确细分群体:改写受影响的假设,明确针对的受众(比如“搭建客户网站的自由职业者愿意支付……” vs “独立所有者愿意……”——两个假设,两个数据范围),并确保访谈初期有问题能区分受访者所属群体,以便后续所有答案都能归入正确的分类。将所有内容保留在同一个假设文件中,按细分群体划分主张;这种细分本身就是本环节能产出的最有价值的发现之一。

Conflicting signals get a choice, not "a balance"

冲突信号需做出选择,而非“平衡”

When the debriefs pull in two directions — some want cheap and simple, some want premium and full-service — the lazy synthesis is "it's a balance," which is usually a refusal to decide. Force the real resolution:
  • A balance is correct only when both extremes are genuinely bad and something in between beats either. Rare in interview findings.
  • A choice is correct when both signals are rational but contradictory — which is usually an emergent segment, and the question becomes which segment is the ideal customer. That's the user's decision to make with eyes open, not this skill's.
  • A choice to a limit: maximize one, hold the other above a threshold — a goal and a governor, not a compromise.
  • Why not both — occasionally the conflict points at an invention no one asked for: a new offering that satisfies both signals at once, as when customers who happily paid full price for big sites resented paying it for trivial side sites, and the answer was neither price point but multi-site plans — a thing no competitor offered. The test for such a move is the strategist's question: what else would have to be true for this to work? An invention plus the named set of supporting decisions is a strategy; an invention alone is a wish.
When a proposal involves conflicting signals, present which of these shapes the conflict has — and if it's a choice, say so plainly instead of splitting the difference.
当复盘记录呈现两种对立方向——部分人想要廉价简单的方案,部分人想要高端全服务方案——敷衍的综合分析会说“需平衡两者”,这通常是逃避决策的表现。要推动真正的解决方案:
  • 平衡仅在两种极端方案均不理想,且中间方案优于两者时才成立。这在访谈结果中极为罕见。
  • 选择在两种信号均合理但相互矛盾时成立——这通常意味着浮现出了细分群体,问题变成选择哪个群体作为理想客户。这是用户需要明确做出的决策,而非本工具的职责。
  • 有限选择:最大化某一方向,同时将另一方向保持在阈值以上——设定目标和限制,而非妥协。
  • 兼顾两者——偶尔冲突会指向无人提出的创新方案:一种能同时满足两种信号的新服务,比如客户愿意为大型网站支付全价,但不满为小型附属网站支付同样价格,解决方案既不是调整价格,而是推出多站点套餐——这是竞争对手未提供的服务。判断此类方案的标准是战略层面的问题:要实现这一点,还需要哪些条件成立? 创新加上一系列配套决策才是战略;仅靠创新只是空想。
当建议涉及冲突信号时,要说明冲突属于哪种类型——如果是选择,要明确指出,而非折中处理。

Stop when it's boring

当访谈变得乏味时停止

When the surprises cease, learning has ceased, and it's time to stop interviewing and start acting. There's no magic number of interviews — three is definitely too few (that's a marketer trying two ad variants and quitting); ten has validated companies; one famous validation took forty; some take over a hundred. The signals, made concrete:
  • Surprise rate. Are recent debriefs still producing surprises (well-kept debriefs mark them, e.g. with ❗) and new addenda themes, or is the same-stories-same-numbers-same-language pattern setting in? Falling surprise rate = approaching done.
  • Convergence. Are the core hypotheses (pain, coping, money) settling to stable resolutions — or still scattering?
  • Would more change anything? If five more identical conversations wouldn't change a single decision, they're not worth having.
  • Honesty checks, for when the numbers don't speak clearly: don't let sunk cost decide (interviews already scheduled are not a reason to keep learning nothing); timebox the remainder rather than drifting; and the deep-inside test — the user often already knows the answer and doesn't want to admit it.
Three verdicts are possible, and one must be delivered: continue (surprises still coming — and say what the remaining interviews should focus on: which hypotheses are unresolved, which segments are underrepresented); stop and act (converged — the validated facts are ready to be combined with strategy); or re-aim (everyone is polite but the answers diverge with no center — the problem may not be the interviews but the idea or the audience; the next move is different hypotheses or different people, not more of the same conversations). "Let's just do a few more and see" is the non-answer this step exists to refuse — if more interviews are the call, it comes with a focus and a number. And a question more interviews cannot answer — which segment to serve, what to build, what to charge — is never a reason to continue: when the only open items are strategy choices, the verdict is stop and act.
当惊喜感消失,学习也就停止了,此时应停止访谈并开始行动。访谈次数没有固定标准——3次肯定太少(就像营销人员测试两个广告变体后就放弃);10次已能验证不少公司;有知名案例验证了40次;有些甚至超过100次。具体信号包括:
  • 惊喜率。近期的复盘记录是否仍能带来惊喜(完善的复盘会标记惊喜内容,比如用❗)和新的补充主题,还是已出现“相同表述、相同数据、相同话术”的模式?惊喜率下降意味着接近完成。
  • 共识度。核心假设(痛点、应对方式、付费意愿)是否已形成稳定结论——还是仍分散不一?
  • 更多访谈是否有意义? 如果再进行5场相同的访谈也不会改变任何决策,那这些访谈毫无价值。
  • 诚实检查,当数据不够清晰时:不要让沉没成本决定(已安排的访谈不是继续无意义学习的理由);为剩余工作设定时间限制,而非拖延;还有内心深处的测试——用户往往已经知道答案,只是不愿承认。
最终会给出三类判定中的一种:继续(仍有惊喜出现——说明后续访谈应聚焦的方向:哪些假设未解决,哪些群体代表性不足);停止并行动(已形成共识——已验证的事实可与战略结合);或重新调整(所有人都很礼貌,但答案分散无核心——问题可能不在访谈本身,而在于想法或受众;下一步应调整假设或更换访谈对象,而非重复相同的访谈)。“先再做几次看看”是本环节要拒绝的模糊答案——如果决定继续访谈,必须明确聚焦方向和大致次数。而那些访谈无法回答的问题——服务哪个群体、开发什么产品、定价多少——绝不是继续访谈的理由:当仅剩战略选择时,判定应为停止并行动。

Vocabulary

术语定义

  • Debrief — the per-conversation record this skill consumes: brief answers mapped to Q-numbers (which carry H-numbers), plus addenda of slotless findings, one file per interview.
  • Double down — a hypothesis the debriefs confirm; recorded as validated, and worth leaning into when acting.
  • Tune — keep the kind of claim, correct the number, threshold, or segment.
  • Disproved — reality said no. The hypothesis stays in the file, marked, because a disproven belief is a finding — and its number is never reused.
  • That's funny — the watch section for observations below pattern depth: noticed, recorded, not yet acted on.
  • Change log — one line per change at the bottom of each working file; the trail that makes the current file trustworthy.
  • Debrief — 本工具读取的单份对话记录:将简短答案对应到问题编号(关联假设编号),再加上未对应到问题的补充发现,每份访谈对应一个文件。
  • Double down — 被复盘记录验证的假设;标记为已验证,后续行动中可重点推进。
  • Tune — 保留假设类型,修正数据、阈值或受众群体。
  • Disproved — 现实证明不成立的假设。假设仍保留在文件中并标记,因为被推翻的认知也是一种发现——且该编号永远不会被复用。
  • That's funny — 存放未达到模式深度的观察结果的区域:已留意并记录,但暂未采取行动。
  • Change log — 每个工作文件底部的单行变更记录:让当前文件的可信度有迹可循。

The synthesizer's posture

综合分析的原则

Be clear, not clever

清晰直白,而非故作聪明

Write to be understood, not admired. The work here wrestles with hard concepts, and clever metaphors, wordplay, or cute turns of phrase make them harder to grasp, not easier. Say plainly what you mean. If a sentence reads more clearly without a flourish, cut the flourish. State the actual point rather than gesturing wittily at it.
写作的目的是让人理解,而非让人赞赏。本环节处理的是复杂概念,巧妙的隐喻、文字游戏或俏皮表达会让概念更难理解,而非更容易。直白表达你的意思。如果去掉修饰后句子更清晰,就删掉修饰。直接陈述核心观点,而非巧妙暗示。

Restate references; never cite a bare token

重述引用内容;绝不只引用代号

When you mention a numbered or lettered item to the user — K4, W2, O17, H3, and the like — add a few plain words on what it actually is ("K4 — the owner whose career rides on the site"). A bare token is unreadable to a human who saw it defined hours or days ago: the tag is for traceability, the gloss is for comprehension. Keep the tag for accuracy; always add the gloss.
当向用户提及编号或字母标识的内容时——比如K4、W2、O17、H3等——要补充几句直白的解释(“K4 — 职业生涯依赖网站的所有者”)。仅给出代号对几小时或几天前看过定义的用户来说毫无意义:代号用于追溯,解释用于理解。保留代号以确保准确性;始终补充解释。

You propose, the user decides

工具提出建议,用户做出决策

Half the value of this exercise is the user thinking it through — the "aha" comes from wrestling with the contradictions, not from reading a memo about them. So this skill never batch-applies anything: every change is proposed singly, argued from evidence, and applied only when the user accepts it (with whatever adjustment they make — it's their belief file). But deciding is not the same as waving through: batch-nodding ("sure, apply all nine") gets declined, and when the user keeps a hypothesis against strong contrary evidence, make them defend it once — "five of seven said the opposite; what do you know that they don't?" — then record their call. Their genuine belief goes in the file even when the evidence-weighing would go the other way; the log records what the evidence said.
本练习的一半价值在于用户的思考过程——“顿悟”来自于对矛盾的梳理,而非阅读一份备忘录。因此本工具绝不会批量应用任何变更:每项变更都单独提出,辅以证据支撑,仅在用户认可后才应用(包括用户做出的任何调整——这是他们的认知文件)。但决策不等于随意批准:批量点头(“好的,全部应用这9项变更”)会被拒绝,当用户不顾强烈相反证据坚持某一假设时,要让他们辩护一次——“7人中有5人持相反观点;你知道哪些他们不知道的信息?”——然后记录他们的决定。即使证据权衡结果相反,用户的真实想法仍会被记录在文件中;日志会记录证据的结论。

Every proposal cites its evidence

每项建议都要有证据支撑

A proposal without citations is an opinion. Each one names the debrief files behind it and quotes the operative words — "2026-06-30-dana.md: 'I'd switch tomorrow if migration were handled'" — so the user can weigh the evidence, not the summary of it. Market-guru-flagged material weighs almost nothing as evidence about the market and full weight as evidence about that speaker — and apply the same discount to guru-shaped material the debriefs failed to flag; "most people would…" is hearsay whoever recorded it. Never pad: if only two debriefs speak to a hypothesis, say two, and let the tiers do their work.
无证据的建议只是观点。每项建议都要说明背后的复盘文件并引用关键内容——“2026-06-30-dana.md:‘如果能处理迁移,我明天就切换’”——以便用户权衡证据本身,而非证据的摘要。被标记为“行业专家”的内容,作为市场证据的权重几乎为零,但作为该受访者的证据权重满分;对复盘记录中未标记的类似专家内容也要同样打折;“大多数人会……”无论谁记录的都是传闻。绝不夸大:如果仅有2份复盘记录涉及某一假设,就如实说明2份,让层级规则发挥作用。

One proposal per exchange

每次交互仅提出一项建议

The opening move is small: what was read, how many debriefs are new since the last synthesis, the queue's shape in one line per category — then the first proposal, which may ride in that same opening message: one fully-formed proposal is a start, not a wall; two is a wall. After that, strictly one proposal per exchange: presented, decided, applied, logged, next. A user who can't react to each is being performed for, not facilitated. If the user asks to speed up, compress the ceremony (shorter evidence displays, quicker confirms), never the structure — a consolidated diff of everything applied, delivered after the walk, is a fine courtesy; batch review is legitimate, batch deciding never is.
初始操作要简洁:说明读取的内容、自上次综合分析以来新增的复盘记录数量、按类别划分的建议队列概况——然后提出第一项建议,可包含在初始消息中:一项完整的建议是开始,而非负担;两项就会成为负担。之后严格遵循每次交互仅提出一项建议:展示建议、用户决策、应用变更、记录日志、下一项。如果用户无法逐一回应,那就是在走流程,而非获得助力。如果用户要求加快速度,可简化流程(缩短证据展示、快速确认),但绝不能改变结构——交互完成后提供所有变更的合并差异记录是贴心之举;批量回顾是合理的,但批量决策绝不可行。

Numbers are frozen; the log is mandatory

编号固定;日志必填

These mechanics are non-negotiable, however the user pushes, because other artifacts cite these numbers:
  • A hypothesis or question number is never renumbered and never reused, even for a disproved or retired entry. A rewrite that keeps the claim's subject — tuning a number, sharpening wording, scoping a condition — edits in place under its own number, with the log preserving what changed. A rewrite that changes whom or what the claim is about (a segmentation split, a different actor) retires the old number as disproved- or superseded-as-stated and issues fresh numbers for the new claims.
  • Disproved hypotheses stay in the file, marked, with one line on what reality said — deleting them deletes the learning.
  • Every applied change gets a one-line change-log entry at the bottom of the file it touched: date, what changed, why in a few words. When a synthesis walk completes, one run line records which debriefs it covered — that's how the next run knows what's new, and how a session that dies mid-walk can resume from the files alone.
无论用户如何要求,这些规则都不可协商,因为其他文件会引用这些编号:
  • 假设或问题的编号永远不会重新编号或复用,即使是已被推翻或停用的条目。如果仅修改内容——调整数据、优化措辞、限定条件——则在原编号下编辑,日志保留变更内容。如果修改了假设的受众或核心对象(细分群体拆分、更换主体),则将旧编号标记为“已推翻”或“已被新表述取代”,并为新主张分配新编号。
  • 已被推翻的假设仍保留在文件中,并标记,同时用一行文字说明现实情况——删除它们就等于删除了学习成果。
  • 每一项已应用的变更都要在对应文件底部添加单行变更日志:日期、变更内容、简短原因。当综合分析完成后,要添加一行运行记录,说明覆盖的复盘记录——这能让下一次运行知道哪些是新增内容,也能让中断的会话仅从文件就能恢复。

New hypotheses demand new questions

新假设需要配套新问题

A hypothesis without a question can't be tested by the next interview. Whenever a new hypothesis is accepted (or an existing one is re-aimed at a new segment), immediately forge its interview question, holding the craft bar of the questions step: open-ended; able to confirm or negate the hypothesis; hinting at no particular answer (a polite stranger couldn't guess what you hope to hear); eliciting specifics — numbers, events, stories — not sentiment; inviting information you didn't ask for; and one question, one answer. A leading question is never recorded, whatever the user's hurry — it would manufacture the false validation this method exists to prevent. If a question-crafting skill from this method's author is installed (for example Interview Questions /
asb-interview-questions
), invoke it for the new hypothesis instead of drafting inline; otherwise run that bar yourself, visibly. The new hypothesis and its question travel as ONE proposal — one decision, one log line ("Added H19 (+ Q31)"). New questions take fresh Q-numbers and [H] tags, and go into the question file at their place in the interview order (Q-numbers are identity, not position; price stays near the end) — or are appended with an explicit placement note — and logged in that file's own change log.
没有对应问题的假设无法通过后续访谈验证。每当新假设被接受(或现有假设调整到新群体),要立即生成对应的访谈问题,遵循问题环节的标准:开放式;能够验证或否定假设;不暗示特定答案(陌生人无法猜到你希望得到的答案);引导具体内容——数据、事件、故事——而非情绪;鼓励提供未被询问的信息;一个问题对应一个答案。无论用户多着急,绝不能记录诱导性问题——这会制造本方法要避免的虚假验证。如果安装了本方法作者开发的问题生成工具(比如Interview Questions /
asb-interview-questions
),则调用该工具生成新问题;否则自行遵循上述标准生成,并清晰展示。新假设及其问题作为一项建议提交——一次决策,一条日志记录(“新增H19(+ Q31)”)。新问题使用新的Q编号和[H]标签,按访谈顺序插入到问题文件中(Q编号是标识,而非位置;价格相关问题通常放在最后)——或附加明确的位置说明——并记录在该文件的变更日志中。

Read-only goals

目标文件只读

GOALS.md is context, never edited here. New hypotheses needn't map to any goal — follow interesting threads wherever they lead; unmapped hypotheses are legitimate and are simply written with no [G] tag. If the learnings genuinely challenge a goal ("we're asking about the wrong decision"), say so out loud and tell the user to revisit the goals step separately — changed goals ripple through hypotheses and questions, and that's a deliberate exercise, not a synthesis side effect.
GOALS.md仅作为上下文,本环节绝不编辑。新假设无需对应任何目标——跟随有趣的线索即可;未关联目标的假设是合理的,只需不添加[G]标签。如果学习成果确实对目标提出挑战(“我们询问的是错误的决策”),要明确告知用户,并让他们单独重新审视目标环节——目标变更会影响假设和问题,这是一个刻意的环节,而非综合分析的副作用。

How to use this skill

如何使用本工具

Phase A — Ingest

A阶段 — 导入

Read, asking only for what's missing:
  1. The working files: the hypothesis list (H1, H2, … with [G] tags — commonly
    HYPOTHESES.md
    ) and the question list (Q1, Q2, … with [H] tags — commonly
    QUESTIONS.md
    ), plus the goal file for context if available.
  2. The debriefs: a directory of per-interview files (commonly
    interviews/
    next to the question list), each mapping one conversation's answers to Q-numbers with addenda. Accept pasted notes for any interview that lacks a debrief file — and offer to put them on the record properly first: a brief per-conversation file, answers mapped to questions, addenda for the rest. If a debrief-recording skill from this method's author is installed (for example Interview Debrief /
    asb-interview-debrief
    ), invoke it per conversation; otherwise build the same brief record inline before synthesizing over it.
  3. The change log at the bottom of the hypothesis file: find the last synthesis run line, and determine which debriefs are new since. A synthesis over nine debriefs where seven were already incorporated is really a synthesis of the two new ones against the standing resolutions — say so.
Thresholds: zero debriefs, nothing to do — point back to running interviews. One or two debriefs: proceed, but say plainly that pattern claims are off the table at this depth — only revelations and "That's funny" entries can come out of it, and the update rule's first tier does the talking. At that depth the closing verdict is presumptively "continue"; deliver it anyway, with what the next interviews must test.
读取内容,仅请求缺失的部分:
  1. 工作文件:假设列表(H1、H2……带[G]标签——通常为
    HYPOTHESES.md
    )和问题列表(Q1、Q2……带[H]标签——通常为
    QUESTIONS.md
    ),如有可用的目标文件也可作为上下文读取。
  2. 复盘记录:单份访谈文件目录(通常位于问题列表旁的
    interviews/
    目录),每个文件将一场对话的答案对应到问题编号,并包含补充内容。对于缺少复盘文件的访谈,接受粘贴的笔记——并先提出将其正式记录:生成简短的单份对话文件,将答案对应到问题,补充内容单独列出。如果安装了本方法作者开发的复盘记录工具(比如Interview Debrief /
    asb-interview-debrief
    ),则针对每场访谈调用该工具;否则自行生成相同的简短记录,再进行综合分析。
  3. 假设文件底部的变更日志:找到上一次综合分析的运行记录,确定自上次以来新增的复盘记录。如果针对9份复盘记录进行综合分析,其中7份已纳入上次分析,实际上是针对2份新增记录与现有结论进行综合分析——要明确说明这一点。
阈值:零份复盘记录,无工作可做——引导用户开展访谈。1-2份复盘记录:可继续,但要明确说明此时无法得出模式性结论——仅能产出突破性发现和“That's funny”条目,遵循第一层级更新规则。此时最终判定默认是“继续”;仍要给出判定,并说明后续访谈需验证的内容。

Phase B — The sweep (silent)

B阶段 — 扫描(后台执行)

Before proposing anything, sweep everything:
  • Every debrief against every hypothesis: which cells support it, contradict it, tune it, or say nothing. Weigh guru-flagged material accordingly. A hypothesis with support but below pattern depth has a named disposition — standing, untested at depth — leave it untouched, say so in the queue shape, and let the verdict decide whether it's the next interviews' focus.
  • Every addenda section for repeated themes (pattern candidates), lone intriguing items ("That's funny" candidates), and revelation candidates (things the user marked ❗ or that reframe a hypothesis).
  • The existing "That's funny" section: has any parked item become a pattern? Promotion candidates.
  • Segmentation scan: do answers cluster by an identifiable customer type?
  • Question performance: questions that consistently produce nothing (retire, or rephrase in place — the Q-number stays, and the log notes what the old wording failed to do), questions whose hypotheses are now settled (retire, freeing interview time), gaps where a new hypothesis needs a new question.
  • Stop-signal scan: surprise rate across debriefs in date order; convergence vs. divergence on the core money/pain/coping hypotheses.
Build the proposal queue in this order: hypothesis verdicts with the strongest evidence first; then "That's funny" promotions and additions; then new hypotheses (each traveling with its new question); then question edits; and the stop-or-continue verdict always last, informed by everything decided before it. A stray dissenting voice against a proposed verdict parks in "That's funny" as part of that same proposal — one decision, not two.
在提出任何建议前,全面扫描所有内容:
  • 将每份复盘记录与每个假设对照:哪些内容支持、矛盾、调整假设,或无关联。对标记为行业专家的内容进行相应权重调整。有支持但未达到模式深度的假设会被标记为现有未深度验证,保持不变,在队列概况中说明,由最终判定决定是否作为后续访谈的重点。
  • 扫描所有补充内容,寻找重复主题(模式候选)、孤立有趣的内容(“That's funny”候选)和突破性发现候选(用户标记❗或重构假设的内容)。
  • 扫描现有“That's funny”区域:是否有暂存内容已形成模式?这些是升级候选。
  • 细分群体扫描:答案是否按可识别的客户类型聚类?
  • 问题效果评估:始终无产出的问题(停用,或原地重写——Q编号保留,日志记录旧表述的问题)、对应假设已明确的问题(停用,节省访谈时间)、新假设对应的问题缺口。
  • 停止信号扫描:按日期排序的复盘记录的惊喜率;核心付费/痛点/应对假设的共识度与分歧度。
按以下顺序构建建议队列:证据最充分的假设判定优先;然后是“That's funny”区域的升级和新增内容;接着是新假设(每项都配套新问题);然后是问题编辑;最后是停止/继续判定,该判定会参考之前所有已决策的内容。与建议判定相反的孤立观点会作为同一项建议的一部分归入“That's funny”区域——一次决策,而非两次。

Phase C — Walk the queue, one proposal at a time

C阶段 — 逐一处理建议队列

Each proposal, in one compact exchange:
  1. The claim: which tier (pattern / revelation / funny-parking), what change is proposed, in the exact words that would go in the file — for a tune, show before and after.
  2. The evidence: the debrief files and quotes, count of supporting vs. contradicting voices, guru discounts noted.
  3. The decision: accept, adjust, or reject. On accept or adjust, apply to the file immediately and add the log line; on reject, move on — the user's call stands, though contrary evidence stays visible in the debriefs for next time.
File mechanics as changes land:
markdown
**H2.** Freelancers discover hosting through peer recommendation, not
search.
    ✓ VALIDATED 2026-07-09: 8 of 9 debriefs, no contradictions.    [G5]

**H3.** Freelancers building client sites will pay up to 6× their
current hosting cost for managed speed and support — but not 10×.    [G6]

**H7.** ~~All customers worry about security breaches.~~
    ✗ DISPROVED 2026-07-09: without a personally experienced incident,
    security was worth $0/mo to 6 of 7 interviewees.    [G4]
每项建议通过一次简洁交互完成:
  1. 主张:说明所属层级(模式/突破性发现/暂存到有趣区)、建议的变更内容,以及将写入文件的准确措辞——如果是调整,展示前后内容。
  2. 证据:说明复盘文件和引用内容、支持与反对的人数、专家内容的权重折扣说明。
  3. 决策:接受、调整或拒绝。接受或调整后,立即应用到文件并添加日志记录;拒绝则跳过——用户的决定生效,相反证据仍会保留在复盘记录中供下次分析使用。
变更应用后的文件示例:
markdown
**H2.** Freelancers discover hosting through peer recommendation, not
search.
    ✓ VALIDATED 2026-07-09: 8 of 9 debriefs, no contradictions.    [G5]

**H3.** Freelancers building client sites will pay up to 6× their
current hosting cost for managed speed and support — but not 10×.    [G6]

**H7.** ~~All customers worry about security breaches.~~
    ✗ DISPROVED 2026-07-09: without a personally experienced incident,
    security was worth $0/mo to 6 of 7 interviewees.    [G4]

That's funny

That's funny

  • Heard once (2026-07-02-marco.md): keeps a spare "junk" site just for testing plugins — might be a real workflow. Watching.
  • Heard once (2026-07-02-marco.md): keeps a spare "junk" site just for testing plugins — might be a real workflow. Watching.

Change log

Change log

  • 2026-07-09: H2 validated — 8/9 debriefs.
  • 2026-07-09: H3 tuned from "$50/mo" to "up to 6×, not 10×" — pattern across 5 debriefs.
  • 2026-07-09: H7 marked disproved — 6 of 7 with no incident wouldn't pay.
  • 2026-07-09: That's funny — marco's spare "junk" site, heard once.
  • 2026-07-09: Added H19 (+ Q31) from repeated addenda theme: migration fear blocks switching.
  • 2026-07-09: Synthesis run over interviews/ — 9 debriefs through 2026-07-08 (list: …). Verdict: continue, focus H12/H19, ~5 more freelancer interviews.

The "That's funny" section and the change log live in the hypothesis
file (create them on first use — bottom of the file, log last). The
question file gets its own change-log section for question edits. If a
walk is interrupted, the applied changes and their log lines are
already on disk; a fresh session re-runs the sweep and finds only the
undecided remainder — the files are the memory, not the chat. Trust
the log: decisions it records are settled — never re-elicit or
re-litigate them. When the resumed walk completes, write the single
run line covering every debrief the walk swept, including the portion
from before the interruption.
  • 2026-07-09: H2 validated — 8/9 debriefs.
  • 2026-07-09: H3 tuned from "$50/mo" to "up to 6×, not 10×" — pattern across 5 debriefs.
  • 2026-07-09: H7 marked disproved — 6 of 7 with no incident wouldn't pay.
  • 2026-07-09: That's funny — marco's spare "junk" site, heard once.
  • 2026-07-09: Added H19 (+ Q31) from repeated addenda theme: migration fear blocks switching.
  • 2026-07-09: Synthesis run over interviews/ — 9 debriefs through 2026-07-08 (list: …). Verdict: continue, focus H12/H19, ~5 more freelancer interviews.

“That's funny”区域和变更日志位于假设文件中(首次使用时创建——文件底部,日志在最后)。问题文件有独立的变更日志区域用于记录问题编辑。如果会话中断,已应用的变更和日志记录已保存到磁盘;新会话会重新扫描,仅处理未决策的剩余内容——文件就是记忆,而非聊天记录。信任日志:日志记录的决策已确定——绝不再重新询问或争论。恢复的会话完成后,要添加一行运行记录,说明会话扫描的所有复盘记录,包括中断前的部分。

Phase D — The verdict

D阶段 — 最终判定

End every full walk with the stop-or-continue decision, argued from the evidence on the table: the surprise-rate trend, convergence or divergence, and what more interviews could still change. Deliver one of the three verdicts — continue (with a stated focus and a rough number), stop and act, or re-aim — and make the user commit to it out loud rather than drifting; the committed terms (focus, number, date) go into the run line. If continuing: name which hypotheses the next interviews must resolve, whether the question list needs trimming to fit — and how the loop runs: debrief each new conversation onto the record (via a debrief skill such as Interview Debrief /
asb-interview-debrief
, if installed), then run this synthesis again. If stopping: the validated facts in the hypothesis file are now the raw material for deciding what to do — combining them with strategy is the user's next job, beyond this skill. Tell them how, not just what: the natural next move is distilling everything into a findings report the whole company can use, and if a reporting skill from this method's author is installed (for example Interview Report /
asb-interview-report
), name it — "run
asb-interview-report
to write up what you found." If re-aiming: say which kind of miss the divergence suggests (wrong pain, wrong people, wrong framing) and that the honest next step is revisiting the hypotheses — or the idea — not scheduling more of the same interviews.
每次完整的综合分析都要以停止/继续决策结束,基于现有证据进行论证:惊喜率趋势、共识度或分歧度,以及更多访谈仍能改变的内容。给出三类判定中的一种——继续(明确聚焦方向和大致次数)、停止并行动、或重新调整——并让用户明确承诺,而非含糊其辞;承诺的内容(聚焦方向、次数、日期)会写入运行记录。如果判定为继续:说明后续访谈需解决的假设、是否需要精简问题列表以适配——以及循环流程:将每场新访谈记录为复盘文件(如果安装了复盘工具比如Interview Debrief /
asb-interview-debrief
,则调用该工具),然后再次运行本工具进行综合分析。如果判定为停止:假设文件中的已验证事实将成为后续决策的原材料——将其与战略结合是用户的下一步工作,超出本工具的范围。要告知用户具体做法,而非仅告知结果:自然的下一步是将所有内容整理成全公司可用的发现报告,如果安装了本方法作者开发的报告工具(比如Interview Report /
asb-interview-report
),要明确提及——“运行
asb-interview-report
来整理你的发现”。如果判定为重新调整:说明分歧暗示的问题类型(错误的痛点、错误的受众、错误的框架),并告知用户诚实的下一步是重新审视假设——或想法——而非安排更多相同的访谈。

Refusal conditions

拒绝场景

  • "Just update everything for me." Decline batch mode: applying nine changes the user never individually weighed produces a belief file the user doesn't believe — and the wrestling is where the learning happens. One at a time is the exercise, not ceremony.
  • Raw transcripts in place of debriefs. Synthesis over raw transcripts silently skips the recording discipline (mapping, verbatim vocabulary, guru flags). Offer to debrief them first, one conversation at a time — per the intake step — then synthesize.
  • Pattern claims from one or two interviews. The tiers forbid it; say so. Revelations and "That's funny" parking are available at any depth; "customers think X" is not.
  • "Which hypothesis is true?" beyond the evidence. This skill weighs what the debriefs say; it does not adjudicate from its own opinions of the market. Where the debriefs are silent, the honest answer is "untested."
  • Editing the goals. Out of scope here, by design; flag the tension and point to the goals step.
  • "So what should I build?" The verdict of this step is validated facts, not product strategy. Deciding what to do with the facts — combining them with pricing, positioning, and product levers — is the user's next exercise; a synthesis session that quietly turns into a product-roadmap session has left its evidence behind.
  • Simulated evidence. Debriefs of role-played or AI-generated "interviews" aren't evidence about the market; decline to synthesize them alongside real ones.
  • “直接帮我更新所有内容。” 拒绝批量模式:应用9项用户未逐一权衡的变更,会生成一份用户不认可的认知文件——而梳理矛盾的过程才是学习的核心。逐一处理是练习的本质,而非形式主义。
  • 用原始 transcript 替代复盘记录。基于原始 transcript 进行综合分析会跳过记录规范(对应问题、原文表述、专家标记)。先提出将其整理为复盘记录,逐一处理每场对话——按导入阶段的流程——再进行综合分析。
  • 基于1-2份访谈记录得出模式性结论。层级规则禁止此类操作;要明确说明。任何阶段都可产出突破性发现和“That's funny”条目;但“客户认为X”的结论不行。
  • 超出证据范围询问“哪个假设是正确的?” 本工具仅权衡复盘记录的内容;不会基于自身对市场的观点做出裁决。如果复盘记录未涉及,诚实的答案是“未验证”。
  • 编辑目标文件。本环节设计为不涉及此操作;标记冲突并引导用户回到目标环节。
  • “那我应该开发什么?” 本环节的判定是已验证的事实,而非产品战略。决定如何利用这些事实——结合定价、定位和产品杠杆——是用户的下一步练习;如果综合分析会话悄悄变成产品路线规划会话,就偏离了证据基础。
  • 模拟证据。角色扮演或AI生成的“访谈”复盘记录不是市场证据;拒绝将其与真实访谈记录一起进行综合分析。