Purpose
中文:一份系统化的清单,罗列所有能把一个人从其表态选择推向其经反思认可选择(反思均衡)的因素。反思游戏的任务,就是逐一把这些因素浮现出来——靠亲历式模拟,而非说教——直到某个决定在充分考量下依然稳定:「我能接受,我不后悔。」把它当作场景/选项生成的覆盖清单,以及调查员智能体的探查目标。它在 ZH 的种子(城市 · 成本 · 主要因素 · 偏差)之上做了扩展。
EN: A systematic inventory of everything that can move a person from their stated choice toward their reflectively-endorsed one (反思均衡). The reflection game’s job is to surface each of these — through lived simulation, not lecturing — until a decision is stable under full consideration: “我能接受,我不后悔.” Use this as the coverage checklist for scene/option generation and for the investigator agent’s probing targets. Extends ZH’s seed (cities · costs · major factors · biases).
如何读这份清单 · How to read this
中文:一个选择处于反思均衡,当它能经受住 (a) 对其本应权衡的每个实质因素的充分暴露,以及 (b) 对那些扭曲它的偏差的纠正——前提是 (c) 偏好部分是在生活中形成的,而非从一个固定先验里读出来。下面三个区块正对应这三点。区块 IV 讲游戏如何触发每一项。
EN: A choice is at reflective equilibrium when it survives (a) exposure to every substantive factor it should weigh, and (b) correction of the distortions biasing it — given that (c) preferences are partly formed in the living, not read off a fixed prior. The three blocks below are exactly those. Block IV is how the game triggers each.
Design principle (from the LfL impossibility result)
中文:一个人若急着给出「显而易见」的答案,他就是不可读的——你只能学到他的最优选项,永远学不到他的排序。所以每个因素都必须经由一个强制权衡来浮现,逼他去给那些原本会跳过的选项排序。别去问;让他亲历一个该因素会咬人的选择。
EN: A person who rushes to the “obvious” answer is illegible — you only learn their top pick, never their ordering. So each factor must be surfaced via a forced tradeoff that makes them rank options they’d otherwise skip. Don’t ask; make them live a choice where the factor bites.
I. 实质因素(他们可能权衡不足的)· Substantive factors (what they may under-weigh)
中文:下表按簇列出需要浮现的实质因素——学业、职业、财务、城市、日常生活、关系、身份/价值、反事实成本、时间与可逆性、风险。
| Cluster | Factors to surface |
|---|---|
| Academic / major | course loads & difficulty · what you actually study day-to-day · 保研/转专业/辅修 flexibility · 学科实力 vs 学校牌子 · 培养质量 · 挂科/劝退风险 |
| Career | 对口率 · 就业前景 & 岗位画像 · 天花板/上升空间 · 35岁问题 · 行业周期与风口 · 体制内 vs 市场 · 创业可能 · 地域就业机会密度 |
| Financial | 应届起薪 vs 长期收入曲线 · 学费/家庭负担 · 城市生活成本 · ROI & 回本周期 · “financial promises” vs 真实分布 |
| City / geography | 城市能级(一线/新一线)· 实习与就业机会 · 气候/饮食/方言/文化适配 · 离家远近 · 落户/户口 · 婚恋与社交圈 |
| Daily / communal life | 日常社群与归属感 · 宿舍/校园文化 · 社团/活动 · 通勤 · 作息与 work-life rhythm(这是最常被忽视的”least-examined”维度之一) |
| Relational | 家庭期望与代际压力 · 父母养老 · 伴侣/异地 · 同辈比较 |
| Identity / values | 兴趣 vs 擅长 · 意义感/使命 · 自我认同 · 8 维价值取向(TRUTH↔PROTECTION … RIGOR↔MERCY) |
| Counterfactual cost | 机会成本——选 A 到底放弃了什么(what to give up) · 对称呈现红利与代价 |
| Time & reversibility | 短期 vs 长期(NOW↔LATER)· 可逆性——转专业/考研/留学/转行 作为”出口”的真实难度 |
| Risk | 最坏情况 · 下行保护 · 波动容忍度 · 路径的脆弱性 |
II. 扭曲(弯曲当前偏好的偏差)· Distortions (biases that bend the current preference)
中文:把这些浮现出来、并温和地命名——本游戏相较推荐系统的优势,正在于它让人在亲历后果中自己抓住偏差。
- deference to authority — 父母/老师/张雪峰 的话当成事实
- confirmation bias — 只接收支持既有选择的信息
- conformity / 从众 — 追热门专业、随大流
- loss aversion — 过度规避”亏”,错估风险
- present bias / 双曲贴现 — 高估眼前、低估长期
- anchoring — 被分数/排名/一本线锚定
- availability — 听来的个案当成普遍分布
- sunk cost — 已投入的备考/兴趣绑架选择
- prestige / status bias — 名校光环压过专业适配
- halo effect — 学校牌子好 ⇒ 误以为专业也好
- optimism / planning fallacy — “我一定能进前5%转专业”
- projection bias — 以为现在的喜好会一直持续
- social desirability — 选”别人觉得对”的而非自己要的
EN: Surface and gently name these — the game’s edge over a recommender is that it lets people catch their own bias by living the consequence.
- deference to authority — 父母/老师/张雪峰 的话当成事实
- confirmation bias — 只接收支持既有选择的信息
- conformity / 从众 — 追热门专业、随大流
- loss aversion — 过度规避”亏”,错估风险
- present bias / 双曲贴现 — 高估眼前、低估长期
- anchoring — 被分数/排名/一本线锚定
- availability — 听来的个案当成普遍分布
- sunk cost — 已投入的备考/兴趣绑架选择
- prestige / status bias — 名校光环压过专业适配
- halo effect — 学校牌子好 ⇒ 误以为专业也好
- optimism / planning fallacy — “我一定能进前5%转专业”
- projection bias — 以为现在的喜好会一直持续
- social desirability — 选”别人觉得对”的而非自己要的
III. 偏好形成动力学(ZH 的关键难点)· Preference-formation dynamics (ZH’s crux)
中文:最难的部分:偏好并不是一个可被读出的固定先验——它部分是在生活中形成的。这份清单必须尊重这一点,而不能把它假设掉。
- 形成的路径依赖 — 你将会想要什么,取决于你先过了什么样的生活;A/B 模拟的存在正是为了暴露这一点(两种都过一遍,再来比较)。
- 适应性偏好(酸葡萄) — 人们在选择之后为之合理化;要把反思性认可和事后适应区分开。
- 是 / 信念 / 偏好 三层边界 — 当一个选择发生转变时,要分清*「我的事实错了」(是)、「我的信念错了」(信念)、和「我的价值本身移动了」*(偏好)。游戏应帮助定位是哪一层变了。
- 最少被审视的维度 — 信息量最高的一步,是浮现那个此人从未权衡过的因素(如日常社群生活)。这是调查员智能体的目标选择规则。
- 反思均衡(反思均衡) — 停止条件:一个在充分暴露 + 偏差纠正之后依然成立的决定。成本接受闸门(「我能接受,我不后悔」)把它操作化。
EN: The hard part: preferences aren’t a fixed prior to be read off — they’re partly formed in the living. The checklist must respect this, not assume it away.
- Path-dependence of formation — what you’ll come to want depends on what you live first; the A/B simulation exists precisely to expose this (live both, then compare).
- Adaptive preferences (酸葡萄) — people rationalize after choosing; distinguish a reflective endorsement from a post-hoc adaptation.
- The is / belief / preference boundary — when a choice shifts, separate “my facts were wrong” (is) from “my belief was wrong” (belief) from “my values themselves moved” (preference). The game should help locate which layer changed.
- Least-examined dimension — the highest-information move is to surface the factor the person has never weighed (e.g. communal daily life). This is the investigator agent’s target-selection rule.
- Reflective equilibrium (反思均衡) — the stopping condition: a decision that holds after full exposure + bias correction. The cost-acceptance gate (“我能接受,我不后悔”) operationalizes it.
IV. 游戏如何浮现每一项(操作层)· How the game surfaces each (operational)
中文:
- 强制权衡 — 每一站都让玩家在某个目标因素会咬人的地方排序选项(可读性要求如此;见上面的设计原则)。
- 亲历式模拟胜过说教 — 让他们亲历35岁天花板 / 冷门专业的日常现实,而不是被告知(地基层提供真实细节)。
- 反事实 A vs B — 两种人生都活一遍,再反向推演,从而让机会成本被感受到。
- 偏差命名式干预 — 当某个选择模式散发出区块 II 偏差的气味时,镜子会温和地把它浮现为一个待核查的假设,绝不下定论。
- 最少被审视维度的探查 — 协调者追踪玩家尚未直面过哪些区块 I 因素,并把场景导向它们。
EN:
- Forced tradeoffs — every stage makes the player rank options where a target factor bites (legibility requires this; see the design principle above).
- Lived simulation over telling — they experience the 35岁 ceiling / the 冷门 daily reality, not get told about it (grounding layer supplies the real specifics).
- Counterfactual A vs B — both lives lived, then backward-reasoned, so opportunity cost is felt.
- Bias-naming interventions — when a choice pattern smells of a Block-II bias, the mirror gently surfaces it as a hypothesis to check, never a verdict.
- Least-examined-dimension probing — the coordinator tracks which Block-I factors the player hasn’t yet confronted and routes scenes toward them.
还缺什么 / 开放问题 · What else / open
中文:
- 区块 I 很可能还漏了 身体与心理健康负荷(专业/职业对身体心理的真实消耗)、性别相关考量(行业性别结构、婚育与职业的张力)、以及 运气/时机(毕业即遇行业寒冬)。随语料增长再补。
- 已整合: ZH 的
叶晓阳_Jan2026_反思均衡deck 现已折叠并入下文(读自reflection-game/developer/)。
EN:
- Block I likely still misses physical & mental health load (专业/职业对身体心理的真实消耗), gender-specific considerations (行业性别结构、婚育与职业的张力), and luck/timing (毕业即遇行业寒冬). Add as the corpus grows.
- Integrated: ZH’s
叶晓阳_Jan2026_反思均衡deck is now folded in below (read fromreflection-game/developer/).
摘自 叶晓阳 的 反思均衡 deck(2026年1月)— 已整合 · From 叶晓阳’s 反思均衡 deck (Jan 2026) — integrated
中文:ZH 的 deck 为整个项目搭了框架,并补入了具体机制。关键要点:
- 理论锚点 — 行为经济学 2.0。 该 deck 建立在 Ludwig, Mullainathan, Pink & Rambachan (2025), “Algorithms as a Vehicle to Reflective Equilibrium: Behavioral Economics 2.0.” 之上。论点是:BE 1.0 助推人们去向规划者认为好的东西(家长主义),而 BE 2.0 用算法帮一个人达到他自己的反思均衡——即他在充分理解选项后会认可的选择。这正是本工具的立场(浮现,而非推荐),并直接接上 LfL / 理想偏好这条线。
- 反思均衡 = Rawls。 「一个人在充分理解信息、充分理解选项、并经过审思之后会做出的选择。」一个处于反思均衡的学生:知道这个专业到底学什么/做什么,知道这个选择放弃了什么,且不会被他人的劝说所动摇。
- 框架稳定性测试
(新机制)(已退役 2026-06-17,见 per-stage-anatomy Legacy)。 把同一个暂定选择放在不同框架下呈现——恐惧 / 机会 / 权威——看它是否动摇。处于反思均衡的选择是框架不变的;动摇则定位出一个区块 II 偏差。→ 把它折进游戏,作为对当前 #1 的稳定性探针。 - 招生简介翻译(新机制)。 把官方 专业 描述 → 一个「说人话」的大白话版本 → 它没告诉你的部分(如「2016 新专业 / 数学极重 / 技术迭代快 / 热门 = 竞争激烈」)。地基层的事实-vs-观点拆分已经提供了其中大部分。
- 那几个框架数字: 70%+ 的学生后悔自己的专业选择;3000+ 所大学;700+ 个本科专业——并且 「没有最好的专业,只有最适合的专业。」
EN: ZH’s deck frames the whole project and adds concrete mechanics. Key takeaways:
- Theoretical anchor — Behavioral Economics 2.0. The deck is built on Ludwig, Mullainathan, Pink & Rambachan (2025), “Algorithms as a Vehicle to Reflective Equilibrium: Behavioral Economics 2.0.” The thesis: where BE 1.0 nudges people toward what a planner thinks is good (paternalism), BE 2.0 uses algorithms to help a person reach their own reflective equilibrium — the choice they’d endorse after fully understanding the options. That is exactly this instrument’s stance (surface, don’t recommend) and it ties straight to the LfL / ideal-preference line.
- 反思均衡 = Rawls. “The choice a person would make after fully understanding the information, fully understanding the options, and deliberating.” A student at reflective equilibrium: knows what the major really studies/does, knows what the choice gives up, and won’t be swayed by others’ persuasion.
- Framing-stability test (new mechanic). Present the same tentative choice under different frames — fear / opportunity / authority — and see if it wavers. A choice at reflective equilibrium is frame-invariant; wavering localizes a Block-II bias. → fold this into the game as a stability probe on the current #1.
- Admissions-blurb translation (new mechanic). Take the official 专业 description → a “说人话” plain-language version → what it didn’t tell you (e.g. “2016 new major / very high math / fast tech churn / 热门 = fierce competition”). The grounding layer’s fact-vs-opinion split already supplies most of this.
- The framing numbers: 70%+ of students regret their major choice; 3000+ universities; 700+ undergrad majors — and “there is no best major, only the best-fit major.”
叶晓阳 的终局自检(作为游戏的反思均衡检查清单逐字使用)· 叶晓阳’s end self-check (use verbatim as the game’s 反思均衡检查清单)
- 我了解这个专业具体学什么课程
- 我知道毕业后主要做什么工作
- 我考虑过这个选择要放弃什么
- 我和在读学生或从业者聊过
- 即使别人反对,我也不会轻易动摇
- 我考虑过城市、学校、专业的优先级