模拟面试日:国内/海外各一套全流程,自评
完整走一遍国内和海外风格各一套模拟面试流程,用评分表给自己打分,列出接下来重点要补的弱项。
今日目标
- 能完整走完一套国内风格的模拟面试全流程(自我介绍到反问环节)
- 能完整走完一套海外风格的模拟面试全流程(含行为面试与系统设计)
- 能用评分表给自己打分,并列出至少 3 条具体的弱项清单
昨天(D27)你把三个仓库包装成了简历要点、README 和一段 demo 视频。那些东西都是静态的:别人可以慢慢看,你也可以改到满意为止,改二十遍没人知道。今天要做的事正好相反——坐下来,开计时器,从第一句话开始不能重来。读完回到页面顶部把三条目标勾掉。
小白版讲解
国内那一套:63 分钟被切成五段,每段考的不是同一件事
考过驾照的人多半有过同一种经历:跟教练练车时开得挺顺,可一坐上考试车,副驾驶换成一位手里拿着扣分表的考官,起步就熄火。车没换,路线也是练过的,变的只是「此刻有人正在按条目给你记分」这件事。
面试是一模一样的场景。你在自己电脑前对着 README 讲得头头是道,换成一个陌生人隔着屏幕盯着你、还在本子上写字,第 30 秒就会开始听见自己的声音发飘。所以真考之前要先模拟一次。但真正让人进步的不是模拟本身,是那张扣分表:它把「开得不好」这种没法练的评价,拆成了「打转向灯不足 3 秒就转向」这种可以单独拎出来练二十遍的具体动作。今天这一整天的产出,就是你自己的那张扣分表。
先看国内技术面这一套的形状。一轮技术面通常是 60 分钟出头,被切成五段,每一段考官在判断的东西完全不同:
| 环节 | 时长 | 面试官在判断什么 | 最常见的扣分 |
|---|---|---|---|
| 自我介绍 | 3 分钟 | 你能不能自己组织一段有结构的表达 | 从大学讲起,把 3 分钟讲成简历朗读 |
| 项目深挖 | 25 分钟 | 这些事到底是不是你做的 | 全程说「我们」,听不出你负责哪一块 |
| 手撕代码 | 20 分钟 | 你写代码时会不会边想边说 | 闷头写 15 分钟,写完才开口 |
| 场景与八股 | 10 分钟 | 你的知识有没有边界感 | 不知道也硬答,被追问两层就崩 |
| 反问 | 5 分钟 | 你有没有真的想清楚要来干什么 | 只问一个官网上就能查到的问题 |
这五段里权重最高的是项目深挖那 25 分钟,而它的判据只有一条:这件事到底是不是你做的。 判断方法不是听你讲做了什么,是听你能不能说出「为什么不那么做」。做过的人身上会留下选择的痕迹——你知道当时还有另外两条路、知道它们各自的代价、知道你为什么放弃。没做过的人只能复述结果,追问到第二层就开始造。所以 D27 那句 STAR 公式在这里要反过来用一次:简历上写的是结论,面试里被问的是产生这个结论的过程。
手撕代码那 20 分钟也常被误解。写不写得出来当然要看,但同等重要的是你写的时候有没有声音:你沉默的每一分钟对面试官都是空白,而一旦你把思路说出来,走错方向时他往往会顺手把你拉回来——那不是作弊,那是他在看你能不能协作。
时间盒是硬的:你在自我介绍上多花的 2 分钟,是从项目深挖里扣掉的,没人给你补时。所以「讲不完」不是意外,是没排练过的必然结果。
麻烦在于,上面这些你坐在里面的时候一样都感觉不到。人在紧张状态下对自己的评价系统是失灵的:讲完只剩下一个模糊的「还行吧」或者「完了」,第二天什么都练不了。
海外那一套:五轮,每一轮换一批人从头判断你
海外的流程和国内最大的差别不在题目,在结构。国内常见的是一天之内两到三轮技术面,同一批人越问越深;海外是拉长到几周的一串独立环节,每一环由不同的人负责,各判各的,最后合起来投票。
| 轮次 | 时长 | 这一轮真正在看什么 | 常见的坑 |
|---|---|---|---|
| recruiter screen | 30 分钟 | 背景是否匹配、动机、薪资区间、工作许可与时区 | 当成闲聊,薪资区间当场编一个 |
| technical / coding | 60 分钟 | 边写边讲、边界用例、复杂度 | 全程沉默,最后 5 分钟才发现思路不对 |
| system design | 60 分钟 | 澄清、取舍、失败模式 | 五步只走到架构图,没留权衡的时间 |
| behavioral | 45 分钟 | STAR,尤其是 R 里的那个数字 | 三个问题用同一个故事,而且都没有 R |
| team match | 视情况 | 双向选择:你也在面这个团队 | 以为是走过场,一个问题都没准备 |
第一轮 recruiter screen 最容易被轻视。它由招聘顾问而不是工程师来做,问的都是「你为什么想换」「你期望的薪资范围」「你能在哪个时区工作」这类问题——但它是一个真实的闸门,答得含糊就没有下一轮。薪资区间尤其要提前想好一个区间和一句依据,现场临时编一个数字,后面所有环节都会被这个数字压着。
behavioral 那 45 分钟是国内候选人最不适应的一轮。它不问技术,只问经历:「讲一次你和同事意见不合的时候」「讲一次你搞砸了的事」「讲一次你在信息不全的情况下做决定」。应对方法是提前准备 2 到 3 个故事,然后让它们复用——同一段经历,换一个切面就能回答不同的问题。三个故事写好后按 D27 的 STAR 公式各自过一遍,重点检查 R:没有数字的 R 等于没有 R。反过来也要防一个坑:别用同一个故事回答三个问题,面试官会记得。
system design 那一轮直接接 D26 的五步和时间盒(澄清 5 分钟、估算 3 分钟、草图 8 分钟、深入 15 分钟、权衡 5 分钟)。海外这一轮尤其在意最后那 5 分钟:你放弃了什么、什么规模下这套设计会被推翻。只把图画完就停下的候选人,通常拿到的评价是「能实现,但看不出判断力」。
最后的 team match 不是走过场。它决定你进哪个组、做什么方向,而且这一轮里你的问题比你的回答更有信息量:这个组现在最头疼的问题是什么、上线节奏是怎样的、代码评审怎么做。问得出这些,对方会认为你在认真挑工作;问不出来,会被理解成你只是想要一个 offer。
两套流程的共同点是:没有任何一轮是靠临场发挥的。 都是可以排练的动作。
怎么给自己打分才不是凭感觉「还行」
多数人没有面试搭子。好消息是模拟面试的核心价值不来自对面那个人,来自事后的回放,而回放你一个人就能做。
一个人跑一轮的最小装备只有三样:录音或录屏、一个倒计时器、一份写好的题目单。 顺序也很重要——按上面两张表的时间盒设好闹钟,题目单提前写好(自我介绍、三个项目各一个深挖问题、一道编码题、一道系统设计题、两个行为面试问题),然后一次性从头走到尾,中间不许暂停、不许重来、不许查资料。中断一次,这次记录就作废了,因为真实面试不会给你暂停键。
想要有人追问,就让大模型来当这个面试官。关键不在提示词写得多漂亮,在于你必须先把追问纪律写死,否则它会变成一个不停夸你的助手:
角色:你是面试官,负责 Agent 工程师岗位的项目深挖环节,共 25 分钟。
纪律 1:一次只问一个问题,问完就停,等我回答,不要自问自答。
纪律 2:我每答完一次,你必须基于我刚说的内容再追问一层,连续追问三层再换话题。
纪律 3:全程不要评价、不要夸奖、不要补充我漏掉的内容。
纪律 4:环节结束后,按五个维度打 1 到 5 分,并逐条指出我哪一句话触发了扣分。
材料:以下是我的简历要点和三个项目的一句话介绍。跑完之后再打分,而且必须对着回放打,不要凭结束时的感觉打。刚讲完的那十分钟里人的自我评价偏差最大,讲得顺就觉得都很好,卡过一次就全盘否定。评分表固定五个维度、每个维度 1 到 5 分,满分 25 分:
| 维度 | 1 分 | 3 分 | 5 分 |
|---|---|---|---|
| 自我介绍 | 超时或跑题,听完不知道你能干什么 | 3 分钟内讲完、有结构,但亮点全靠形容词 | 90 秒讲完主线,每个亮点带一个数字,还自然引出了对方想追问的点 |
| 项目讲解深度 | 只说做了什么,讲不出为什么这么选 | 能说出方案和取舍,追问两层还答得上 | 主动说出放弃了什么、失效模式是什么、什么规模下会推翻 |
| 编码 | 长时间沉默,或者写完跑不通 | 能写出可运行的版本,被提醒后能补上边界 | 边写边讲,主动给出复杂度和三个边界用例,写完自己走一遍 |
| 系统设计 | 上来就画框,没有澄清环节 | 五步走完,但停在架构图,没有权衡 | 澄清带数字、估算有依据、深入包挑得准、结尾说清放弃了什么 |
| 沟通与反问 | 答非所问,反问环节说「我没有问题」 | 表达清楚,能问出一两个真实问题 | 卡壳时主动说出思路、会先确认题意、反问看得出做过功课 |
两条阈值规则是硬的,不要自己放宽:任何一个维度低于 3 分,就必须进弱项清单;总分低于 18 分,建议隔两天把整套流程重跑一次,而不是继续往下走。隔两天是为了给中间的针对性练习留出时间,当天重跑只是把同一段话再背熟一点。
把「讲得不好」拆成明天能练的动作
打完分,最后一步是把低分翻译成动作。用固定的四列,一条弱项一行:现象 → 根因 → 明天要练的最小动作 → 怎么验证。这四列是明天(D29)的输入,格式不要改。
第一列「现象」是最容易写坏的一列。判据很简单:现象必须是回放里能被指着看的一个事实,带时间、带原话。「讲得不好」不是现象,「项目讲解到第 8 分钟还没说到我自己做了什么」才是现象;「系统设计答得虚」不是现象,「开场 40 秒就开始画框,全程没问过一个澄清问题」才是。写不出这种句子,通常说明你没有真的回放,只是在凭印象总结。
第二列「根因」要往下追一层,追到行为习惯或缺失的准备,而不是停在「不熟练」。「8 分钟没说到自己」的根因往往是按时间顺序讲——从立项讲起,自然要讲很久才轮到你;「不敢开口」的根因往往是从来没练过边想边说,习惯了想清楚再动手。根因写对了,第三列几乎是自动生成的。
第三列「最小动作」有一条硬约束:必须是明天一天之内能做完的一件具体的事。「加强表达能力」不是动作,「把三条项目要点按 STAR 重写,开口第一句就是结论」才是。第四列「怎么验证」必须是可观察的现象,最好带一个数字门槛,否则你明天照样不知道自己练没练成。
一份填好的清单长这样:
| 现象 | 根因 | 明天要练的最小动作 | 怎么验证 |
|---|---|---|---|
| 项目讲解到第 8 分钟还没说到我做了什么 | 按时间顺序从立项讲起,没有先给结论 | 用 STAR 公式把三条项目要点各重写一遍,开口第一句就是结论 | 重录一遍,90 秒内必须出现「我负责的是」,超时算不通过 |
| 手撕代码前 3 分钟一句话没说 | 习惯想清楚再动手,没练过边想边说 | 拿一道已经会的题,全程出声讲思路 | 回放录音,最长的一段沉默不超过 15 秒 |
| 系统设计开场 40 秒就开始画框 | 没把「澄清 5 分钟」当成必答题 | 背下 D26 的五个澄清问题,练三遍开场 | 计时,第 5 分钟之前不出现任何架构名词 |
清单至少 3 条,但也不要贪多。超过 5 条就说明你在罗列缺点,而不是在挑要练的东西——明天只有一天,能真正练完的通常就是 3 条。挑的原则是先修那些「一改就能被听出来」的:开口顺序、有没有数字、有没有声音,这三类的投入产出比远高于再多背一个知识点。
源码导读
动手实验
今天的产出不是代码,是一张评分表和一张弱项清单。所以没有 starter 和 solution,只有录音文件和两张表——但它们是明天唯一的输入,糊弄过去的部分明天全都练不了。
动手之前先把三样东西准备好:题目单(自我介绍、三个项目各一个深挖问题、一道编码题、一道系统设计题、两个行为面试问题)、倒计时器、录音或录屏工具。国内那套建议上午跑,海外那套下午跑,中间隔开,避免第二套只是把第一套的话再说一遍。
- 按国内五段的时间盒跑完整一轮:自我介绍 3 分钟、项目深挖 25 分钟、手撕代码 20 分钟、场景与八股 10 分钟、反问 5 分钟。全程录音,中途不暂停、不查资料、不重来。
- 按海外五轮跑第二套:recruiter screen 30 分钟、coding 60 分钟、system design 60 分钟、behavioral 45 分钟。behavioral 前先写好 2 到 3 个可复用的故事,system design 直接套 D26 的五步时间盒。
- 回放录音,对照五个维度各打 1 到 5 分,把总分算出来。每个分数旁边写一句「因为第几分钟说了什么」,写不出来的分数不算数。
- 把所有低于 3 分的维度,按现象、根因、最小动作、怎么验证四列展开成弱项清单,至少 3 条,不超过 5 条。
- 把清单存成一个文件放在手边。如果总分低于 18 分,在日历上标一个「隔两天重跑」的时间,明天先照清单练。
面试题
今天 3 道题在下方题库区,全是流程与自评类的题——技术题前 27 天已经出满了,今天补的是「你怎么准备、怎么复盘」这一面,而这类问题在 recruiter screen 和 behavioral 环节出现的频率比你想象的高。展开后先看「分析过程」再看要点。
检查清单与明日预告
- 能完整走完一套国内风格的模拟面试全流程(自我介绍到反问环节)
- 能完整走完一套海外风格的模拟面试全流程(含行为面试与系统设计)
- 能用评分表给自己打分,并列出至少 3 条具体的弱项清单
- 能说出国内五段和海外五轮各自在判断什么,以及两者结构上的根本差别
- 五个维度的分数旁边都写得出「因为第几分钟说了什么」
- 弱项清单四列填满,每条的最小动作都能在一天内做完、都有可观察的验证方式
- 3 道面试题不看要点也能答出至少 2 道
明天(D29)我们照着这张清单一条条补,同时把四道高频编码题(限流器、LRU、并发控制、流式 JSON 解析)过一遍热身。顺序是有意的:先有清单再练,练的才是你真正缺的那两小节;今天不把弱项落到纸上,明天的练习就会退化成从头到尾泛泛复习一遍,看着很努力,弱的地方一个都没动。
面试题库
国内和海外技术面试的流程差异主要在哪里?你会怎么分别准备?How do domestic Chinese and overseas tech interview loops differ structurally, and how would you prepare for each?
国内高频海外高频基础#interview-process#career分析过程 · 先想清楚再作答
- 这题看着像常识题,区分度其实在于你有没有真的按流程准备过。只答「海外有 behavioral、国内有八股」是在复述听说,面试官听不出你排练过。
- 先给结构这条主线,其余差异都是它的推论:国内通常是一天之内两到三轮,同一批人越问越深,单轮 60 分钟出头切成五段——自我介绍 3 分钟、项目深挖 25 分钟、手撕代码 20 分钟、场景与八股 10 分钟、反问 5 分钟;海外是拉长到几周的五个独立环节——recruiter screen 30 分钟、technical/coding 60 分钟、system design 60 分钟、behavioral 45 分钟、team match,每一环由不同的人负责,各判各的,最后合票。
- 由结构推准备策略,这一步才是答案的价值所在:同一批人越问越深,意味着国内的胜负手在项目深挖那 25 分钟,要练的是被追问三层还答得上;独立环节合票意味着海外任何一轮都能单独把你否掉,所以短板比长板重要,尤其是多数人从没排练过的 behavioral。
- 第三条差异是评价载体:国内更依赖面试官当场的主观印象,海外多数公司有结构化的评分维度和书面反馈,所以「边写边讲」「主动说出取舍与失败模式」这类能被写进反馈的行为,在海外权重更高。
- 要主动澄清一个常见误区:差异不是「海外不考算法」。coding 那 60 分钟照样是算法题,区别在于它明确要求你全程出声,沉默本身就会被扣分。
- 可以预期的追问:那准备时间怎么分配?答共同部分先练——项目深挖和系统设计两套流程都要考,投入产出比最高;剩下的按目标市场补,投海外就补 2 到 3 个可复用的 STAR 故事,投国内就补知识的边界感(不知道就说不知道,再说出你会怎么查)。
How to reason about it · think before answering
- This looks like trivia, but the discriminator is whether you actually rehearsed against a loop. Answering only 'overseas has behavioral, China has fundamentals drilling' sounds like hearsay.
- Lead with structure, because every other difference follows from it. A domestic loop is usually two or three rounds in a single day with the same people digging deeper each round, and one round of roughly 60 minutes splits into five segments: 3 minutes of self-introduction, 25 of project deep-dive, 20 of live coding, 10 of scenario and fundamentals, 5 of candidate questions. An overseas loop is five independent stages spread over weeks: a 30-minute recruiter screen, 60 minutes of technical/coding, 60 of system design, 45 of behavioral, then team match, each run by different people who score independently and vote at the end.
- Derive preparation from that structure, which is where the answer earns its keep. Same people digging deeper means the domestic loop is decided in that 25-minute deep-dive, so rehearse surviving three layers of follow-up. Independent stages plus a vote means any single overseas round can sink you, so weakest link beats strongest link, especially behavioral, which most engineers never rehearse.
- A third difference is how judgment is recorded: domestic outcomes lean on the interviewer's live impression, while most overseas companies use structured rubrics and written feedback. That makes behaviors which can be written down — narrating while coding, volunteering trade-offs and failure modes — worth more overseas.
- Correct a common misconception before they raise it: the difference is not that overseas skips algorithms. That 60-minute coding round is still an algorithm round; what changes is the explicit requirement to think out loud, where silence itself costs points.
- Expect the follow-up on time allocation: train the overlap first — project deep-dive and system design appear in both loops and give the best return — then specialize, adding two or three reusable STAR stories for overseas, or the habit of naming the edge of your knowledge for domestic rounds.
答题要点
- 结构差异是主线:国内一天内两三轮、同一批人越问越深;海外五个独立环节跨几周,不同的人各判各的最后合票
- 国内单轮的五段与时间盒:自我介绍 3 分钟、项目深挖 25 分钟、手撕代码 20 分钟、场景与八股 10 分钟、反问 5 分钟
- 海外五轮:recruiter screen 30 分钟、coding 60 分钟、system design 60 分钟、behavioral 45 分钟、team match
- 准备策略由结构推出:国内练被追问三层,海外补短板(尤其 behavioral 的 2 到 3 个可复用故事)
- 海外更依赖结构化评分与书面反馈,所以边写边讲、主动说取舍这类可被记录的行为权重更高;但算法一样要考
Key points
- Structure is the through-line: domestic loops run two or three rounds in one day with the same panel going deeper; overseas loops are five independent stages over weeks, scored separately and voted on
- Domestic segments and time boxes: 3 minutes intro, 25 project deep-dive, 20 live coding, 10 scenario and fundamentals, 5 candidate questions
- Overseas stages: 30-minute recruiter screen, 60 coding, 60 system design, 45 behavioral, then team match
- Preparation follows from structure: domestic means surviving three layers of follow-up; overseas means fixing your weakest round, especially two or three reusable STAR stories
- Overseas relies on rubrics and written feedback, so narrating while coding and volunteering trade-offs count for more — but algorithms are still tested
自我介绍环节最容易出的问题是什么?一段好的自我介绍应该长什么样?What most commonly goes wrong in the self-introduction, and what does a good one look like?
国内高频海外高频进阶#self-presentation#communication分析过程 · 先想清楚再作答
- 先看清面试官在这 3 分钟里做什么:一是看你能不能自己组织一段有结构的表达,二是决定接下来 25 分钟挖你哪个项目。看懂第二件事,答案就不是「讲短一点」这么浅了。
- 最容易出的问题有一个统一的根因——按时间顺序讲。从大学讲起、顺着简历从上往下念,于是 3 分钟到点时你还没讲到最近、最有价值的那段经历,而那恰恰是唯一有人想听的部分。
- 由此推出正确形态:倒序,只留三块——你现在是什么方向的工程师、一到两个带数字的代表作、你为什么来面这个岗位。90 秒讲完主线,把剩下的时间让给对方追问,而不是把 3 分钟填满。留白是主动权,不是浪费。
- 第二个高频问题是通篇形容词、没有一个数字。「我做过一个高性能的 Agent 服务」几乎不携带信息;换成一句带约束和指标的话(在什么约束下、为了什么目标、做了什么、把哪个指标从多少改善到多少),才会让对方接着问下去。
- 第三个问题最隐蔽:自我介绍里埋了自己不想被问的东西。你说出口的每一个技术名词都是一张邀请函,不熟的栈别写也别说;反过来,希望被问的点要主动埋进去,这是全场唯一由你控制议题的机会。
- 可以预期的追问:面试官打断你说「再简单说一下」,说明你已经超时或跑题了。所以要提前排练两个版本,一个 60 秒、一个 90 秒,现场直接切,不要临场压缩——临场压缩的结果通常是把结论也一起删掉了。
How to reason about it · think before answering
- Start from what the interviewer is doing during those three minutes: judging whether you can structure a piece of speech unaided, and deciding which project to spend the next twenty-five minutes on. Once you see the second one, the answer stops being 'keep it short'.
- The usual failures share one root cause: telling it chronologically. Starting at university and reading the resume top-down means the three minutes expire before you reach the recent work, which is the only part anyone wants to hear.
- That gives the correct shape: reverse order, three blocks only — what kind of engineer you are now, one or two signature pieces of work with a number attached, and why this role. Land the main line in ninety seconds and leave room for follow-up rather than filling the slot. The silence is leverage, not waste.
- The second frequent failure is adjectives with no numbers. 'I built a high-performance agent service' carries almost no information; a sentence with a constraint, a goal, an action, and a metric moved from X to Y is what makes someone ask the next question.
- The third is the subtlest: seeding things you do not want to be asked about. Every technology you name is an invitation, so leave unfamiliar stacks out — and conversely, plant the topics you want to be asked about, since this is the only moment in the loop where you set the agenda.
- Expect this follow-up: if the interviewer cuts in with 'just briefly', you have already run long or drifted. Rehearse two versions, sixty and ninety seconds, and switch between them rather than compressing live — live compression usually deletes the conclusion too.
答题要点
- 面试官在这 3 分钟里同时做两件事:判断你的表达结构,决定接下来挖哪个项目
- 最常见的错是按时间顺序讲,时间用完还没讲到最近最有价值的经历;正确做法是倒序
- 结构只留三块:现在的技术方向、一到两个带数字的代表作、为什么来面这个岗位
- 90 秒讲完主线、主动留白给对方追问,比把 3 分钟填满更有利
- 每个说出口的技术名词都是邀请函:不熟的不提,想被问的主动埋进去;提前排练 60 秒和 90 秒两个版本
Key points
- The interviewer is doing two things at once: assessing structure and choosing which project to dig into
- The common failure is chronological order, which burns the clock before reaching recent work; reverse it
- Keep three blocks: current engineering focus, one or two signature results with numbers, and why this role
- Landing the main line in ninety seconds and leaving room for follow-up beats filling all three minutes
- Every technology you name is an invitation: omit unfamiliar stacks, plant the topics you want asked, and rehearse a sixty-second and a ninety-second version
一次模拟面试之后,你怎么做一次客观的自我评估,而不是停在「感觉还行」?After a mock interview, how do you assess yourself objectively instead of settling for 'that felt okay'?
国内高频海外高频进阶#self-assessment#deliberate-practice分析过程 · 先想清楚再作答
- 这题在考你有没有把练习工程化。答「录下来多听几遍」只是及格线,真正的区分度在于有没有可重复的评分口径和事先定好的阈值——没有口径,两次模拟之间就没法比较,也就谈不上进步。
- 客观的前提是有可回放的证据,所以先固定三件事:全程录音或录屏、按时间盒计时、事后对着回放打分而不是结束时凭感觉打。刚讲完的十几分钟里自我评价偏差最大,讲得顺就全盘肯定,卡过一次就全盘否定。
- 然后用固定维度代替整体印象:自我介绍、项目讲解深度、编码、系统设计、沟通与反问,各 1 到 5 分,满分 25。关键是每个维度要写好 1 分、3 分、5 分各长什么样的锚点描述,否则同一个「4 分」在两次之间根本不是同一件事。
- 阈值要在打分之前定好,这是防止事后给自己找理由的唯一办法:任何单项低于 3 分就进弱项清单,总分低于 18 分就隔两天把整套流程重跑一次,而不是硬着头皮往下走。
- 最后一步才是全部价值所在——把低分翻译成四列:现象、根因、最小动作、怎么验证。现象必须是回放里能指着看的事实(带时间、带原话),「讲得不好」不算,「项目讲解到第 8 分钟还没说到我做了什么」才算;最小动作必须一天内做得完;验证必须可观察,最好带数字门槛。
- 可以预期的追问:一个人怎么产生追问?用大模型当面试官,但必须先把追问纪律写进指令——一次只问一个问题、基于我的回答连追三层、全程不评价不夸奖不给答案。不写纪律,它会退化成一个不停鼓励你的助手,那就失去了模拟的意义。
How to reason about it · think before answering
- This asks whether you have engineered your practice. 'Record it and listen again' is the passing floor; the discriminator is a repeatable rubric plus thresholds fixed in advance, because without a rubric two sessions are not comparable and improvement is unmeasurable.
- Objectivity requires reviewable evidence, so fix three things first: record audio or screen throughout, run the real time boxes, and score against the recording afterwards rather than on feeling at the buzzer. Self-assessment is at its most distorted in the minutes right after you finish.
- Then replace overall impression with fixed dimensions: self-introduction, project depth, coding, system design, and communication plus candidate questions, each scored 1 to 5 for a total of 25. What makes it work is writing anchor descriptions for what a 1, a 3, and a 5 look like — otherwise the same '4' means different things in different sessions.
- Set thresholds before scoring, which is the only defense against rationalizing afterwards: any dimension below 3 goes on the weakness list, and a total below 18 means rerunning the whole loop two days later instead of pressing on.
- The final step carries all the value: translate low scores into four columns — observation, root cause, smallest drill for tomorrow, and how to verify. An observation has to be a fact you can point at in the recording, with a timestamp and the actual words: 'explained it badly' does not qualify, 'eight minutes into the project story and still had not said what I personally did' does. The drill must fit in one day, and verification must be observable, ideally with a numeric bar.
- Expect the follow-up: how do you generate follow-up questions alone? Use a model as the interviewer, but write the interrogation rules into the instructions first — one question at a time, three consecutive layers of follow-up grounded in what you just said, and no praise, no evaluation, no supplying the answer. Without those rules it degrades into an encouraging assistant, which defeats the point.
答题要点
- 先有证据再有判断:全程录音或录屏、按时间盒计时、事后对着回放打分,不在结束当场凭感觉打
- 用固定五个维度(自我介绍、项目讲解深度、编码、系统设计、沟通与反问)各 1 到 5 分、满分 25,并给每个维度写 1/3/5 分的锚点描述
- 阈值先定后打:任何单项低于 3 分进弱项清单,总分低于 18 分隔两天重跑整套流程
- 把低分翻译成四列:现象、根因、最小动作、怎么验证;现象必须是回放里能指着看的事实,动作必须一天内做得完
- 一个人练时用大模型当面试官,但要先写死追问纪律:一次一问、连追三层、不评价不夸奖不给答案
Key points
- Evidence before judgment: record throughout, run real time boxes, and score against the replay rather than on feeling at the buzzer
- Use five fixed dimensions (intro, project depth, coding, system design, communication and candidate questions) scored 1 to 5 out of 25, each with anchors for what 1, 3 and 5 look like
- Fix thresholds before scoring: any dimension below 3 goes on the weakness list, a total below 18 means rerunning the loop two days later
- Translate low scores into four columns — observation, root cause, smallest drill, verification — where the observation is a pointable fact and the drill fits in one day
- Practicing alone, use a model as interviewer but write the rules first: one question at a time, three layers of follow-up, no praise, no evaluation, no answers
评论
登录后即可参与讨论
还没有评论,来说第一句。