2026 年初我走了 Stripe 的 ML engineer 面试 loop,面的是 L5 左右、fraud/risk ML 团队。几乎没人写过相关信息,所以把我记得的都写下来。
Stripe 的 ML 团队比你想象的小。大部分 ML 工作在支付欺诈、风险评分和一些内部工具上。他们不打算做 foundation models。如果你想做纯研究或 CV/NLP research,这里不是。
整个 loop: recruiter screen(30 分钟):常规,想要什么、地点、时间线、薪酬大概范围 technical phone screen(60 分钟):ML 概念 + coding 各一半。ML 部分是模型评估:欺诈分类这么严重 class imbalance 的情况下怎么评估?哪些指标重要,哪些会误导?coding 部分是一道 leetcode medium,不是 ML 相关 onsite(4 个 panel): ml depth:深挖一个项目。不只是 "what did you do(你具体做了什么)",而是 "why that model architecture(为什么用这种模型架构)"、"what did you try that didn't work(你试过哪些不奏效的方法)"、"how did you measure production impact(你怎么衡量线上影响)"。他们想确认你做过真实工作,不是只调过超参 ml system design:我拿到的是 "design an ML system to detect fraudulent transactions at stripe's scale.(设计一个能在 Stripe 这种规模下检测欺诈交易的 ML 系统。)"。这一轮信息量很大:特征工程、在线 vs. 离线推理、延迟要求、怎么处理 concept drift、如何在 prod 里安全做 A/B test software engineering:一轮 coding,更偏标准数据结构。他们会确认 Stripe 的 MLE 需要写真实生产代码,不是只写 jupyter notebooks behavioral:tell me about a model you deployed that had unexpected behavior in production. what did you do(你具体做了什么)。这是我在任何地方遇到过最好的 behavioral 题,能直接看出你是不是真的上线过东西
他们看重什么: class imbalance、fraud ML 领域知识、模型监控/可观测性。有做过 fraud 或 risk 的是大加分 能把 precision 和 recall 的取舍放到业务语境里讲清楚,不是只讲学术 软件工程是真的要做,代码要写得干净
我确实拿到了 offer。最后因为别家 equity 更好拒了。comp 大概是 base $215k,equity $500k/4yr,L5,SF。