
澳蒂华
Production Operations Engineer (Trading Systems / SRE / Application Support)
Production Operations Engineer (Trading Systems / SRE / Application Support)
发布于 大约 15 小时前普通员工/个人贡献者
上海市
高级经验
全职员工
仅现场办公
学历未注明
信息技术与基础设施
CI/CD
Sre
低延迟
分布式系统
网络
自动化
高频交易
AI 估算 · 30k–50k
上海高级运维工程师,高频交易行业技术门槛高,薪资竞争力强,base 30-50K,年终奖丰厚。
职位详情
关于这个职位
该职位属于 Optiver 全球生产运维团队,主要负责维护和扩展低延迟交易系统,确保交易环境的稳定性与高性能
你将与交易员、开发团队紧密协作,进行部署、监控、故障排查及自动化优化,是保障高频交易业务顺利运行的核心角色
最低要求
Technical Skills:
Solid experience managing complex distributed systems in Linux environments (L2/L3 support), with a strong understanding of networking fundamentals, build, deployment, and version control systems.
Fundamental efficiency with CI/CD tooling and scripting in Python or similar languages to automate tasks.
Experience with relational databases (e.g., PostgreSQL) or other database systems.
Demonstrated ability to drive system improvements and work across global teams to deliver changes.
Non-Technical Skills:
A collaborative team player with a positive attitude, capable of coordinating effectively across departments.
A pragmatic and logical thinker who thrives under pressure in high-intensity environments, with strong troubleshooting skills and the ability to resolve incidents efficiently.
Effective communicator, able to clearly explain technical concepts and ask insightful questions.
Highly organized, able to manage multiple responsibilities while keeping stakeholders informed.
Passionate about technology, curious about the interaction between hardware and software, and driven to learn and grow by embracing frequent change.
工作职责
Provide technical expertise and support to high-frequency trading floors, troubleshooting complex issues and resolving incidents in a fast-paced, dynamic setting.
Working hours will involve a combination of standard business hours and rotational weekend for weekend trading support and test (with time off during the week).
Deploy, maintain, and monitor proprietary trading applications, ensuring stability and performance in the trading environment.
Review and analyze incidents to identify recurring themes, driving improvements to enhance automation and operational efficiency.
Coordinate deployments of new trading products, systems, and exchanges, ensuring readiness for events such as corporate actions or special announcements.
Architect and maintain services to maximize reliability, ensuring the trading systems operate within performance thresholds.
Collaborate with development teams to establish best practices and ensure the long-term stability of production systems, even under extreme conditions.
Set standards for system configuration and deployment to minimize risks and support smooth scaling.
Manage the colocation to ensure the right level of redundancy and capacity for growth is available
Automate system management processes, while maintaining control over critical operations.
优先资格
Prior experience in high-frequency trading environments is a plus but not required.
AI 洞察
优缺点分析
优点
- 加入全球顶级做市商,接触最前沿的低延迟交易技术和大规模分布式系统
- 与优秀同事共事,技术氛围浓厚,内部培训和个人发展机会丰富
- 薪资和奖金极具竞争力,福利完善(如免费餐食、健身、按摩等)
- 需适应轮班制(包括周末),工作强度大,对快速响应能力要求高
- 技术栈深入且复杂,学习曲线陡峭,要求持续跟进新知识
- 该职位适合热爱技术、抗压能力强、渴望在高速变化环境中成长的运维或 SRE 工程师,对金融交易系统有浓厚兴趣者尤佳
缺点 / 挑战
- 高频交易环境压力大,错误成本高,需保持高度专注和严谨
角色解读
- 可从运维工程师向高级 SRE 或架构师发展,深入优化系统性能和可靠性
- 有机会转向交易系统开发或量化研究,接触核心业务逻辑
- 在全球化团队中积累跨地域协作经验,未来可晋升为技术主管或团队负责人
- 负责高频交易生产环境的日常运维,包括交易系统部署、监控和故障排查,确保低延迟和高可用性
- 与交易员和开发团队紧密合作,快速响应事故并驱动根本原因分析,提升系统自动化水平
- 参与周末轮班支持交易测试,并协调新产品上线和交易所变更,保证业务连续性
- 熟练掌握 Linux 系统管理及分布式系统故障排查,精通网络协议和 CI/CD 工具链
- 具备 Python 或类似语言的脚本开发能力,能够编写自动化运维工具
- 熟悉关系数据库(如 PostgreSQL)及版本控制系统,有大规模集群管理经验
- 具有强烈的责任心和抗压能力,能在快节奏环境中高效解决问题
申请策略
- 展示对 Optiver 公司文化和业务的了解,强调自己适应快节奏、注重协作的团队环境
- 准备一个你曾主导的系统稳定性或自动化改善项目,详细说明过程和量化结果
- 突出 Linux 系统管理、分布式系统故障排查的实战经验,用具体案例说明如何解决复杂问题
- 强调 Python 或 CI/CD 自动化脚本的项目成果,展示提升效率的能力
- 如果有金融行业或低延迟系统背景,务必重点描述
- 体现跨团队协作和事故驱动改进的经历
- 深入学习网络协议(如 TCP/IP、UDP 多播)和低延迟优化技术
- 补充 SRE 常用工具(如 Prometheus、Grafana、Kubernetes)的实践
面试指南
- 使用 STAR 法则(情境-任务-行动-结果)结构化回答经历类问题
- 对于技术方案类问题,先明确目标,再分步骤阐述设计思路,强调权衡和验证
- 体现系统性思维:不仅解决眼前问题,还考虑根本原因和后续预防
- 描述一次你如何定位和解决一个复杂分布式系统故障的经历
- 谈谈你如何设计自动化方案来减少人工干预
- 你如何平衡系统稳定性和快速迭代的需求?
- 如果你发现交易系统延迟异常,你会怎么排查?
- 你对高频交易环境下的监控和告警有什么看法?
职位点评
73
综合评分
顶级做市商的核心运维岗,薪资优厚、技术前沿,但需要轮班,工作强度大。
从薪资福利、成长空间、工作节奏和岗位方向综合评估,方便横向比较。
更适合这类人
最看重技能成长与前沿技术挑战,能接受高强度工作节奏的求职者。
表现最好
成长发展
相对薄弱
工作生活
薪资福利85
成长发展90
工作生活40
使命价值60
薪资福利
85较高
公司提供行业领先的绩效奖金结构和丰富福利,薪资竞争力强,但未明确具体薪资范围。
薪资信号未披露(AI估算:30K-50K/月)
福利待遇Performance-based bonus structure、Training, mentorship and personal development opportunities、Daily breakfast, lunch and snacks、Gym membership, sports and leisure activities、Weekly in-house chair massages、Regular social events, clubs and Friday afternoon drinks
成长发展
90较高
接触前沿低延迟技术和复杂分布式系统,公司提供培训、导师制和个人发展机会,成长空间巨大。
技术前沿前沿/新兴技术
技术栈Linux、分布式系统、低延迟、Python、CI/CD、PostgreSQL、网络
成长机会Training, mentorship and personal development opportunities
业务类型profit_center
工作生活
40较低
需要周末轮班,工作强度高,办公地点在上海可能通勤时间较长,但提供调休和丰富福利。
工作模式仅现场办公
办公地点市区核心地段
加班情况JD含高强度暗示词
使命价值
60中等
所在行业稳定增长,公司注重市场流动性提供和社会责任,但职位主要为技术运维,社会意义感一般。
行业发展稳定成熟行业
社会影响中性/一般
创新程度积极采用新技术
澳蒂华 的其他在招职位
相似职位推荐
Watch Jobs