远程工作雷达

事件运营指挥官

Incident Operations Commander

职能支持全球可投
公司alpaca
薪资未公开
工作地点Remote - Global
地域资格全球可投
时区要求无特别要求
用工类型未标注
发布时间6 天前
数据来源Greenhouse
前往企业招聘页投递 →
全球可投:该职位未限制候选人所在地区。仍需注意薪资可能按地区折算,以及实际签约方式(正式雇佣 / 独立合同)。

我们是谁:

Alpaca 是一家总部位于美国的全球领先代理优先经纪基础设施公司,提供股票、ETF、期权、加密货币、固定收益、24/5交易等服务。

在我们的子公司中,Alpaca 是一家获得许可的金融服务公司,通过我们机构级 API 为全球 40 个国家的数百家金融机构提供服务。这包括经纪交易商、投资顾问、财富管理公司、对冲基金和加密货币交易所,总计超过 1000 万笔经纪账户。

我们的全球团队由经验丰富的工程师、交易员和经纪专业人士组成,他们致力于实现我们让地球上每个人都能享受金融服务的使命。我们高度重视开源贡献,积极培育活跃的社区,持续提升我们获奖的、开发者友好的 API 以及其背后的强大基础设施。

Alpaca 获得了来自顶级全球投资者的 4 亿美元资金支持,包括 Portage Ventures、Spark Capital、Tribe Capital、Social Leverage、Horizons Ventures、Opera Tech Ventures、SBI Group、Derayah Financial、Unbound、Peak XV、Elefund 和 Y Combinator。

我们的团队成员:

我们是一个由 400 多名分布在世界各地的成员组成的活力团队,大家在自己喜欢的地方工作,团队成员遍布美国、加拿大、日本、匈牙利、尼日利亚、巴西、英国等地!

我们正在寻找渴望为 Alpaca 的快速发展做出贡献的热情人士。如果你认同我们的核心价值观——保持好奇、富有同理心、承担责任,并准备好产生重大影响,我们鼓励你申请。

职位描述

作为 Alpaca 最关键事件的值班指挥官,协调跨职能响应以快速恢复服务,确保相关人员参与并了解情况,并确保每个事件后组织都能采取行动。

你不需要修复停机。你要确保响应可靠:正确评估严重性,让合适的工程师参与,缓解措施不会停滞,领导层及时获知,并且后续工作在电话结束后仍能持续进行。

你将负责的工作内容

  • 全程指挥事件。从声明到缓解,全程掌控,确保响应人员专注于尽快阻止客户和合作伙伴的影响。主持会议,让观察者不干扰响应人员,并在看到停滞时明确指出。
  • 分类并确定严重性。在事件声明时设定严重性,并根据事实重新评估。
查看英文原文

Who We Are:

Alpaca is a US-headquartered, global leader in agent-first brokerage infrastructure for stocks, ETFs, options, crypto, fixed income, 24/5 trading, and more.

Amongst our subsidiaries, Alpaca is a licensed financial services company, serving hundreds of financial institutions across 40 countries with our institutional-grade APIs. This includes broker-dealers, investment advisors, wealth managers, hedge funds, and crypto exchanges, totalling over 10 million brokerage accounts.

Our global team is a diverse group of experienced engineers, traders, and brokerage professionals who are working to achieve our mission of opening financial services to everyone on the planet. We're deeply committed to open-source contributions and fostering a vibrant community, continuously enhancing our award-winning, developer-friendly API and the robust infrastructure behind it.

Alpaca is proudly backed by $400 million in funding from top-tier global investors including Portage Ventures, Spark Capital, Tribe Capital, Social Leverage, Horizons Ventures, Opera Tech Ventures, SBI Group, Derayah Financial, Unbound, Peak XV, Elefund, and Y Combinator.

Our Team Members:

We're a dynamic team of 400+ globally distributed members who thrive working from our favorite places around the world, with teammates spanning the USA, Canada, Japan, Hungary, Nigeria, Brazil, the UK, and beyond!

We're searching for passionate individuals eager to contribute to Alpaca's rapid growth. If you align with our core values—Stay Curious, Have Empathy, and Be Accountable—and are ready to make a significant impact, we encourage you to apply.

Role

Serve as the on-duty commander for Alpaca's most critical incidents, directing cross-functional response to restore service quickly, keeping the right people engaged and informed, and making sure every incident leaves behind something the organisation can act on.

You do not fix the outage. You make the response reliable: correct severity, the right engineers in the room, mitigation that does not stall, leaders informed in time, and follow-up work that survives the call.

Things You Get To Do

  • Command incidents end to end. Take command from declaration to mitigation, keeping responders focused on stopping customer and partner impact as fast as possible. Run the bridge, keep observers out of the responders' way, and name a stall out loud when you see one.
  • Classify and hold the line on severity. Set severity at declaration and re-check it as facts arrive. Risk advises on financial and regulatory materiality; the call is yours.
  • Engage the right people, fast. Identify the owning team by service, symptom and blast radius, page them, and expand the responder set the moment the first team is wrong or not enough. When a page goes unanswered, escalate - and escalate the escalation. Bring in the leaders who must make business calls: feature flags, traffic shedding, failover, freeze-or-ship.
  • Hold the bridge, and protect the people fixing it. Keep engineering and technical support uninterrupted - questions from stakeholders, partners and executives come to you. Be the single source of truth to the partner communications team on impact, severity and timing: you decide when a status page update or partner contact is needed, they write and send it, and chasing a late or stale update is yours.
  • Run follow-the-sun handoffs. Deliver warm, high-fidelity handoffs across regions: current impact and severity, mitigation path and next actions, who is in the room, outstanding decisions, and what must not be dropped. The incoming commander confirms ownership before you step away - command never goes dark at a region boundary.
  • Close the loop, on the clock. Maintain the timeline of facts as the incident runs rather than reconstructing it afterwards - in a regulated business that record has to hold up long after the call ends. Once mitigated, make sure a blameless retrospective is scheduled with a named owner and a timebox, and record where the cause sits - that choice sets which follow-up items are mandatory. Every action item needs a real ticket, one named accountable, a priority and a category, delivered inside the agreed service level. If a postmortem produces nothing but low-priority items, treat that as a signal the analysis stopped early and escalate to SRE rather than passing it on.
  • Automate the coordination away. Coordination is the part of this job that should eventually belong to a machine. Every manual prompt you send - the update that is due, the question nobody answered, the partner nobody contacted - is a candidate for automation, and the direction we are heading is AI handling the routine so commanders can spend their attention on judgement. You get us there by working to the decision trees, saying where they are wrong, and being honest about which of your instincts are actually rules.

Who You Are (Must-Haves)

  • 4+ years commanding or co-commanding high-severity incidents in a production engineering, SRE or technical operations environment.
  • You direct technical responders under pressure without being the person writing the fix.
  • You make and defend crisp severity and escalation decisions, and you take charge without waiting to be asked. Command means waking senior people at 03:00, interrupting an executive, and telling an experienced engineer to stop what they are doing - with an audience watching. It is a visible, directive role and it needs to be instinctive.
  • You can read a dashboard and judge for yourself whether impact has actually stopped.
  • You communicate clearly with engineers, executives and partner-facing stakeholders - and you know the difference between briefing the comms function and speaking for the company.
  • You are comfortable holding other teams to account in the moment, across a reporting line that is not yours, without turning it into friction.
  • You thrive in a follow-the-sun model with clean cross-region handoffs.
  • You understand FinTech concepts and the trust stakes of API-driven financial platforms.
  • You use AI tools and agentic automation to reduce manual toil and speed up response.
  • You will work a regional coverage window as part of a global 24x7 Incident Commander roster.

Who You Might Be (Nice-to-Haves)

  • Formal incident command training - ITIL, Major Incident Management or crisis management.
  • Experience with modern incident management and on-call platforms.
  • You have written severity rubrics, decision trees, escalation matrices, runbooks or incident playbooks.
  • You have commanded in game days, tabletop exercises or incident simulations, not only in production.
  • You have partnered with problem management or reliability programme functions to roadmap incident follow-ups.
  • Online securities trading or capital markets experience, or another regulated, market-hours-sensitive domain.

How We Take Care of You:

  • Competitive Salary & Stock Options
  • Health Benefits
  • New Hire Home-Office Setup: One-time USD $500
  • Monthly Stipend: USD $150 per month via a Brex Card

Alpaca is proud to be an equal opportunity workplace dedicated to pursuing and hiring a diverse workforce.

Recruitment Privacy Policy

本页面信息整理自 Greenhouse,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

该公司其他在招职位

← 返回全部职位