高级数据工程师
Senior Data Engineer
我们是谁:
Alpaca 是一家总部位于美国的全球领先代理优先经纪基础设施公司,提供股票、ETF、期权、加密货币、固定收益、24/5 小时交易等服务。
在我们的子公司中,Alpaca 是一家获得许可的金融服务公司,通过我们机构级 API 为全球 40 个国家的数百家金融机构服务。这包括经纪自营商、投资顾问、财富管理公司、对冲基金和加密货币交易所,总计超过 1000 万个交易账户。
我们的全球团队是由经验丰富的工程师、交易员和经纪专业人士组成的多元化团队,致力于实现我们让地球上每个人都能享受金融服务的使命。我们高度重视开源贡献,积极培育活跃的社区,持续提升我们获奖的、开发者友好的 API 及其背后的强大基础设施。
Alpaca 获得了来自顶级全球投资者的 4 亿美元资金支持,包括 Portage Ventures、Spark Capital、Tribe Capital、Social Leverage、Horizons Ventures、Opera Tech Ventures、SBI Group、Derayah Financial、Unbound、Peak XV、Elefund 和 Y Combinator。
我们的团队成员:
我们是一个由 400 多名分布在世界各地的成员组成的充满活力的团队,大家喜欢在世界最喜爱的地方工作,团队成员遍布美国、加拿大、日本、匈牙利、尼日利亚、巴西、英国等!
我们正在寻找热忱的个人,希望为 Alpaca 的快速发展做出贡献。如果你认同我们的核心价值观——保持好奇、富有同理心、承担责任,并准备好产生重大影响,我们鼓励你申请。
你的角色:我们正在寻找一位高级数据工程师,帮助设计和构建我们下一代数据平台,随着我们继续拓展更大的客户和新的司法管辖区。在 Alpaca,数据工程涵盖金融交易、客户数据、API 日志、系统指标、增强数据以及影响内部和外部利益相关者决策的第三方系统。我们每天处理数亿个事件,随着我们引入新客户和产品,这一数字还在持续增长。
我们在数据栈中优先使用开源技术,同时以 Google Cloud Platform(GCP)作为数据基础设施的基础。这包括批处理和流式摄入、转换和消费层,用于 BI/报告、AI/代理接口(MCP)和外部第三方数据源。我们还负责数据实验。
查看英文原文
Who We Are:
Alpaca is a US-headquartered, global leader in agent-first brokerage infrastructure for stocks, ETFs, options, crypto, fixed income, 24/5 trading, and more.
Amongst our subsidiaries, Alpaca is a licensed financial services company, serving hundreds of financial institutions across 40 countries with our institutional-grade APIs. This includes broker-dealers, investment advisors, wealth managers, hedge funds, and crypto exchanges, totalling over 10 million brokerage accounts.
Our global team is a diverse group of experienced engineers, traders, and brokerage professionals who are working to achieve our mission of opening financial services to everyone on the planet. We're deeply committed to open-source contributions and fostering a vibrant community, continuously enhancing our award-winning, developer-friendly API and the robust infrastructure behind it.
Alpaca is proudly backed by $400 million in funding from top-tier global investors including Portage Ventures, Spark Capital, Tribe Capital, Social Leverage, Horizons Ventures, Opera Tech Ventures, SBI Group, Derayah Financial, Unbound, Peak XV, Elefund, and Y Combinator.
Our Team Members:
We're a dynamic team of 400+ globally distributed members who thrive working from our favorite places around the world, with teammates spanning the USA, Canada, Japan, Hungary, Nigeria, Brazil, the UK, and beyond!
We're searching for passionate individuals eager to contribute to Alpaca's rapid growth. If you align with our core values—Stay Curious, Have Empathy, and Be Accountable—and are ready to make a significant impact, we encourage you to apply.
Your Role: We are seeking a Senior Data Engineer to help design and build the next generation of our Data Platform as we continue to scale to larger customers and new jurisdictions. At Alpaca, Data Engineering encompasses financial transactions, customer data, API logs, system metrics, augmented data, and third-party systems that impact decision making for both internal and external stakeholders. We process hundreds of millions of events daily, and this number continues to grow as we onboard new customers and products.
We prioritize open-source technologies in our data stack while leveraging Google Cloud Platform (GCP) as the foundation for our data infrastructure. This spans batch and stream ingestion, transformation, and consumption layers for BI/Reporting, AI/agent interfaces (MCP), and external third-party sinks. We also oversee data experimentation, cataloging, and monitoring/alerting systems.
Our team is 100% distributed and remote.
Responsibilities:
- Design, build, and evolve the core data platform infrastructure e.g., distributed query engines, orchestration, warehousing, cataloging, and more.
- Own our lakehouse infrastructure as code, managing deployments through Terraform and Ansible on Kubernetes.
- Build and maintain low-latency streaming and CDC ingestion pipelines, as well as batch ingestion paths landing in Iceberg.
- Develop and scale our BI landscape so downstream teams and agents get performant, self-serve access to lakehouse data.
- Enforce platform reliability best practices, including monitoring and alerting, on-call rotations, incident response, maintenance windows, runbooks, and SLAs.
- Partner with DevOps, Analytics Engineering, and other stakeholders to close infrastructure gaps and support new data requirements.
Must-Haves:
- 5+ years of experience in Data Engineering, including 2+ years building and operating scalable, low-latency data platforms handling > 100M events/day.
- Strong hands-on experience running data infrastructure on Kubernetes, with cloud-native tooling like Docker and Helm.
- Production experience with IaC: Terraform, Ansible, and ArgoCD (or equivalents).
- Deep knowledge of distributed systems (storage, transactions, and query processing) with hands-on experience operating open-source query engines like Trino or Presto.
- Strong experience with object storage and open table formats, specifically Apache Iceberg.
- Experience with streaming and CDC systems: Kafka, Redpanda, and Debezium.
- Hands-on experience with orchestration frameworks (Airflow) and ELT tools (Airbyte).
- Strong working knowledge of Python and SQL for building pipelines and platform tooling.
- Experience with Google Cloud Platform and its data services (GCS, Cloud Build, Cloud SQL, Dataproc, etc); or related experience with other cloud services.
- Ability to thrive in a fast-paced startup environment and adapt infrastructure to rapidly changing needs.
Nice to Haves:
- Experience with semantic/metrics layers (Cube, dbt, Looker).
- Familiarity with transformation frameworks (dbt).
- Familiarity with reverse ETL tooling (Hightouch)
- Familiarity with data catalog and lineage tooling (OpenMetadata, Datahub)
- Experience with data access control and governance frameworks (Apache Ranger)
How We Take Care of You:
- Competitive Salary & Stock Options
- Health Benefits
- New Hire Home-Office Setup: One-time USD $500
- Monthly Stipend: USD $150 per month via a Brex Card
Alpaca is proud to be an equal opportunity workplace dedicated to pursuing and hiring a diverse workforce.
Recruitment Privacy Policy