高级数据工程师
Senior Data Engineer
### 职位描述
Wheelhouse 是一个面向超过 5000 亿美元灵活租赁市场的收入管理平台。我们的技术赋能短期和中等时长住宿的运营商,他们管理独栋房屋、公寓楼,以及在某些情况下还有酒店——这是一个广泛且庞大的可触达市场。Wheelhouse 相信更加灵活和互联的世界既是不可避免的,也是重要的,我们构建技术以确保推动这种生活方式和社会的业务能够蓬勃发展。
作为一个团队,我们喜欢按时或提前交付客户喜爱的产品,同时平衡工作与生活并一起享受乐趣。我们被描述为透明且协作的,我们努力让团队成员在职业和个人方面都能取得成功。我们是一家以远程办公为主、可以随时随地工作的公司,我们相信“健康的奋斗”是良好增长的关键。
我们核心的工作是处理大量数据,为客户提供竞争优势。我们跟踪全球超过 1400 万条房源的日历数据,每天获取每条房源的未来日历可用性和定价信息。这转化为一个数百 TB 的高度动态的数据集。
我们正在寻找一名 **高级数据工程师** 来负责这个庞大的每日数据摄入引擎。在 Ruby on Rails 和 Resque 环境中工作,后端使用 AWS RDS Postgres 和 Aurora 以及 Athena,你将确保我们的数据管道稳健、可扩展且高度优化。你不仅会监督和改进我们现有的每日数据摄入,还会设计系统来整合新的数据源,并无缝地将这些数据传递到我们的核心应用中。
你的 **使命** 是确保我们数据基础设施的可靠性、速度和规模。你将成为我们数百 TB 数据集的技术守护者,确保随着我们追踪的房源数量增加以及新增数据源,我们的数据摄入引擎能够优雅且成本有效地扩展,而不会出现任何中断。我们的收入管理算法完全依赖于这 1400 万条房源的每日日历同步,以准确地为客户的房产定价。如果我们的数据管道变慢或失败,我们的核心产品将无法运行。
### 我们寻找的人选需要具备:
**独立工作并承担责任:** 你不会等待被告知要优化什么。你将数据基础设施视为自己的产品,主动发现改进点并推动其落地。
查看英文原文
### Job Description
Wheelhouse is a revenue management platform for the $500B+ flex rental space. Our technology empowers short & mid-length stay operators, who manage single-family homes, apartment buildings, and(in some cases) hotels - a broad & massive addressable market. Wheelhouse believes a more flexible & connected world is both inevitable & important, and we build technology to make sure the businesses that enable this lifestyle & society can thrive.
As a team, we enjoy shipping products our customers love, on time or ahead of schedule, while balancing work/life & having fun together. We’re best described as transparent & collaborative, and we strive to set our teammates up for success - both professionally & personally. We’re a remote-first, work-anywhere, and“yes - you should make time for that adventure/vacation” company, who believes that“healthy hustle” is the key to good growth.
At our core, we process massive amounts of data to give our clients a competitive edge. We track calendar data for over 14 million listings worldwide, ingesting the future calendar of availability and pricing for every single listing, every single day. This translates to a multi-terabyte, highly dynamic dataset.
We are looking for a **Senior Data Engineer** to take ownership of this massive daily ingestion engine. Working within a Ruby on Rails and Resque environment, backed by AWS RDS Postgres and Aurora as well as Athena, you will ensure our data pipelines are robust, scalable, and highly optimized. You will not only supervise and improve our existing daily ingestion but also design systems to integrate new data sources and seamlessly deliver this data to our core application.
Your **mission** is to guarantee the reliability, speed, and scale of our data infrastructure. You will be the technical guardian of our multi-terabyte dataset, ensuring that as we grow to track even more listings and add new data sources, our ingestion engine scales gracefully and cost-effectively without missing a beat. Our revenue management algorithms rely entirely on the daily calendar sync of these 14 million listings to price our clients' properties accurately. If our data pipelines slow down or fail, our core product cannot function.
### We’re looking for someone who:
**Works independently and takes ownership:** You don't wait to be told what to optimize. You treat the data infrastructure as your own product, proactively identifying areas for improvement and driving them to completion.
**Thinks in systems:** You understand that scaling isn't just about writing faster code; it's about how Resque, PostgreSQL, and our AWS infrastructure interact. You see the big picture and understand the cascading effects of your changes.
**Stays calm under pressure:** Because this daily data sync is our lifeblood, pipeline hiccups are critical. You can troubleshoot complex, multi-terabyte issues methodically and keep a cool head when the pressure is on.
**Communicates technical concepts clearly:** You can easily translate complex data engineering constraints and architecture decisions into plain language for the wider engineering and product teams.
**Values pragmatism over perfection:** You know when to invest in an elegant, long-term architectural solution, and when the business just needs a robust, practical fix to keep the daily sync running smoothly.
**Is obsessed with observability:** You believe a system isn't finished until it's properly monitored. You build alerts and dashboards to catch anomalies and bottlenecks before they ever impact our clients.
**Is super detail oriented**: You intrinsically see that certain data problems can be needles in haystacks and you are process driven to build safeguards to flag the vast majority of issues before our customers notice the same.
### What you'll be doing
- Architect, optimize, and supervise the daily ingestion pipelines that process calendar data for 14+ million listings.
- Tune and scale our AWS RDS PostgreSQL and Aurora databases to handle extreme high-throughput read/write operations and multi-TB storage.
- Optimize background job orchestration to ensure timely and efficient data processing.
- Design and build robust ETL processes to integrate new, complex data sources into our ecosystem.
- Collaborate with application engineers to make diverse versions and aggregations of our data easily accessible to the core application.
- Apply DevOps best practices to maintain and improve our AWS infrastructure, CI/CD pipelines, and infrastructure-as-code.
- Build comprehensive monitoring, alerting, and observability tooling to catch data bottlenecks before they impact the business.
### Requirements
- **5+****years of Data Engineering or Backend Engineering experience**, specifically dealing with massive, high-velocity datasets(multi-TB scale) and **10+ years** overall engineering experience at technology companies.
- **Deep database expertise:** Advanced knowledge of relational databases, specifically PostgreSQL and AWS Aurora. You must know how to tune databases, optimize complex queries, and manage large-scale indexing.
- **Strong programming skills:** Proficiency in Ruby and Ruby on Rails(or a strong willingness to learn, backed by expert-level experience in a similar language like Python) to navigate our core stack.
- **Job Orchestration:** Experience with background processing frameworks at scale(Resque, Sidekiq, Celery, or similar).
- **DevOps & Cloud:** Hands-on experience with AWS cloud services, infrastructure provisioning(Terraform/CloudFormation), and CI/CD pipelines.
- **Data Modeling:** Strong ability to design data models that balance fast ingestion with efficient application querying.
- **Problem Solver:** A track record of identifying performance bottlenecks in complex distributed systems and implementing elegant solutions.
- **Remote DNA**: Demonstrated ability to work autonomously, communicate asynchronously, and manage your own"healthy hustle."
### Why Join Us
- Own the big data pipelines in a fast-growing SaaS company.
- Directly improve the reliability and usability of our product.
- Enjoy the power to shape your own projects and help set the direction for new development.
- Collaborative, small-team environment where your impact is highly visible
### Perks and Benefits
- The expected base salary range for this position is between $156,000 and $180,000 depending on experience.
- Competitive salary, cash bonus potential and equity
- Full medical, dental, and vision benefits for each US employee
- Fidelity 401k available for each US employee
- Unlimited PTO
- And…more!
#### About Us:
Wheelhouse is a fintech platform for the $500B+ flex rental space. Most specifically, we enable short & mid-length stay providers with 1 to 100,000+ listings to earn 20%+ more from their rental properties.
In 2021, our target customer segment voted our platform“Innovation of the Year” at the Data & Revenue Management conference. This sentiment is shared by our customers, as evidenced by our platforms low churn and rapidly growing ARPU. In 2022, we closed a significant funding round, with participation from many of the best tech, travel & real estate investors. We’re lucky to have a long runway, low burn, and rapidly growing revenue.
As a team, we enjoy shipping products our customers love, on time or ahead of schedule, while balancing work/life & having fun together. We’re best described as transparent & collaborative, and we strive to set our teammates up for success - both professionally & personally. We’re a remote-first, work-anywhere, and“yes - you should make time for that adventure/vacation” company, who believes that“healthy hustle” is the key to good growth.
We’re experienced business & product builders who have founded multiple companies together, know our category extremely well, and recognize how rare/special it is to be perfectly positioned around a big opportunity with a very strong cross-functional team.
We’d be eager to say hello and learn more about you!