远程工作雷达

数据工程师,拉美地区

Data Engineer LATAM

开发工程限定地区(需当地身份)与中国几乎无重叠,需长期倒时差
公司Onebeat
薪资未公开
工作地点Brazil
地域资格限定地区(需当地身份)
时区要求与中国几乎无重叠,需长期倒时差
用工类型Full Time
发布时间今天
数据来源Himalayas
前往 Himalayas 查看并投递 →
注意地域限制:该职位明确限定在 Brazil 招聘。如果你是位于中国大陆的求职者,通常需要当地工作身份才能投递,或需与雇主确认是否接受独立合同(Contractor)形式合作。
作息提示:与中国几乎无重叠,需长期倒时差。

Description
我们正在寻找一位经验丰富的数据工程师。理想的候选人应具备自我驱动力,能够同时处理多项任务,并有良好的团队合作经验。您将负责设计、开发、管理和维护我们的开源数据平台,包括我们的数据湖仓(S3、Apache Iceberg 和 ClickHouse)、ETL 流程以及编排工具(Temporal Workflow)。

What You Will Do

  • 开发一个可扩展的数据平台,整合多个数据源以便于访问。
  • 设计和优化数据工具(编排、治理、数据湖仓、BI 等)。
  • 确保数据分析、科学和工程团队的数据系统顺畅运行。
  • 在微服务环境中优化数据管道(摄入、处理和输出)。
  • 使用 Temporal 构建、维护和监控 ETL/ELT 流程并编排工作流。
  • 排查并提升数据基础设施(S3、Apache Iceberg、ClickHouse)的性能、可扩展性和可靠性。
  • 与数据科学家、分析师和后端工程师跨职能协作,理解数据需求并提供解决方案。
  • 在整个平台上实施并推广数据质量、治理和安全的最佳实践。

Requirements

  • 3 年以上数据工程师或类似数据基础设施相关岗位的经验。
  • 精通 SQL,并有数据建模的实际经验。
  • 有数据湖/湖仓架构经验(如 Apache Iceberg、S3 或类似技术)。
  • 有分析型/列式数据库经验(如 ClickHouse 或类似技术)。
  • 有构建和编排 ETL/ELT 流程的经验(如 Temporal、Airflow 或类似工具)。
  • 精通 Python 和/或 Scala/Java 编程。
  • 有在微服务架构和云环境(AWS 优先)中工作的经验。
  • 自我驱动,具备强大的多任务处理能力,并有良好的团队合作经验。
  • 优秀的沟通能力,能够独立工作并协同合作。
  • 有使用 Apache Spark(或类似技术)进行大规模数据处理的实际经验。
  • 具备英语书面和口语的专业水平。
  • 注意:该职位专注于批处理数据(非实时流处理)。

Nice to Have

  • 有参与开源数据平台和工具的经验和贡献。
  • 熟悉 BI 和可视化工具(如 Superset、Looker、Tableau、Metabase 或类似工具)。
  • 有容器化和编排经验(Docker、Kubernetes 等)。
查看英文原文

Description
We are seeking an experienced Data Engineer. The ideal candidate is self-motivated, a multitasker, and a demonstrated team player. You will be responsible for designing, developing, managing, and maintaining our open-source data platform, including our Data-Lakehouse (S3, Apache Iceberg, and ClickHouse), ETL processes, and orchestration tool (Temporal Workflow).
What You Will Do

  • Develop a scalable data platform integrating multiple sources for easy access.
  • Design and enhance data tools (orchestration, governance, Data-Lakehouse, BI, etc.).
  • Ensure smooth operation of data systems for analysts, scientists, and engineers.
  • Optimize data pipelines (ingestion, processing, and output) in a microservices environment.
  • Build, maintain, and monitor ETL/ELT processes and orchestrate workflows using Temporal.
  • Troubleshoot and improve the performance, scalability, and reliability of the data infrastructure (S3, Apache Iceberg, ClickHouse).
  • Collaborate cross-functionally with data scientists, analysts, and backend engineers to understand data needs and deliver solutions.
  • Implement and champion data quality, governance, and security best practices across the platform.

Requirements

  • 3+ years of experience as a Data Engineer or in a similar data infrastructure role.
  • Strong proficiency in SQL and hands-on experience with data modeling.
  • Experience with data lake/lakehouse architectures (e.g., Apache Iceberg, S3, or similar).
  • Experience with analytical / columnar databases (e.g., ClickHouse or similar).
  • Experience building and orchestrating ETL/ELT pipelines (e.g., Temporal, Airflow, or similar).
  • Strong programming skills in Python and/or Scala/Java.
  • Experience working within a microservices architecture and cloud environments (AWS preferred).
  • Self-motivated, strong multitasking skills, and a demonstrated team player.
  • Excellent communication skills and the ability to work both independently and collaboratively.
  • Hands-on experience with Apache Spark (or similar technologies) for large-scale data processing.
  • Professional proficiency in written and spoken English.
  • Note: this role is focused on batch data processing (not real-time streaming).

Nice to Have

  • Experience working with and contributing to open-source data platforms and tools.
  • Familiarity with BI and visualization tools (e.g., Superset, Looker, Tableau, Metabase, or similar).
  • Experience with containerization and orchestration (Docker, Kubernetes).
  • Experience with infrastructure-as-code and CI/CD practices.
  • Experience with AWS EMR and running Apache Spark workloads in a cloud environment.
  • Experience leveraging AI-assisted development tools (e.g., GitHub Copilot, Cursor, or similar) to boost engineering productivity.

Originally posted on Himalayas

本页面信息整理自 Himalayas,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

该公司其他在招职位

← 返回全部职位