远程工作雷达

数据工程师 (Spark)

Data Engineer (Spark)

开发工程限定地区(据职位描述推断)
公司Addepto
薪资未公开
工作地点Worldwide
地域资格限定地区(据职位描述推断)
时区要求日间重叠约 9 小时,基本正常作息
用工类型Full Time
发布时间今天
数据来源Himalayas
前往 Himalayas 查看并投递 →
注意地域限制:该职位明确限定在 招聘。如果你是位于中国大陆的求职者,通常需要当地工作身份才能投递,或需与雇主确认是否接受独立合同(Contractor)形式合作。

Addepto 是一家领先的 AI 咨询()和数据工程()公司,为全球最大的企业及创新初创公司打造可扩展、注重投资回报率的 AI 解决方案,包括 Rolls Royce、Continental、Porsche、ABB 和 WGU。专注于人工智能和大数据,Addepto 通过可衡量业务影响和长期增长的系统,帮助组织释放数据的全部潜力。
公司的业务不仅限于客户项目。基于现实世界中的挑战和洞察,Addepto 自主开发了产品 ContextClue,并积极为 AI 社区贡献开源解决方案。这种将实践经验转化为可扩展创新的承诺,使 Addepto 获得了福布斯的认可,成为全球十大 AI 咨询公司之一。
作为 KMS Technology(一家美国-based 全球科技集团)的一部分,Addepto 将深度 AI 专业知识与企业级交付能力相结合,使合作伙伴能够安全、高效地将客户从 AI 实验推进到生产阶段。
作为数据工程师,你将有机会与技术专家团队一起,在多个行业的挑战性项目中工作,利用前沿技术。以下是我们在寻找有才华的人才加入的一些项目:

  • 开发和维护一个用于处理汽车数据的大型平台。大量数据以流式和批处理模式进行处理。技术栈包括 Spark、Cloudera、Airflow、Iceberg、Python 和 AWS。
  • 为全球航空航天公司设计和开发一个通用数据平台。该基于 Azure 和 Databricks 的项目整合了多种企业及公共数据源。该数据平台尚处于开发初期,涵盖架构和流程设计,同时提供技术选择的自由度。
  • 为一家快速增长的美国电信公司构建集中式报告平台。该项目涉及在 BigQuery 和 Looker 上实现中央数据报告平台,重点在于集中数据、整合各类 CRM,并构建高管报告解决方案以支持决策和业务增长。

🚀 主要职责:

  • 开发和维护高性能的汽车数据处理平台,确保其可扩展性和可靠性。
  • 设计和实现数据管道,处理
查看英文原文

Addepto is a leading AI consulting () and data engineering () company that builds scalable, ROI-focused AI solutions for some of the world's largest enterprises and pioneering startups, including Rolls Royce, Continental, Porsche, ABB, and WGU. With an exclusive focus on Artificial Intelligence and Big Data, Addepto helps organizations unlock the full potential of their data through systems designed for measurable business impact and long-term growth.
The company's work extends beyond client engagements. Drawing from real-world challenges and insights, Addepto has developed its own product - ContextClue - and actively contributes open-source solutions to the AI community. This commitment to transforming practical experience into scalable innovation has earned Addepto recognition by Forbes as one of the top 10 AI consulting companies worldwide.
As part of KMS Technology, a US-based global technology group, Addepto combines deep AI specialization with enterprise-scale delivery capabilities—enabling the partnership to move clients from AI experimentation to production impact, securely and at scale.
As a Data Engineer, you will have the exciting opportunity to work with a team of technology experts on challenging projects across various industries, leveraging cutting-edge technologies. Here are some of the projects we are seeking talented individuals to join:

  • Development and maintenance of a large platform for processing automotive data. A significant amount of data is processed in both streaming and batch modes. The technology stack includes Spark, Cloudera, Airflow, Iceberg, Python, and AWS.
  • Design and development of a universal data platform for global aerospace companies. This Azure and Databricks powered initiative combines diverse enterprise and public data sources. The data platform is at the early stages of the development, covering design of architecture and processes as well as giving freedom for technology selection.
  • Centralized reporting platform for a growing US telecommunications company. This project involves implementing BigQuery and Looker as the central platform for data reporting. It focuses on centralizing data, integrating various CRMs, and building executive reporting solutions to support decision-making and business growth.

🚀 Your main responsibilities:

  • Develop and maintain a high-performance data processing platform for automotive data, ensuring scalability and reliability.
  • Design and implement data pipelines that process large volumes of data in both streaming and batch modes.
  • Optimize data workflows to ensure efficient data ingestion, processing, and storage using technologies such as Spark, Cloudera, and Airflow.
  • Work with data lake technologies (e.g., Iceberg) to manage structured and unstructured data efficiently.
  • Collaborate with cross-functional teams to understand data requirements and ensure seamless integration of data sources.
  • Monitor and troubleshoot the platform, ensuring high availability, performance, and accuracy of data processing.
  • Leverage cloud services (AWS) for infrastructure management and scaling of processing workloads.
  • Write and maintain high-quality Python (or Java/Scala) code for data processing tasks and automation.

Requirements
🎯 What you'll need to succeed in this role:

  • At least 4 years of commercial experience implementing, developing, or maintaining Big Data systems, data governance and data management processes.
  • Strong programming skills in Python (or Java/Scala): writing a clean code, OOP design.
  • Hands-on with Big Data technologies like Spark, Cloudera, Kafka, Data Platform, Airflow, NiFi, Docker, and Iceberg.
  • Excellent understanding of dimensional data and data modeling techniques.
  • Experience implementing and deploying solutions in cloud environments.
  • Consulting experience with excellent communication and client management skills, including prior experience directly interacting with clients as a consultant.
  • Ability to work independently and take ownership of project deliverables.
  • Fluent English (at least C1 level).
  • Bachelor’s degree in technical or mathematical studies.

➕ Nice to have:

  • Experience with an MLOps framework such as Kubeflow or MLFlow.
  • Familiarity with Databricks and/or dbt.

🎁 Discover our perks & benefits:

  • Work in a supportive team of passionate enthusiasts of AI & Big Data.
  • Engage with top-tier global enterprises and cutting-edge startups on international projects.
  • Enjoy flexible work arrangements, allowing you to work remotely or from modern offices and coworking spaces.
  • Accelerate your professional growth through career development paths, knowledge-sharing initiatives, language classes, and sponsored training and conferences. Benefit from partnerships with Databricks and Anthropic, which provide access to industry-leading training materials and certification programs.
  • Participate in team-building events and utilize the integration budget.
  • Celebrate work anniversaries, birthdays, and milestones.
  • Access medical and sports packages, eye care, and well-being support services, including psychotherapy and coaching.
  • Get full work equipment for optimal productivity, including a laptop and other necessary devices.
  • With our backing, you can boostyourpersonal brand by speaking at conferences, writing for our blog, or participating in meetups.
  • Experience a smooth onboarding with a dedicated buddy, and start your journey in our friendly, supportive, and autonomous culture.

Originally posted on Himalayas

本页面信息整理自 Himalayas,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

该公司其他在招职位

← 返回全部职位