数据工程师负责人 - Databricks - Eng
Data Engineer Lead - Databricks - Eng
🚀 加入我们的远程数据产品和机器学习开发初创公司! 🚀
Muttdata 是一家致力于使用前沿大数据和机器学习技术打造创新系统的动态初创公司。
我们正在寻找一位数据工程师负责人,帮助我们将专业知识提升到新的高度。如果你也像我们一样是数据爱好者,我们很乐意与你联系! 🐶🚀
🚀 我们做什么
- 借助我们的专业技能,我们为需求预测和预算预测构建现代机器学习系统。
- 开发可扩展的数据基础设施,根据每个客户的需求提升高层次决策能力。
- 提供全面的数据工程和定制AI解决方案,优化云基础系统。
- 使用生成式AI,我们帮助电商平台和零售商更快地创建更高质量的广告。
- 构建深度学习模型,我们提升各个行业的视觉识别和自动化能力,改善产品分类、质量控制和信息检索。
- 开发推荐模型,我们在电商、流媒体和数字平台上个性化用户体验,提高参与度和转化率。
🌟 我们的合作伙伴
- 亚马逊网络服务
- Astronomer
- Databricks
🌟 我们的价值观
- 📊 我们是数据爱好者
- 🤗 我们是开放协作的团队成员
- 🚀 我们有主人翁精神
- 🌟 我们保持积极心态
🔍 想了解我们正在做什么吗?查看我们的案例研究,并深入阅读我们的博客文章,了解更多关于我们的文化以及我们正在进行的令人兴奋的项目! 🚀
职责 🤓
- 领导端到端的数据迁移项目,确保平稳、安全且符合业务需求的过渡。
- 定义并传达数据架构决策,评估在AWS Redshift、Databricks等云数据平台及相关技术中的影响、风险和性能考量。
- 协调发现和记录依赖关系、模式目录、数据流程以及迁移工作中的关键流程。
- 负责迁移和切换计划,确保所有活动、验证和应急措施都得到明确定义和执行。
- 作为工程、分析、产品和业务利益相关者之间的主要联络人,促进整个项目生命周期中的清晰沟通和对齐。
- 将数据管道、数据仓库和平台架构中的技术变更转化为清晰的业务影响评估。
- 定义并记录假设、需求和接受标准。
查看英文原文
🚀 Join Our Data Products and Machine Learning Development Remote Startup! 🚀
Muttdata is a dynamic startup committed to crafting innovative systems using cutting-edge Big Data and Machine Learning technologies.
We’re looking for a Data Engineer Lead to help take our expertise to the next level. If you consider yourself a data nerd like us, we’d love to connect! 🐶🚀
🚀 What We Do
- Leveraging our expertise, we build modern Machine Learning systems for demand planning and budget forecasting.
- Developing scalable data infrastructures, we enhance high-level decision-making, tailored to each client.
- Offering comprehensive Data Engineering and custom AI solutions, we optimize cloud-based systems.
- Using Generative AI, we help e-commerce platforms and retailers create higher-quality ads, faster.
- Building deep learning models, we enhance visual recognition and automation for various industries, improving product categorization, quality control, and information retrieval.
- Developing recommendation models, we personalize user experiences in e-commerce, streaming, and digital platforms, driving engagement and conversions.
🌟 Our Partnerships
- Amazon Web Services
- Astronomer
- Databricks
🌟 Our Values
- 📊 We are Data Nerds
- 🤗 We are Open Team Players
- 🚀 We Take Ownership
- 🌟 We Have a Positive Mindset
🔍 Curious about what we’re up to? Check out our case studies and dive into our blog post to learn more about our culture and the exciting projects we’re working on! 🚀
Responsibilities 🤓
- Lead end-to-end data migration initiatives, ensuring a smooth, secure, and business-aligned transition.
- Define and communicate data architecture decisions, evaluating impacts, risks, and performance considerations across cloud data platforms such as AWS Redshift, Databricks, and related technologies.
- Coordinate the discovery and documentation of dependencies, schema inventories, data flows, and critical processes involved in migration efforts.
- Own migration and cutover planning, ensuring all activities, validations, and contingency measures are properly defined and executed.
- Serve as the primary liaison between engineering, analytics, product, and business stakeholders, facilitating clear communication and alignment throughout the project lifecycle.
- Translate technical changes in data pipelines, data warehouses, and platform architecture into clear business impact assessments.
- Define and document assumptions, requirements, and acceptance criteria to ensure successful migration outcomes.
- Oversee data validation processes to guarantee the accuracy, consistency, and integrity of migrated datasets against source-of-truth systems.
- Proactively identify, assess, and mitigate migration-related risks, escalating blockers and dependencies when necessary to maintain project timelines and quality standards.
- Drive problem-solving and decision-making in fast-paced, ambiguous environments, helping cross-functional teams stay focused on priorities and deliverables.
- Create, maintain, and improve project documentation, including migration runbooks, testing strategies, rollback plans, and operational procedures.
- Provide regular project status updates to technical and executive stakeholders, ensuring transparency around progress, risks, and next steps.
- Promote data governance, quality, and operational best practices throughout the migration lifecycle.
- Align engineers, analysts, and product stakeholders around priorities, timelines, and delivery expectations to ensure successful project execution.
Required Skills 💻
- Experience in Data Engineering, including building and optimizing data pipelines.
- Experience working with Databricks, Python, and Spark, building batch and streaming.
- Advanced English Level.
- Familiarity with a BI visualization tool like Looker or Tableau.
🎁 Perks
- Remote-first culture – work from anywhere! 🌍
- AWS, DBT, Google Cloud, Azure & Databricks certifications fully covered
- In-Company English Lessons.
- Birthday off + an extra vacation week (Mutt Week! 🏖️)
- Referral bonuses – help us grow the team & get rewarded!
- Maslow: Monthly credits to spend in our benefits marketplace.
- ✈️🏝️ Annual Mutters' Trip – an unforgettable getaway with the team.
- 👶 Monthly Childcare Reimbursement – Because supporting families matters too