数据平台工程师
Data Platform Engineer
我们致力于帮助医疗机构构建高效、专业的临床团队,提供行业领先的科技产品、人才与薪酬分析以及自动化工作流程解决方案。依托深厚的技术实力和行业领先的数据资源,我们为客户提供创新的解决方案,助力他们在临床团队生命周期的每个阶段进行规划、教育和参与。我们致力于为客户提供卓越的指导和支持,使其专注于塑造医疗行业的未来。
作为数据平台工程师,你的核心职责是设计、推进并维护管理数据作为企业资产的关键基础设施和系统。在此职位中,你将创建和维护工具、抽象层和服务,使企业用户能够细致地管理其数据产品的整个生命周期,包括数据的高效收集、存储、处理和检索,同时确保严格的安全性、治理和数据质量标准。你需要具备在多种数据存储技术(包括数据库、数据湖和数据仓库)以及流式架构和服务层技术方面丰富的经验。你精通Python、Scala、Java和SQL等编程语言。你的技能涵盖数据集成、ETL(抽取、转换、加载)、ELT(抽取、加载、转换)、Restful/GRPC服务、流式处理,以及构建和抽象强大的数据管道。作为数据平台工程师,总体目标是使企业产品团队能够交付高质量的数据产品,为客户带来业务洞察和价值。这包括构建连接数据源和数据消费者的工具和服务,确保数据的可访问性、结构完整性和可靠性,以满足分析和业务需求。
主要职责
- 与数据平台架构师合作,设计和定义平台工具、服务和集成方案
- 构建数据湖和湖仓架构
- 构建流式数据架构,以支持数据平台工具和服务
- 与软件团队合作,将流式和批处理架构与软件产品集成
- 与数据科学家、分析师和业务相关方协作,实现最优的最小可行产品解决方案
- 协作
查看英文原文
ABOUT US AND ABOUT YOU
Clinician Nexus enables health care organizations to build thriving clinician teams with industry-leading technology products, workforce and compensation analytics, and automated workflow solutions. Backed by extensive technical expertise and industry-leading data, we deliver innovative approaches to help clients to plan, educate, and engage their clinical workforce at every stage of the lifecycle. We are committed to providing our clients with outstanding guidance and support as they focus on shaping the future of health care.
As a Data Platform Engineer your core responsibility revolves around crafting, advancing, and maintaining the infrastructure and systems essential for managing and optimizing data as an enterprise asset. In this capacity, you will create and maintain tooling, abstractions, and services that allow enterprise users to meticulously oversee the entire lifecycle of their data product, encompassing the efficient collection, storage, processing, and retrieval of data, all while upholding stringent standards for security, governance, and data quality. You will possess extensive expertise in working with diverse data storage technologies, including databases, data lakes, and data warehouses, as well as streaming architectures and service layer technologies. You are proficient in programming languages such as Python, Scala, Java, and SQL. Your skillset extends to encompass data integration, ETL (Extract, Transform, Load) processes, ELT (Extract, Load, Transform), Restful/GRPC services, Streaming, and the construction and abstraction of robust data pipelines. The overarching objective as a Data Platform Engineer is to enable enterprise product teams to deliver quality data products that drive business insight and value to our customers. This includes building tooling and services that connects data sources to data consumers and ensures the accessibility, structural integrity, and reliability of data to fulfill analytical and business imperatives.
PRIMARY ACCOUNTABILITIES
- Collaborate with data platform architects to design and specify platform tooling, services, and integrations
- Build data lake and lake house architectures
- Build streaming data architectures to support data platform tooling and services
- Collaborate with software teams to integrate streaming and batch architectures with software products
- Collaborate with data scientists, analysts, and business stakeholders to MVP optimal solutions
- Collaborate with data governance teams and data platform architects to build governance into technical solutions
- Maintain systems and services that provide transparency and observability into our critical systems
- Implement and promote engineering and architectural patterns, perform code reviews, and collaborate in architectural reviews
- Collaborate with product team data engineers and architects to MVP data products, maintain data models, and promote data modeling best practice
- Provide technical leadership and mentorship to product and data engineers, guiding their growth and professional development and enabling their ability to use platform tooling and services
- Lead by example through hands-on contributions to designing, coding, and troubleshooting complex data systems
- Identify and address performance bottlenecks and optimization opportunities within data pipelines, databases, and processing frameworks
- Optimize data processing workflows to improve efficiency and reduce latency
- Lead efforts to diagnose and resolve data-related incidents in a timely manner
- Will participate in the grooming of stories
- Will be responsible for their own tasking towards the completion of stories
- Will mentor junior members of the team and guide junior members on best practice
- Will participate in design
- Will be responsible for the quality of their own code and will participate in code review of others product
- Will be responsible for integrating their own code with the team's DevOps plan and implementation
KNOWLEDGE, SKILLS & ABILITIES
Minimum Required Qualifications
- 5+ years relevant experience
- Degree in Computer Science, Software Engineering, Information Systems, Information Technology or a related computer degree or equivalent experience. Master’s degree is a plus
- Must know and have proficiency in one object or object/functional programing language. Preferably Python
- Must know common object and object/functional design patterns. Builder, factory, façade, context, etc.
- Must have proficiency with Apache Spark
- Must know data lake and lake house design principles, OLTP (Online Transaction Processing) design principles, document data stores, and graph
- Must know and understand how to build ELT/ETL patterns in a distributed compute system. Preferably Databricks
- Must know and have proficiency with common data quality tooling and the design of configurable systems to front end that tooling
- Must know and have proficiency with common systems and data observability tooling
- Must know standard practices of the Software Development Lifecycle (SDLC)
- Must have proficiency in standard SDLC concepts and tooling including unit and integration testing frameworks like Pytest, IDE features like debugging, testing practices like mocking, CICD tooling like Github actions, build tooling (poetry), GIT
- Must have proficiency in creating and using basic Restful or GRPC services
- Must have proficiency in message bus and streaming technologies and understand common event stream architectures
- Knowledge of the health care industry is a plus
WORK ENVIRONMENT
Work location is remote. Minimal travel. Must be physically able to perform the essential functions of the job.
SALARY, BENEFITS & PERKS
- Competitive total compensation package
- Medical and dental coverage at no premium cost for employees
- 401(k) and profit-sharing retirement plans
- Flexible spending accounts
- Generous paid time off (PTO)
- Company paid holidays
- Gender-neutral parental leave
- Bereavement and pet leave
- Continuing education and professional accreditation sponsorship
- Life and AD&D insurance
- Short- and long-term disability
- Employee assistance program
- Mental health support program
- Additional perks
Reflected below is the base salary range offered for this position. Actual salaries may vary depending on factors including but not limited to academic achievements, skills and experience. The range listed is just one component of the compensation package offered to candidates.
- $120,000.00 - $160,000.00 annually
Our Values In Action: How We CARE
We live our values daily through four commitments:
- Connect: Collaborate selflessly to support others, advance ideas, and solve problems using critical thinking.
- Act: Bring integrity and respect to every interaction—no exceptions.
- Reach: Commit to continuous learning and knowledge sharing that strengthens teams and clients.
- Embrace: Foster inclusion and belonging so everyone can thrive and contribute.
SullivanCotter Holdings (Clinician Nexus, SCH Services, SullivanCotter) is an Equal Employment Opportunity/Affirmative Action employer, and all qualified applicants will receive consideration for employment without regard to race, color, religion, sex, age, national origin, protected veteran status, disability status, sexual orientation, gender identity or expression, marital status, genetic information, or any other characteristic protected by law or marital status.
Originally posted on Himalayas