远程工作雷达

资深平台工程师,AI/ML基础设施

Staff Platform Engineer, AI/ML Infrastructure

AI开发工程限定地区(需当地身份)日间重叠约 2 小时,需偶尔早起或晚睡
公司Pfizer
薪资€65,250 - €108,750/年
工作地点France
地域资格限定地区(需当地身份)
时区要求日间重叠约 2 小时,需偶尔早起或晚睡
用工类型Full Time
发布时间昨天
数据来源Himalayas
前往 Himalayas 查看并投递 →
注意地域限制:该职位明确限定在 France 招聘。如果你是位于中国大陆的求职者,通常需要当地工作身份才能投递,或需与雇主确认是否接受独立合同(Contractor)形式合作。
作息提示:日间重叠约 2 小时,需偶尔早起或晚睡。

Staff Platform Engineer, AI/ML Infrastructure
部门:AI软件与运维
职位概述
Staff Platform Engineer, AI/ML Infrastructure 将为支撑企业级生成式AI应用的云平台、部署系统和运维基础架构提供技术领导力。
该职位将定义并演进在AWS、Kubernetes、无服务器和容器化环境中运行的AI/ML平台的基础设施架构。工程师将主导平台标准,包括可靠性、可扩展性、可观测性、CI/CD、安全性及开发者赋能,同时与软件工程、AI工程、安全和运维团队紧密合作。
理想的候选人兼具深入的云工程实践经验与资深级别的技术影响力。他们熟悉设计基础设施模式、编写基础设施即代码、优化交付流水线、指导工程师,并做出提升多个团队AI平台运营成熟度的架构决策。
主要职责
定义并推动支持生成式AI应用、LLM集成、模型路由和企业AI服务的AI/ML平台基础设施的技术策略。
使用AWS服务(如EKS、ECS Fargate、Lambda、DynamoDB、S3、OpenSearch、Secrets Manager、CloudWatch、ALB和MWAA)构建和操作可扩展的云平台。
利用CloudFormation、Helm和Terraform建立可重用的基础设施模式,以支持可靠的多环境和多区域部署。
使用GitHub Actions、可重用的工作流、基于OIDC的AWS认证、自动化质量门禁、部署晋升和环境审批来主导CI/CD架构。
设计并改进AI平台的可观测性,包括CloudWatch仪表板、日志、警报、Prometheus/Grafana、OpenSearch、Langfuse和针对LLM的特定操作指标。
构建适用于GenAI工作负载的平台能力,包括模型可用性监控。
与软件工程团队合作,提升部署可靠性、回滚策略、健康检查、自动扩展、负载测试和运行时性能。
定义并执行基础设施的安全与合规实践,包括IAM权限边界、Secrets Manager的使用、密钥扫描、审计日志、标签标准和变更管理控制。
为成本优化、容量规划、环境标准化和运维提供技术领导力。

查看英文原文

Staff Platform Engineer, AI/ML Infrastructure
Department:AI Software & Operations
Role Summary
The Staff Platform Engineer, AI/ML Infrastructure will provide technical leadership for thecloud platforms, deployment systems, and operational foundations that power enterprise-scalegenerative AI applications.
This role will define and evolve the infrastructure architecture for AI/ML platforms running across AWS,Kubernetes, serverless, and containerized environments. The engineer will lead platform standards forreliability, scalability, observability, CI/CD, security, and developer enablement, while partnering closelywith software engineering, AI engineering, security, and operations teams.
The ideal candidate combines deep hands-on cloud engineering experience with staff-level technicalinfluence. They are comfortable designing infrastructure patterns, writing infrastructure-as-code,improving delivery pipelines, mentoring engineers, and making architectural decisions that raise theoperational maturity of AI platforms across multiple teams.
Key Responsibilities
Define and drive the technical strategy for AI/ML platform infrastructure supporting generative AIapplications, LLM integrations, model routing, and enterprise AI services.
Architect, build, and operate scalable cloud platforms using AWS services such as EKS, ECSFargate, Lambda, DynamoDB, S3, OpenSearch, Secrets Manager, CloudWatch, ALB, and MWAA.
Establish reusable infrastructure patterns using CloudFormation, Helm, and Terraform to supportreliable multi-environment and multi-region deployments.
Lead CI/CD architecture using GitHub Actions, reusable workflows, OIDC-based AWSauthentication, automated quality gates, deployment promotion, and environment approvals.
Design and improve observability across AI platforms, including CloudWatch dashboards, logs,alarms, Prometheus/Grafana, OpenSearch, Langfuse, and LLM-specific operational metrics.
Build platform capabilities for GenAI workloads, including model availability monitoring.
Partner with software engineering teams to improve deployment reliability, rollback strategies,health checks, autoscaling, load testing, and runtime performance.
Define and enforce security and compliance practices for infrastructure, including IAM permissionboundaries, Secrets Manager usage, secret scanning, audit logging, tagging standards, andchange-management controls.
Provide technical leadership for cost optimization, capacity planning, environment standardization,and operational resilience across development, test, production, and sandbox environments.
Mentor engineers, review architecture and infrastructure designs, and influence platformengineering practices across teams.
Basic Qualifications
Bachelor’s degree in Computer Science, Engineering, Information Technology, or a relatedtechnical field, or equivalent practical experience.
7+ years of experience in DevOps, platform engineering, cloud infrastructure, site reliabilityengineering, or software engineering roles.
Strong hands-on experience with AWS/Azure/GCP infrastructure and services, including container,serverless, networking, storage, observability, and security services.
Experience designing and operating production systems on Kubernetes, ECS/Fargate, orcomparable container orchestration platforms.
Proficiency with infrastructure-as-code, especially CloudFormation, Terraform, Helm, or similartooling.
Strong CI/CD experience with GitHub Actions or similar platforms, including reusable workflows,automated testing, deployment gates, and cloud authentication.
Experience building and operating observability solutions using CloudWatch, Prometheus/Grafana,OpenSearch, or similar tools.
Strong understanding of cloud security practices, IAM, secrets management, least-privilegeaccess, audit logging, and compliance requirements.
Experience supporting distributed systems, microservices, APIs, asynchronous workloads, andmulti-environment deployments.
Demonstrated ability to lead technical design, mentor engineers, and influence engineeringpractices across teams.
Preferred Qualifications
Experience supporting AI/ML or generative AI platforms, including LLM gateways, model routing,prompt observability, token metering, or model failover.
Experience operating platforms in regulated enterprise environments, ideally healthcare,pharmaceutical, finance, or life sciences.
Experience with multi-account, multi-region AWS architectures and enterprise governancepatterns.
Experience with cost optimization, autoscaling strategies, capacity planning, and cloud budgetmonitoring.
Experience with load testing and performance validation using tools such as Locust or comparableframeworks.
Strong Python or scripting skills for platform automation, operational tooling, and CI/CD extensions.
Ability to communicate complex technical decisions clearly to engineering, security, operations,and leadership audiences.
Technical Environment
This role works across a modern AI platform ecosystem including: Cloud:
AWS EKS, ECS Fargate, Lambda, DynamoDB, S3, OpenSearch, CloudWatch, SecretsManager, ALB, VPC, IAM
Infrastructure-as-Code: CloudFormation, Helm, Terraform
CI/CD: GitHub Actions, reusable workflows, OIDC federation, environment approvals, automatedrelease promotion
AI/ML Platform: AWS Bedrock, Azure OpenAI, LiteLLM, Langfuse
Observability: CloudWatch dashboards and alarms, Prometheus, Grafana, OpenSearch, Langfuse,custom metrics
Security & Governance: IAM permission boundaries, secret scanning, audit logging, taggingcompliance, change-management automation
Engineering Practices: Docker, Python, pre-commit, automated testing, load testing, code qualitygates, monorepo service standards
Leadership Expectations
As a J090 Staff-level engineer, this role is expected to operate beyond individual delivery. The engineerwill identify systemic platform gaps, define technical direction, create reusable standards, and raiseengineering maturity across multiple teams.
Success in this role requires strong judgment, ownership, and communication. The engineer should beable to balance hands-on implementation with architectural leadership, guide teams through ambiguoustechnical decisions, and build platform capabilities that make AI product teams faster, safer, and morereliable.
Work location assignment : Remote
This role is posted in multiple locations. If you are applying for the role in an secondary job posting location where pay transparency regulations apply, your Talent Advisor will share the local pay information with you during the first interview.Pfizer is an equal opportunity employer and complies with all applicable equal employment opportunity legislation in each jurisdiction in which it operates.
Égalité des chances & Emploi
Nous croyons que des équipes diversifiées et inclusives sont essentielles à la réussite d'une entreprise. En tant qu'employeur, Pfizer s'engage à valoriser la diversité et l’inclusion sous toutes ses formes. Cette diversité se reflète également à travers les patients et les communautés que nous servons. Ensemble, continuons à bâtir une culture qui encourage, soutient et responsabilise nos employés.
Handicap & Inclusion
Notre mission est de libérer le potentiel de nos collaborateurs et nous sommes fiers d'être un employeur inclusif pour les personnes handicapées, garantissant ainsi l'égalité des chances en matière d'emploi pour tous les candidats. Nous vous encourageons à donner le meilleur de vous-même en sachant que nous apporterons tous les ajustements raisonnables pour soutenir votre candidature et votre carrière future. Votre expérience avec Pfizer commence ici !
Pfizer endeavors to make www.pfizer.com/careers accessible to all users. If you would like to contact us regarding the accessibility of our website or need assistance completing the application process and/or interviewing, please email . This is to be used solely for accommodation requests with respect to the accessibility of our website, online application process and/or interviewing. Requests for any other reason will not be returned.Pour mieux comprendre les usages autorisés et interdits de l’intelligence artificielle tout au long du processus de recrutement, nous vous invitons à consulter nos bonnes pratiques dédiées à l’utilisation de l’IA par les candidats sur Pfizer Careers.
Information & Business TechOriginally posted on Himalayas

本页面信息整理自 Himalayas,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

该公司其他在招职位

临床研究助理

PfizerPoland100,500/年 PLNFull Time昨天
市场运营职能支持限定地区(需当地身份)日间重叠约 2 小时,需偶尔早起或晚睡

← 返回全部职位