远程工作雷达

系统工程师资深

Principal Systems Engineer

AI开发工程职能支持限定地区(需当地身份)日间重叠仅 1 小时,需熬夜配合
公司Harris
薪资未公开
工作地点United Kingdom
地域资格限定地区(需当地身份)
时区要求日间重叠仅 1 小时,需熬夜配合
用工类型Full Time
发布时间今天
数据来源Himalayas
前往 Himalayas 查看并投递 →
注意地域限制:该职位明确限定在 United Kingdom 招聘。如果你是位于中国大陆的求职者,通常需要当地工作身份才能投递,或需与雇主确认是否接受独立合同(Contractor)形式合作。
作息提示:日间重叠仅 1 小时,需熬夜配合。

Site Reliability Engineer (SRE) - 远程
概述
作为Altera的Site Reliability Engineer(SRE),您将负责确保我们托管的医疗保健平台的可靠性、可扩展性和性能。该职位结合软件和系统工程,以提高服务可用性,自动化运维,并提升客户体验。您将在我们的云和混合环境中担任监控、故障排除、事件响应和持续改进的技术领导者。
主要职责

  • 维护并提升生产环境的可靠性、可用性和性能。
  • 领导复杂应用程序、数据库和基础设施问题的调查与解决。
  • 参与事件管理,进行根本原因分析(RCA),并为事件后评审做出贡献,以防止未来再次发生。
  • 定义并衡量服务级别指标(SLIs)和服务目标(SLOs),以满足我们的服务承诺。
  • 制定主动监控和警报策略,以在问题影响客户之前识别并解决。
  • 使用脚本和基础设施即代码(IaC)自动化运维任务,以提高效率。
  • 与工程和云团队合作,优化部署、监控和支持流程。
  • 在重大事件中提供技术领导,并作为关键问题的升级联系人。

资格要求
经验:

  • 7年以上支持企业应用、基础设施或云环境的经验。
  • 监控与可观测性:熟悉APM工具,如LogicMonitor、AppDynamics、Azure Monitor、SentryOne、Dynatrace、Datadog或New Relic。
  • Microsoft堆栈:深入理解Windows Server管理、IIS、.NET应用、Windows群集、MSMQ、事件日志和PerfMon。
  • 数据库技能:熟练使用SQL Server,包括性能调优、查询优化、阻塞分析和Always On可用性组。
  • 云与网络:有Azure云环境经验,并具备扎实的网络基础知识(DNS、TCP/IP、负载均衡、防火墙)。
  • ITSM与ITIL:熟悉ServiceNow(或其他ITSM平台)和ITIL原则。

优先技能:

  • 使用PowerShell、Python或其他类似语言进行脚本编写。
  • 基础设施即代码(Terraform、ARM模板、Bicep)。
  • CI/CD流水线和部署自动化(Azure DevOps、GitHub Actions)。
  • 有相关经验
查看英文原文

Site Reliability Engineer (SRE) - Remote
Overview
As a Site Reliability Engineer (SRE) at Altera, you will be responsible for ensuring the reliability, scalability, and performance of our hosted healthcare platforms. This role blends software and systems engineering to enhance service availability, automate operations, and improve the customer experience. You will act as a technical leader in monitoring, troubleshooting, incident response, and continuous improvement across our cloud and hybrid environments.
Key Responsibilities

  • Maintain and improve the reliability, availability, and performance of our production environments.
  • Lead the investigation and resolution of complex application, database, and infrastructure issues.
  • Participate in incident management, conduct root cause analysis (RCA), and contribute to post-incident reviews to prevent future occurrences.
  • Define and measure Service Level Indicators (SLIs) and Objectives (SLOs) to meet our service commitments.
  • Develop proactive monitoring and alerting strategies to identify and resolve issues before they impact customers.
  • Automate operational tasks using scripting and Infrastructure-as-Code (IaC) to improve efficiency.
  • Partner with engineering and cloud teams to refine deployment, monitoring, and support processes.
  • Provide technical leadership during major incidents and act as a key escalation point for critical issues.

Qualifications
Experience:

  • 7+ years of experience supporting enterprise applications, infrastructure, or cloud environments.
  • Monitoring & Observability: Strong experience with APM tools such as LogicMonitor, AppDynamics, Azure Monitor, SentryOne, Dynatrace, Datadog, or New Relic.
  • Microsoft Stack: Deep knowledge of Windows Server administration, IIS, .NET applications, Windows Clustering, MSMQ, Event Logs, and PerfMon.
  • Database Skills: Strong SQL Server experience, including performance tuning, query optimization, blocking analysis, and Always On Availability Groups.
  • Cloud & Networking: Experience with Azure cloud environments and a solid understanding of networking fundamentals (DNS, TCP/IP, load balancing, firewalls).
  • ITSM & ITIL: Familiarity with ServiceNow (or other ITSM platforms) and ITIL principles.

Preferred Skills:

  • Scripting with PowerShell, Python, or similar languages.
  • Infrastructure as Code (Terraform, ARM Templates, Bicep).
  • CI/CD pipelines and deployment automation (Azure DevOps, GitHub Actions).
  • Experience with Kubernetes and containerized workloads.
  • Experience implementing SLOs, SLIs, and Error Budgets.
  • Experience in a healthcare technology or patient care environment.

Education:
· Bachelor's Degree in Computer Science, Information Technology, or Engineering is preferred; equivalent professional experience will be considered.
Working Arrangements

  • This is a remote position open to candidates within the UK
  • You will participate in an on-call rotation to support our 24x7 healthcare environment.
  • Occasional after-hours work is required for activations, upgrades, and major incidents.

Travel
· Travel is not a requirement for this role.

Why Altera?
At Altera Digital Health, you will have the opportunity to profoundly impact the lives of patients by empowering healthcare providers to deliver superior care. You will join a passionate and gifted team committed to innovation and excellence. We offer a competitive compensation and benefits package and the opportunity to work in a fast-paced and dynamic environment.
Originally posted on Himalayas

本页面信息整理自 Himalayas,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

该公司其他在招职位

高级项目经理

HarrisCanadaFull Time今天
开发工程限定地区(需当地身份)

项目经理

HarrisCanada90,000 - 110,000/年 CADFull Time今天
职能支持限定地区(需当地身份)

← 返回全部职位