远程工作雷达

高级软件工程师,SRE/可观测性工具

Senior Software Engineer, SRE / Observability Tooling

开发工程未标注地域
公司Glia
薪资未公开
工作地点Estonia
地域资格未标注地域
时区要求日间重叠约 3 小时,需偶尔早起或晚睡
用工类型permanent
发布时间2026-07-17
数据来源4dayweek.io
前往 4dayweek.io 查看并投递 →

**关于 Glia**

Glia 是排名第一的银行人工智能平台,帮助社区和区域金融机构提高效率、加速贷款增长、推动存款,并提供能够战胜大银行和金融科技公司的体验。

Glia 的银行人工智能操作系统是建立在现有技术栈之上的核心智能层,激活了一个由专业代理组成的 AI 工作团队,这些代理从银行数据、交互历史和系统记录中获取信息。这些经过银行训练的代理在语音和数字渠道(从前台到后台)自动化工作流程,从而降低运营成本并实现通用银行员模式。

Glia 以其坚不可摧的安全性和可靠性受到 700 多家银行和信用合作社的信任,Glia 提供了行业首个合同保证的无幻觉承诺。这就是为什么 Glia 客户能够迅速且自信地将银行 AI 投入使用,并从第一天起就获得可衡量的结果。有关 Glia 的更多信息,请访问 [glia.com](http://glia.com)。

**团队介绍**

你将加入我们专注的 **可观测性团队**,该团队构建用于监控 Glia 云原生核心基础设施的可观测性平台和 SRE 工具。我们的团队专注于赋能和自动化:我们为工程团队提供标准、工具和平台,使他们能够保持自身系统的可用性和最佳性能。

**使命与工作内容**

作为该团队的站点可靠性工程师,你的重点将是构建 SRE 和可观测性工具——其他团队用来保持服务健康的平台、自动化和标准。这是一个工具和赋能角色,而不是生产运维角色:你不会直接操作 Glia 的生产服务。职责包括:

- 开发用于仪表板、警报和监控的标准化、基础设施和自动化工具。

- 与开发团队合作,建立生产就绪和运营就绪。

- 构建团队用于定义、测量和报告其服务的服务级别目标(SLO)和服务级别指标(SLI)的工具和模板。

- 开发工具以自动化可观测性和运营工作流程,减少工程团队的重复劳动。

- 构建和改进事件响应工具和工作流程,帮助团队更快地解决故障并从中学习。

**我们的协作模式**

Glia 工程团队采用远程优先的工作方式,成员分布在加拿大、葡萄牙、波兰

查看英文原文

**About Glia**

Glia is the #1 Banking AI platform, empowering community and regional financial institutions to create efficiencies, accelerate loan growth, drive deposits, and deliver experiences that win against megabanks and fintechs.

Glia's Banking AI Operating System is a central intelligence layer on top of existing tech stacks, activating an AI workforce of specialized agents that draw from banking data, interaction history, and integrated systems of record. These banking-trained agents automate workflows across voice and digital–from front office to back office–resulting in decreased operational costs and the Universal Banker model.

Trusted by 700+ banks and credit unions for its ironclad security and reliability, Glia delivers the industry’s first contractual no-hallucination guarantee. It’s why Glia customers quickly and confidently put Banking AI to work with measurable results from day one. More information about Glia can be found at [glia.com](http://glia.com).

**The Team**

You'll be joining our dedicated **Observability Team**, which builds the observability platform and SRE tooling used to monitor Glia’s cloud-native core infrastructure serving the conversational AI. Our team focuses on enablement and automation: we provide the standards, tooling, and platform that engineering teams use to keep their own systems available and performing optimally.

**The Mission & The Work**

As a Site Reliability Engineer on this team, your focus will be on building SRE and observability tooling — the platform, automation, and standards other teams use to keep their services healthy. This is a tooling and enablement role, not a production operations role: you will not be directly operating Glia’s production services. Responsibilities will include:

- Developing standards, infrastructure and automation for dashboards, alerts, and monitors as code.

- Partnering with development teams to establish production readiness and operational readiness.

- Building the tooling and templates teams use to define, measure, and report on Service Level Objectives (SLOs) and Service Level Indicators (SLIs) for their services.

- Developing tooling to automate observability and operational workflows, eliminating manual toil for engineering teams.

- Building and improving the incident response tooling and workflows that help teams resolve outages faster and learn from them.

**Our Collaboration Model**

Glia Engineering is remote-first, spanning Canada, Portugal, Poland, and Estonia. The Observability team is based in Estonia, with optional offices in Tallinn and Tartu. We thrive on flexible remote collaboration, but we still bring the whole team together in Estonia twice a year for in-person innovation and connection.

**Our tech stack**

- **Infrastructure:** AWS, Kubernetes (AWS EKS), Istio, EFK

- **Persistence:** Amazon Aurora Serverless for Postgres, RabbitMQ, Amazon RDS

- **Cache:** Amazon ElastiCache

- **Monitoring & Observability:** DataDog with a focus on dashboards and alerts for system health

- **CI/CD:** Github Actions, ArgoCD, Jenkins, Helm, with a focus on automation and pipeline optimization.

- **Infrastructure as Code:** Terraform

- Additionally, our Engineering teams use:

- **Backend:** Python, Elixir, [Node.js](http://node.js), Ruby, Go

- **Frontend:** Javascript and React.js

- **Native mobile SDKs:** Java and Swift

**What We’re Looking for**

- Expert-level proficiency with AWS and Kubernetes (EKS), particularly in areas of observability, networking, and auto-scaling.

- Experience with modern observability platforms (e.g., DataDog, Prometheus) and a deep understanding of metrics, logging, and tracing.

- Deep, practical understanding of Site Reliability Engineering (SRE) principles (SLOs, error budgets, toil reduction).

- Demonstrable experience analyzing and troubleshooting large-scale distributed systems.

- Strong software development skills in a language like Python or Go, used to build operational tools, services, or automation.

- Expertise in designing and operating robust CI/CD pipelines for a microservices architecture (e.g., using ArgoCD, Github Actions, Helm).

- A systematic, data-driven approach to problem-solving and root cause analysis.

- Proficiency in using AI tools thoughtfully, maintaining ownership of the final output while recognizing the tools' limitations.

Glia is an equal-opportunity employer. Glia does not discriminate against any employee or applicant because of race, creed, color, religion, gender, sexual orientation, gender identity/expression, national origin, disability, age, genetic information, veteran status, marital status, pregnancy or related condition (including breastfeeding), or any other basis protected by law.

_The Glia Talent Acquisition team uses **@**_ [_**glia.com**_](http://glia.com) _and_ [_**@**_](mailto:darina.danchenko@gliatalent.com) [_**gliatalent.com**_](http://gliatalent.com) _email addresses for coordinating interviews, providing updates, and sending documents._

_Our hiring process involves an introduction, practical and team interviews, and a decision and offer. For more information, visit our_ [_Recruitment Privacy Notice page_](https://www.glia.com/eu-recruitment-privacy-notice) _or contact our talent team via **talent@glia.com**_

本页面信息整理自 4dayweek.io,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

该公司其他在招职位

技术客户经理

GliaMexicopermanent7 天前
职能支持限定地区(需当地身份)与中国几乎无重叠,需长期倒时差

生态合作伙伴经理

GliaUnited Statespermanent27 天前
市场运营限定地区(需当地身份)与中国几乎无重叠,需长期倒时差

解决方案架构师

GliaMexicopermanent2026-08-13
开发工程限定地区(需当地身份)与中国几乎无重叠,需长期倒时差

现场营销经理

GliaUnited Statespermanent2026-08-04
市场运营限定地区(需当地身份)与中国几乎无重叠,需长期倒时差

助理销售工程师

GliaUnited Statespermanent2026-08-03
开发工程市场运营限定地区(需当地身份)与中国几乎无重叠,需长期倒时差

← 返回全部职位