远程工作雷达

高级/资深DevOps工程师

Senior/Staff DevOps Engineer

开发工程限定地区(需当地身份)日间重叠仅 1 小时,需熬夜配合
公司MEDvidi
薪资未公开
工作地点Portugal
地域资格限定地区(需当地身份)
时区要求日间重叠仅 1 小时,需熬夜配合
用工类型Full Time
发布时间今天
数据来源Himalayas
前往 Himalayas 查看并投递 →
注意地域限制:该职位明确限定在 Portugal 招聘。如果你是位于中国大陆的求职者,通常需要当地工作身份才能投递,或需与雇主确认是否接受独立合同(Contractor)形式合作。
作息提示:日间重叠仅 1 小时,需熬夜配合。

MEDvidi 是一个基于人工智能的心理健康护理平台,正在为美国的医疗保健设定新的安全、有效和可扩展的精神病学护理标准。
我们结合持牌医护人员与专有 AI 工具,为 ADHD、焦虑、抑郁等状况提供一致且以结果为导向的治疗。MEDvidi 的技术自动化了病历记录、随访和治疗计划,使医护人员能够专注于患者护理,同时提高效率和临床质量。
作为我们的高级/资深 DevOps 工程师,你将决定我们基础设施的下一步发展,并带领整个工程团队前进。这是一个真正意义上的亲力亲为的角色:你制定技术方向,亲手交付,并对公司所有方面的可靠性、成本、安全性以及开发者体验负责。你将直接与我们的产品工程团队合作,深入他们正在构建的内容,消除基础设施的阻力,并根据实际产品需求塑造基础设施。

为什么选择这个职位

  • 真正的全流程负责制。你负责结果,而不是任务,包括你的指标。可靠性、成本、性能、部署健康状况和开发者速度都由你负责:你设定目标并主动带来数据。
  • 无需管理负担的技术领导力。通过专业知识和主动性进行领导,最资深的工程师同时也是在代码中的那个人。你制定方向并提升标准,同时保持亲力亲为。
  • 默认支持 AI 原生。自主代理 AI 是我们工作方式的第一优先级。你将把任务委托给自主代理,将它们的输出整合到生产变更中,并帮助塑造整个组织如何使用 AI 进行开发。

职责

  • 主动制定并推动我们基础设施的技术愿景和季度路线图,明确权衡并设定可衡量的目标。
  • 运行并演进我们的 AWS + Kubernetes(EKS)基础设施:集群管理、自动扩展(Karpenter)、策略执行(Kyverno)和零停机操作。
  • 全流程负责基础设施即代码(Terraform,TypeScript 中的 AWS CDK)以及我们的 GitLab CI/CD(可重用/共享模板、OIDC、自托管 GitLab)。
  • 构建并负责团队真正使用的可观测性(Prometheus、Grafana、OpenTelemetry、OpenSearch、CloudWatch;日志管道、APM),提供共享的系统健康视图,而非没人打开的仪表盘。
  • 保持蓝绿部署和基于健康状态的自动化回滚快速而稳定;负责零停机的 PostgreSQL 模式迁移(扩展/
查看英文原文

Description
MEDvidi is an AI-powered mental healthcare platform setting a new standard for safe, effective, and scalable psychiatric care in the United States.
We combine licensed providers with proprietary AI tools to deliver consistent, outcomes-driven treatment for conditions like ADHD, anxiety, depression, and more. MEDvidi's technology automates charting, follow-ups, and treatment planning, freeing providers to focus on patient care while improving efficiency and clinical quality.
As our Senior/Staff DevOps Engineer, you'll decide where our infrastructure goes next and take the rest of engineering there. This is a hands-on role in the truest sense: you set the technical direction, ship it with your own hands, and own the outcomes for reliability, cost, security, and developer experience across the company. You'll partner directly with our product engineering teams, embedded in what they're building, removing infrastructure friction, and shaping infrastructure around real product needs.
Why this role

  • Real ownership, end to end. You own outcomes, not tasks, including your metrics. Reliability, cost, performance, deployment health, and developer velocity belong to you: you set the targets and bring the numbers to the table proactively.
  • Technical leadership without the management overhead. Lead by expertise and initiative where the most senior engineers are also the ones in the code. You set direction and raise the bar while staying hands-on.
  • AI-native by default. Agentic AI is a first-class part of how we work. You'll delegate to autonomous agents, integrate their output into production changes, and help shape how the whole org builds with AI.

Responsibilities

  • Set and drive the technical vision and quarterly roadmap for our infrastructure proactively, with clear trade-offs and measurable goals.
  • Run and evolve our AWS + Kubernetes (EKS) infrastructure: cluster management, autoscaling (Karpenter), policy enforcement (Kyverno), and zero-downtime operations.
  • Own Infrastructure as Code end to end (Terraform, AWS CDK in TypeScript) and our GitLab CI/CD (reusable/shared templates, OIDC, self-managed GitLab).
  • Build and own observability that teams actually use (Prometheus, Grafana, OpenTelemetry, OpenSearch, CloudWatch; log pipelines, APM) a shared view of system health, not a dashboard nobody opens.
  • Keep blue-green deployments and health-gated automated rollback fast and boring; own zero-downtime PostgreSQL schema migrations (expand/contract) and CI migration gating.
  • Own security engineering in a HIPAA environment: secrets hygiene (rotation, short-lived credentials, leak scanning), PHI-aware handling of logs and data, and Vault managed as code.
  • Partner directly with product teams to remove infrastructure friction and improve developer experience by design.
  • Use agentic AI as a core part of your workflow, integrating autonomous-agent output into production.

Requirements

  • 6+ years in DevOps/infrastructure engineering, with strong systems fundamentals and solid Linux administration and troubleshooting (performance analysis, resource management, process debugging).
  • Hands-on AWS (EC2, EKS, RDS, ElastiCache, Lambda, SQS, EventBridge, API Gateway, ALB, S3) and production Kubernetes/EKS (cluster management, node scaling, policy enforcement; Karpenter, Kyverno, or similar).
  • Strong Infrastructure as Code (Terraform and AWS CDK in TypeScript) and CI/CD ownership (GitLab CI/CD: reusable/shared templates, OIDC id_tokens, self-managed GitLab).
  • Monitoring and observability in practice (Prometheus, Grafana, OpenTelemetry, OpenSearch, CloudWatch; log-shipping and error tracking/APM).
  • Practical security engineering (secrets rotation, short-lived credentials, leak scanning, PHI-aware logging) and HashiCorp Vault as code (KV, JWT/OIDC auth for CI, policy design).
  • Blue-green deployments with automated, health-gated rollback; PostgreSQL zero-downtime schema migrations (expand/contract) and migration gating in CI.
  • Containers (Docker, ECR, immutable tags, image lifecycle) and network/protocol fundamentals (load balancing, TLS, DNS).
  • Hands-on agentic AI workflows (Claude Code or similar): delegating to autonomous agents and integrating their output into production.
  • A developer-focused mindset, strong problem-solving for complex system issues, and strong technical writing (docs-as-code, ADRs, design docs via MRs).
  • Fluent Russian and English (B1).
  • Experience working effectively in remote, distributed teams.

Would be a plus

  • Experience in a regulated/compliance-heavy environment (HIPAA, SOC 2, or similar).
  • Configuration management (Ansible) for VM fleet management.
  • Node.js application operations (pm2, npm), our stack is Node.js + TypeScript.
  • GitOps tooling (ArgoCD, Flux) and deeper PostgreSQL database administration.
  • AWS certifications.

What we offer

  • High-impact environment: A chance to contribute to a product-driven company in the medical tech space.
  • Growth & compensation: Clear growth opportunities and a competitive compensation package.
  • Remote-first: Fully remote long-term collaboration under a B2B model.
  • Health & wellness: Health insurance after the probation period, plus sports & wellness compensation.
  • Career development: Clear growth opportunities, including personalized English lessons via Preply.
  • Time off: 19 paid vacation days annually, 4 additional wellness days each year, and paid sick leave for the first 5 working days.
  • Culture: Thoughtful gifts for key life events and offline corporate events.

Ready to make an impact? Send us your profile, let's build something meaningful together.
MEDvidi is an Equal Opportunity Employer. We celebrate diversity and are committed to creating an inclusive environment for all employees and contractors. All qualified applicants will receive consideration without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status.
Originally posted on Himalayas

本页面信息整理自 Himalayas,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

← 返回全部职位