站点可靠性工程师I
Site Reliability Engineer I
关于Backblaze
Backblaze是开放云运动中的对象存储领导者,通过专为释放预算、减轻管理员负担和激发创新者而设计的云存储推动客户成功。与我们的合作伙伴一起,我们帮助客户摆脱限制性、价格过高的传统解决方案,掌握开放云的全部力量。
成立于2007年,我们在不到300万美元的外部资金支持下发展业务,直到2021年在纳斯达克股票交易所进行了传统的IPO。如今,Backblaze年收入超过1亿美元,是领先的专用存储云服务,为175多个国家的50多万客户(包括企业、开发者、IT专业人士和个人)管理超过三千亿千兆字节的数据存储。
但尽管我们过去有很多值得庆祝的地方,前方的机会几乎同样多。我们正在寻找一名初级站点可靠性工程师加入我们的团队!
你将负责:
- 作为所有影响客户的事件的第一联系人
- 成为解决技术问题的关键推动者
- 确保事件管理流程得到遵循,并完成事件事后分析以记录流程偏差和改进领域
- 向管理层提供一致的沟通
- 响应Zabbix警报/定期监控Zabbix,通过直接处理警报或升级来响应。如果采取了直接行动,需确认每个警报,或与升级联系人联系
- 确保升级被成功交接
- 确保所有站点的Pod健康状况(在Zabbix上定义Pod警报)
- 处理每日Pod文件系统检查
- 为DC Techs排错技术问题 -> 高级Pod问题、部署问题、迁移排错和Ansible剧本问题
- 识别并升级任何与网络相关的潜在问题
- Vault预部署配置和测试
- 开始Vault迁移,监控迁移Pod,处理相应的迁移Pod健康检查
- 文档/自动化日常事项
- 文档/提供即将部署的网络IP地址
- 监控服务器农场的发布/更新,发现问题时及时升级
- 参与值班轮班
- 协助其他TechOps团队成员处理任务
- 对组织生产力的改进提出建议
- 能够在正常工作时间外工作
查看英文原文
About Backblaze
Backblaze is the object storage leader in the open cloud movement, fueling customer success with cloud storage built purposefully to unlock budgets, unburden administrators, and unleash innovators. Together with our partners, we’re helping customers break free from the restrictive, overpriced legacy solutions that hold them back, and blaze forward with the full power of the open cloud in their hands.
Founded in 2007, we scaled the business with less than $3 million in outside funding until 2021, when we did a traditional IPO on the Nasdaq stock exchange. Today, Backblaze generates over $100m in revenue and is the leading specialized storage cloud - managing over three billion gigabytes of data storage for 500K+ customers in 175+ countries, including businesses, developers, IT professionals, and individuals.
But while there is a lot to celebrate in our past, there is almost as much opportunity ahead of us. We are seeking a Site Reliability Engineer I to join our team!
What You’ll Do:
- Act as first point of contact for all customer affecting issues
- Be a Key Driver for managing the resolution of technical problems
- Ensure that incident management processes are following and that incident post-mortems are completed to capture process deviations and areas for improvement
- Deliver consistent communication to Management
- Respond to zabbix alerts/regular monitoring of zabbix, either by taking direct action on alerts or escalating. Acknowledge every alert if direct action taken, or with escalation point of contact.
- Make sure escalations are handed off successfully.
- Ensure health of pods across all sites (define pod alerts on zabbix).
- Work through daily filesystem checks for pods.
- Troubleshoot technical issues for DC Techs -> advanced pod questions, deployment questions, migration troubleshooting, and ansible playbook issues.
- Identification and escalating any potential issues regarding the network.
- Vault pre-deployment configuration and testing.
- Start Vault Migrations, monitor migration pods, handle applicable migration pod health checks.
- Document/Work on automating Daily Items.
- Document/Provide Network IP's for upcoming deployments.
- Monitor Releases/Updates to the Server Farm, escalate issues as they arise.
- Engaging in on-call rotation shifts.
- Assist fellow TechOps team members in handling tasks.
- Making recommendations for improvements in organizational productivity.
- Be able to work outside of normal business hours(weekend shift, holidays & evenings) as needed
The Right Fit:
- Must be located in Bangalore.
- 2 - 4 years of relevant experience.
- Knowledge of Sysadmin and Linux skills.
- Desire to learn and develop all necessary technical skills.
- Strong analytical thinking.
- Strong skills in working with different teams and communication.
- Knowledge of network cabling, network classification, and network topology.
At this point, we hope you're feeling excited about the job description you're reading. Even if you don't meet every requirement, we still encourage you to apply. Learning, developing, and growing are key parts of our culture. We're eager to meet people who believe in our mission and can contribute to our team in various ways. We want people to feel comfortable expressing their true selves and to come, stay, and do their best work here.
At Backblaze, we value being fair and good to our customers, partners, and employees. That’s why diversity, equity, and inclusion are at the core of our values. We are committed to fostering a workforce where all employees feel a sense of belonging regardless of race, ethnicity, nationality, gender, sexual orientation, age, religion, socio-economic status, ability, veteran status, and education. We believe that our dedication to cultivating a diverse workspace not only allows us to better serve our customers in over 175 countries, but further reinforces our commitment to doing the right thing. We are proud to be an Equal Opportunity Employer.
To understand more about the data we collect and process as part of your application, please view our Backblaze Employee Privacy Notice.