Senior Site Reliability Engineer, Tenant Services: Geo
GitLab · Remote, India
This listing is from the archive and may be closed. Browse the latest experienced jobs for current openings.
GitLab is hiring a Senior Site Reliability Engineer, Tenant Services: Geo in India.
GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, in
Responsibilities
- Execute Dedicated Geo migrations and cutovers end-to-end, including planning, pre-cutover validation, execution, and post-cutover verification and cleanup.
- Join the team’s shift and weekend coverage rotation for Dedicated cutovers across EMEA and US hours, and participate in the SaaS Site Reliability Engineering (SRE) on-call rotation to respond to incidents that impact GitLab.com availability.
- Operate and improve the Geo operational surface for Dedicated, including:
- Environment preparation and data hygiene checks prior to migrations.
- Execution of replication, validation, and cutover procedures.
- Handling Geo-related escalations from Support and internal partners.
- Design, build, and maintain automation, tooling, and runbooks that make migrations, cutovers, and Geo escalations as “boring” and repeatable as possible.
- Run our infrastructure with tools such as Ansible, Chef, Terraform, GitLab CI/CD, and Kubernetes; contribute improvements back to GitLab’s product and infrastructure where appropriate.
- Build and maintain monitoring, alerting, and dashboards that:
- Detect symptoms early, not just outages.
- Track migration and cutover success rates, duration, rollback frequency, and related SLOs.
- Collaborate closely with:
- The core Geo team on improving Geo features and operability.
- Dedicated migrations and Support on migration planning, customer communications, and escalation handling.
- Other Infrastructure teams on capacity planning, disaster recovery, and reliability improvements.
- Contribute to readiness reviews, incident reviews, and root cause analyses, turning learnings into changes in automation, process, or product.
Preferred qualifications
- Experience working with disaster recovery technologies.
- Experience with managed/hosted environments similar to GitLab Dedicated, including regulated or compliance-sensitive customers (e.g., SOC2, ISO).
- Prior work on large-scale data migrations or cutovers where customer data integrity, performance, and downtime risk had to be carefully balanced.
- Hands-on experience designing and operating database replication, backup/restore, and cutover workflows (for example, PostgreSQL or cloud-managed equivalents such as AWS RDS), including planning and executing low-risk migrations for large datasets.
- Experience with multi-tenant architectures, sharding, or routing strategies in high-traffic SaaS platforms.
- Familiarity with GitLab (self-managed or SaaS), and/or contributions to open source projects.
How GitLab Supports Full-Time Employees
- Benefits to support your health, finances, and well-being
- Flexible Paid Time Off
- Team Member Resource Groups
- Equity Compensation & Employee Stock Purchase Plan
- Growth and Development Fund
- Parental Leave
Please note that we welcome interest from candidates with varying levels of experience; many successful candidates do not meet every single requirement. Additionally, studies have shown that people from underrepresented groups are less likely to apply to a job unless they meet every single qualificatio
About GitLab
See the company's official careers page for full details, then apply using the button below.