Skip to content
GitLabGitLabBengaluru, India

Backend Engineer, Geo Team

Build and operate Ruby on Rails backend services for GitLab Geo, disaster recovery, and backup and restore. The role focuses on PostgreSQL replication, distributed systems, production resilience, incident response, and large-scale migrations.

Salary not listed
RemoteBackend Engineering

About the job

Responsibilities

  • Design, build, and maintain backend functionality for Geo, Disaster Recovery, and Backup and Restore using Ruby on Rails and related services.
  • Improve replication and verification workflows for repositories and related data, focusing on performance, correctness, and operational simplicity.
  • Use PostgreSQL features, including logical replication, to support Geo, GitLab Dedicated migrations, and future Cells architectures.
  • Measure and reduce replication lag and verification failures through logs, metrics, and alerts, identifying disaster recovery risks early.
  • Collaborate with Database Engineering, Infrastructure, GitLab Dedicated, Tenant Scale, and other teams on initiatives affecting Geo and migrations.
  • Participate in incident response and post-incident reviews, turning findings into product and operational improvements.
  • Partner with Support, Site Reliability Engineering, and Customer Success on help requests, customer escalations, complex migrations, and incidents where Geo capabilities are critical.
  • Own projects from proposal and design through implementation, review, rollout, and production monitoring.
  • Provide constructive feedback on merge requests.

Requirements

  • Experience building and maintaining Ruby on Rails applications in production environments.
  • Experience with PostgreSQL or similar relational databases, including replication, indexing, and performance tuning.
  • Understanding of distributed systems and data replication concepts, including consistency, eventual consistency, and failure modes.
  • Experience building and operating background job or worker systems such as Sidekiq, including monitoring and retry strategies.
  • Knowledge of replication and verification patterns, such as ordered event delivery and checksums, and experience debugging issues such as replication lag.
  • Experience operating or debugging large production deployments across installation methods, operating systems, or cloud providers.
  • Familiarity with recovery point objective (RPO), recovery time objective (RTO), planned failover, and backup and restore for stateful systems.

Compensation and Benefits

  • Flexible paid time off.
  • Team member resource groups.
  • Equity compensation and employee stock purchase plan.
  • Growth and development fund.
  • Parental leave.

Skills

Ruby on RailsRubyPostgresLogical ReplicationDistributed SystemsData ReplicationSidekiqBackground JobsDatabase IndexingPerformance TuningObservabilityDisaster RecoveryBackup And RestoreRpo/RtoGitlab Dedicated