Senior Site Reliability Engineer, Environment Automation

GitLab
Remote, Canada; Remote, USPosted 5 March 2026

Job Description

<div class="content-intro"><p>GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster.</p> <p>The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our <a href="https://handbook.gitlab.com/handbook/values/">values</a> and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. <a href="https://www.youtube.com/watch?v=OuZIb5zszQI">Co-create the future with us</a> as we build technology that transforms how the world develops software.</p> <p>*<em>Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab.</em></p></div><p><strong>An overview of this role</strong></p> <p>As a Site Reliability Engineer (SRE) at GitLab, you’ll help keep all user-facing services and production systems reliable, scalable, and efficient. Our SREs combine a pragmatic operations mindset with strong software engineering practices to drive automation, reduce toil, and improve resilience across our platform.</p> <p>In the Environment Automation specialization, your focus is on operating and automating hundreds of GitLab environments—from initial provisioning to day-to-day maintenance tasks. </p> <p>Unlike other SRE roles, this position centers on automating the lifecycle of many tenant environments, ensuring they remain secure, consistent, and reliable at scale.</p> <ul> <li>Some examples of the projects you could work on:</li> <li>Designing infrastructure automation that provisions and operates GitLab environments using Terraform, Ansible, and Kubernetes</li> <li>Creating and maintaining deployment packages for GitLab, such as Helm Charts and omnibus-gitlab</li> <li>Building and operating Dedicated GitLab instances integrated with cloud-native services (e.g., GCP, AWS)</li> <li>Developing tools to orchestrate infrastructure-as-code workflows across multiple tenants</li> <li>Deploying and managing microservices on Kubernetes clusters at scale</li> <li>Enhancing GitLab’s observability stack (e.g., Prometheus, ELK) to support proactive monitoring and incident response</li> <li>Integrating with and operating infrastructure in cloud provider ecosystems (e.g., IAM, networking, storage)</li> <li>Championing and implementing cloud security best practices across automated infrastructure</li> </ul> <p><strong>What You'll Do</strong></p> <ul> <li><strong>Build Scale Multi-Tenant Infrastructure</strong>: Design and implement automation that provisions and manages hundreds of isolated GitLab environments using Terraform, Ansible, and Kubernetes. Manage complex state strategies and workspace configurations to support scale and maintainability.</li> <li><strong>Debug Resolve Production Issues</strong>: Troubleshoot issues across Kubernetes clusters, cloud services, and GitLab apps—identifying root causes of failed deployments, crash loops, and scheduling conflicts to ensure service continuity.</li> <li><strong>Automate Operations at Scale: </strong>Replace manual workflows with infrastructure-as-code solutions, including automated version upgrades, configuration rollouts, and ... (truncated, view full listing at source)
Apply Now

Direct link to company career page

Share this job