Intermediate Site Reliability Engineer, Environment Automation
Gitlab
GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster.
The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software.
* Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab.
An overview of this role
You'll join the Dedicated team as a Site Reliability Engineer focused on Environment Automation , where your work will help power hundreds of isolated GitLab environments for our customers. In this role, you'll help keep these environments reliable, scalable, secure, and consistent by treating everything as code and contributing to automation across the entire lifecycle, from initial provisioning to day-to-day operations. Instead of operating a single platform, you'll collaborate with senior SREs to solve the unique challenges of managing many tenant environments in parallel, each with its own constraints and integration points.
You'll help define, deploy, and maintain GitLab environments across cloud providers using infrastructure as code, deployment packages, and Kubernetes. You'll contribute to automation that reduces manual work, assist in building tooling that orchestrates upgrades and configuration changes safely at scale, and support an observability stack that lets us understand and improve the health of every environment. Your work will directly impact how customers experience GitLab Dedicated and other managed offerings, enabling them to focus on building software while we ensure their GitLab environments are always production ready.
Some examples of work you'll do:
- Contribute to the design and evolution of infrastructure automation using Terraform, Ansible, and Kubernetes to provision, upgrade, and operate many GitLab environments with minimal manual effort
- Help debug and resolve production issues across Kubernetes clusters, GitLab components, and cloud services, then assist in building automation and safeguards that prevent similar issues from recurring
- Assist in creating and maintaining deployment and orchestration tools, such as Helm Charts, omnibus-gitlab configurations, and multi-tenant workflows, that make it easy for teams to manage GitLab environments at scale
What you'll do
- Contribute to automating operational tasks across many GitLab environments, from initial provisioning and configuration updates to upgrades and routine maintenance, helping reduce manual work and improve reliability at scale under the guidance of senior team members.
- Help build and refine the observability stack for multi-tenant GitLab environments so we monitor the right signals across Kubernetes, cloud services, and GitLab applications, supporting early issue detection and basic capacity tracking.
- Assist in responding to platform alerts and incidents, collaborating with Environment Automation SREs and engineering teams to troubleshoot production issues across multiple tenants and document findings.
- Support planning and implementation of infrastructure changes, capacity expansions, and new service rollouts for Dedicated and other managed GitLab environments, contributing to efforts that improve resource efficiency and environment isolation.
- Develop and maintain scripts, automation tools, and infrastructure-as-code workflows that manage parts of the GitLab environment lifecycle, enabling more repeatable, self-service operations over time.
- Apply and help implement best practices for running GitLab on Kubernetes and cloud platforms, focusing on day-to-day reliability, performance, and security while learning how to keep environments consistent.
- Participate in the on-call rotation for production GitLab environments with appropriate support, helping triage and mitigate incidents across clusters and cloud providers and contributing to post-incident reviews.
- Document operational tasks, runbooks, and lessons learned so they become clear, repeatable processes and can be candidates for future automation, improving shared knowledge and reducing manual toil across the team.
What you'll bring
- Experience working as an SRE or in a similar role operating production infrastructure, with an interest in automating the lifecycle of many environments or tenants in parallel, even if you have not yet done so at large scale.
- Hands-on experience with backend programming languages such as Golang, with the ability to read, understand, and modify infrastructure tools.
- Hands-on experience running Kubernetes-based workloads in production, including basic understanding of deployments, rollouts, and debugging common issues like crash loops, failed health checks, and scheduling problems.
- Familiarity with infrastructure automation and configuration management tools such as Terraform and Ansible, including experience working with modules, variables, and managing state safely for multiple environments.
- Solid understanding of Git-based workflows and infrastructure-as-code practices, with the ability to contribute to reusable modules, templates, and pipelines that make automation safer and more consistent.
- Experience working in distributed systems or cloud-based production environments, ideally in SaaS or managed service settings, with comfort participating in incident response and on-call rotations under guidance from more senior team members.
- A proactive mindset focused on automation and documentation—you look for opportunities to remove manual steps, improve runbooks, and turn repetitive tasks into reliable, self-service tools.
- Comfort working asynchronously across distributed teams and a desire to contribute to GitLab's values of collaboration, transparency, and iteration.
About the team
We are responsible for building, running, and evolving the entire lifecycle of the GitLab environments that power the GitLab Dedicated platform. You'll be part of our team focused on owning the reliability, scalability, performance, and security of automated single-tenant GitLab instances and their supporting services. GitLab Dedicated provides fully managed, isolated environments for customers around the world, which means your work directly impacts how organizations of all sizes run their mission-critical software delivery on GitLab. We operate in a fully distributed, asynchronous environment across multiple regions, collaborating on everything from infrastructure automation and environment lifecycle design to incident response and capacity planning. You'll be solving novel challenges at scale, from orchestrating infrastructure-as-code workflows across hundreds of tenants to designing the automation that keeps those environments consistent, secure, and up to date. We continuously seek to reduce complexity and improve efficiency by leveraging cloud vendor managed products and services where appropriate, ensuring GitLab Dedicated remains a best-in-class managed platform for our customers. For more on how we operate, see the relevant GitLab Dedicated and infrastructure handbook pages.
The base salary range for this role’s listed level is currently for residents of the United States only. This range is intended to reflect the role's base salary rate in locations throughout the US. Grade level and salary ranges are determined through interviews and a review of education, experience, knowledge, skills, abilities of the applicant, equity with other team members, alignment with market data, and geographic location. The base salary range does not include any bonuses, equity, or benefits. See more information on our benefits and equity . Sales roles are also eligible for incentive pay targeted at up to 100% of the offered base salary.
United States Salary Range
$103,600 - $222,000 USD
How GitLab Supports Full-Time Employees
- Home office support
Please note that we welcome interest from candidates with varying levels of experience; many successful candidates do not meet every single requirement. Additionally, studies have shown that people from underrepresented groups are less likely to apply to a job unless they meet every single qualification. If you're excited about this role, please apply and allow our recruiters to assess your application.
Country Hiring Guidelines: GitLab hires new team members in countries around the world. All of our roles are remote, however some roles may carry specific location-based eligibility requirements. Our Talent Acquisition team can help answer any questions about location after starting the recruiting process.
Privacy Policy: Please review our Recruitment Privacy Policy. Your privacy is important to us.
GitLab is proud to be an equal opportunity workplace and is an affirmative action employer. GitLab’s policies and practices relating to recruitment, employment, career development and advancement, promotion, and retirement are based solely on merit, regardless of race, color, religion, ancestry, sex (including pregnancy, lactation, sexual orientation, gender identity, or gender expression), national origin, age, citizenship, marital status, mental or physical disability, genetic information (including family medical history), discharge status from the military, protected veteran status (which includes disabled veterans, recently separated veterans, active duty wartime or campaign badge veterans, and Armed Forces service medal veterans), or any other basis protected by law. GitLab will not tolerate discrimination or harassment based on any of these characteristics. See also GitLab’s EEO Policy and EEO is the Law . If you have a disability or special need that requires accommodation , please let us know during the recruiting process .
- ...please visit . The Role CMG is looking for a Site Reliability Engineer (SRE) with a strong focus on monitoring, observability,... ...production rollouts. Besides the standard pull requests, test automation, code coverage tracking, containerization, and one-click...SuggestedRemote jobFull timeLocal area
$101.2k - $136.9k per year
...platforms and member facing services. We are looking for a highly skilled Site Reliability Engineer (SRE) who will help build, operate and continuously improve the reliability, performance, security and automation of our cloud platforms. This is a hands-on engineering role and you...SuggestedPermanent employmentFull timeInternshipWork at officeImmediate startHome officeFlexible hours2 days per week- ...location and working hours are no barrier. If this sounds like an environment that you believe could work for you then read on to find out more. The role: We’re looking for a Site Reliability Engineer to manage, maintain, improve and provide support on our platform....SuggestedFull timeRemote work
- ...productivity, reduced commuting on employee wellbeing and the environment, we are proud to be a digital-first company. The... ...good fit for you! About the Role We are looking for a Site Reliability Engineer to join our Network and Security Operations Center (NOC), a...SuggestedLong term contractPermanent employmentFull timeWork at officeRemote work
$20 per day
...our careers page to see how you can grow with us! As a Site Reliability Engineer at Hiive, you will be responsible for ensuring the... ...contributor, you will build scalable and resilient infrastructure, automate processes, and respond to incidents efficiently and effectively...SuggestedFull timeSummer holidayRelocation$120k - $200k per year
...and many more. ABOUT THE ROLE At LayerZero, our Site Reliability Engineering (SRE) team is at the intersection of software and systems... ...robust infrastructure, and reduce manual workload through automation. As a member of the SRE team, you'll tackle the unique challenges...Full time- ...might just be in the right place! We’re looking for a Staff Site Reliability Engineer to join our Data team in Canada. As a Staff Data SRE,... ...data infrastructure for batch/streaming workloads, ML/AI environments (Vertex AI, model serving, GPU-backed compute), and the BI...Full timeWork at officeRemote workFlexible hoursShift work
$65k - $130k per year
...recruit candidates from different backgrounds and foster a work environment that encourages employees to collaborate and learn from each... ...Your role: As a DevOps/SRE, you will be responsible for the reliability and smooth operation of your service in both production and...Full time$110k - $160k per year
...hear from you! Role Overview We're seeking a Senior Site Reliability Engineer to join our SaaS-Ops team within Shared Services... ...our highly available SaaS platform, a production Kubernetes environment serving law enforcement and government customers globally....Full timeInternshipWork at officeLocal areaFlexible hours- ...Intelligence platform for industrial and frontline environments. We are a modern, IoT-enabled, cloud-based tool for reliability, safety, and operations of physical... ...company at $2.5 billion. We’re looking for a Site Reliability Engineer (SRE) to help advance MaintainX’s...Full time
$120k - $160k per year
...ABOUT YOU We are looking for a Site Reliability Engineer (Monetization) who is pragmatic, product-minded, and equally comfortable writing... ...with experience in operating production services in a cloud environment (GCP/GKE or comparable) and partnering closely with product...Long term contractFull time- ...We are seeking a Senior DevOps & Site Reliability Engineer to own the reliability, scalability, performance, and operational excellence of Medeloop... ...using AWS CDK, CloudFormation, or Terraform, ensuring all environments are version-controlled and reproducible. Architect multi...Hourly payFull time
$123k - $160k per year
...teammates who take accountability and drive results in an environment where their work truly moves the business forward. We are... ...place here. We are seeking a highly experienced Senior Site Reliability Engineer to own the reliability, performance, and operational excellence...Remote jobLong term contractFull timeRelocation$140k - $180k per year
...and online privacy for all. Right now we are looking for a Site Reliability Engineer to help us tame DNS. About the Position Linux system... ...software to admin hosts Write Python scripts for automation Use your comprehensive understanding of DNS to troubleshoot...Full timeDirect hireWork at officeRemote work- ...relentless about the high quality, reliability, and security our customers... ...will reach the entire engineering organization to enable product... ...confidence. Exemplify cloud-native site reliability best practices... ...emphasis on testability, automation, and resilience in distributed...Long term contractFull timeRemote workFlexible hours
$69k - $90k per year
...-endorsed Zero Project. About the role As a Junior Site Reliability Engineer (SRE) at Fable, you will help support the reliability, performance... ...to build skills in cloud infrastructure, observability, automation, and platform engineering. Requirements...Full timeInternship- ...name a few. About the role We’re looking for a Senior Site Reliability Engineer (SRE) to help strengthen and scale our multi-cloud platform... ...review follow-through Reduce operational toil through automation across secret rotation, access management, and environment...Full timeInternshipRemote workWork from home
- ...committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building... ...or Docker Swarm. Strong experience with automation and configuration management tools such as...Full timeLocal areaRemote workHome officeFlexible hours
$260k - $275k per year
...enterprises • Solve complex reliability challenges at scale • Influence architecture and engineering culture at a company level... ...features faster in a multi-cloud environment Design and build core... ...robust, internal-facing tools and automation for infrastructure...Full time$150k - $240k per year
...achieving the high quality and reliability our customers demand. You... ...deliverables will reach the entire engineering organization to enable... ...effective. Exemplify cloud-native site reliability best practices.... ...utilizing CI/CD platforms to automate provisioning infrastructure,...Long term contractFull timeRemote work- ...visualizing relationships between entities in the system. As a Site Reliability Engineer you will be responsible for the availability, latency,... ...big impact on the company 2. Implement your ideas in an environment that strives for continuous improvement 3. Be part of a...Full timeFlexible hours
- ...in 2024 – but we're just getting started. As a Sr. Site Reliability Engineer, you'll be the guardian of our platform's reliability and... ...ambitious hotels running 24/7, while fostering a culture of automation, resilience, and continuous improvement across our engineering...Full timeWork at officeLocal areaRemote workWork from homeHome officeWeekend work
$180.4k - $230.4k per year
...Coalition. About the role We are looking for a Staff Site Reliability Engineer to lead AI enablement across our engineering organization.... ...— from AI-assisted code review and developer workflow automation to establishing security standards for emerging frameworks...Full timeRemote workHome officeFlexible hoursShift work$145k - $185k per year
...generation of autonomous systems in high-fidelity virtual environments, and the infrastructure underneath that platform is what makes simulation at scale possible. We're hiring a Senior Site Reliability Engineer to help build and operate that infrastructure. This role sits...Remote jobFull time- ...Gauss Labs is seeking a highly skilled Site Reliability Engineer to join our team in Vancouver. As an SRE at Gauss Labs, you will play a critical... ...to minimize downtime and restore service quickly. Automation: Developing and implementing automation tools and scripts to...Full time
- ...America, Europe, Australia, and Japan. As one of our Lead Site Reliability Engineers, you will combine hands-on technical expertise with... ...and beyond. Responsibilities: Define and drive the automation strategy for infrastructure tooling, establishing standards...Long term contractFull time
- ...The Site Reliability Engineering organization at Pinterest is accountable for ensuring overall Pinterest availability as well as enhancing Engineering... ...and opportunities for remediation Build tools and automation to eliminate toil and reduce operational overhead. Create...Full timeWork at officeRelocationRelocation package
$197.5k - $225k per year
...GV and Riverwood Capital. About the Team: As a Senior Site Reliability Engineer, you will be a key technical leader driving the design and... ...ensure production reliability, and embed best practices for automation, observability, and resilience. About the Role:...Full time- ...tiptoeing into it. We are rebuilding our engineering culture around a simple belief: AI... ...this role is at the center of keeping it reliable, fast, and scalable. As a Staff SRE, you... ...ownership — manage and evolve our AWS environment using Terraform, keeping EKS clusters, databases...Remote jobFull timeInternshipWork at officeLocal areaFlexible hoursShift workWeekend work
$153k - $187k per year
...resilient businesses. We’re looking for an incredible Senior Site Reliability Engineer to join our SRE team. We aim to make reliability, security... .../regulatory experience; Experience working in compliant environments such as SOC2 or PCI Experience driving large reliability...Full timeInternship
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Intermediate Site Reliability Engineer, Environment Automation. Be the first to apply!
- site reliability engineer intern Remote
- site reliability engineer remote Remote
- site reliability engineer sre Remote
- site reliability engineer Remote
- senior site reliability engineer Remote
- automation specialist Remote
- marketing automation specialist Remote
- test automation developer Remote
- test automation engineer Remote
- automation engineer - process controls systems Remote
