Senior Site Reliability Engineer
Scalepad
Who We’re Looking For
At ScalePad, we hire thoughtful builders who want their work to matter. Our roles are designed for people who thrive on driving impact, see ambiguity as an opportunity, and believe that raising the bar is a team sport.
We don’t bring people in to run playbooks. We hire people who want to rewrite them . And in this role, you’ll get to do that, while shaping the future of managed services for our global partners. (That’s what we call our customers.)
What is ScalePad
At ScalePad, we’re building more than software; we’re building confidence and clarity for the people who manage the technology businesses rely on every day.
Our mission: help MSPs evolve into MVPs (their clients’ most valuable partner) . Our tools turn them from reactive service providers into strategic advisors through a consistent, scalable Customer Success motion.
Our product suite unifies risk insights, client planning, and service delivery so MSPs can have smarter conversations, show clients their value, and grow their revenue.
But our purpose goes beyond our software. We’re creating a workplace where curious, growth-minded people can do their best work, where ideas are valued, progress is shared, and everyone belongs. Together, we’re creating a future where MSPs don’t just keep businesses running, they help them thrive. We believe that when our partners succeed, we all do.
With offices in Vancouver, Toronto, Montreal, and Phoenix and a global-first mindset. ScalePad has grown into a category leader trusted by 12,000+ partners across 60+ countries. We’ve been recognized for our products and corporate culture by MSP Today, G2, and Great Place to Work™, to name a few.
About the role
We’re looking for a Senior Site Reliability Engineer (SRE) to help strengthen and scale our multi-cloud platform and developer experience. This is a hands-on senior individual contributor role for an engineer who enjoys solving complex infrastructure challenges, improving reliability, and helping teams ship and operate software more effectively.
You’ll work closely with engineering leadership and alongside SREs across product domains. Reliability, infrastructure as code, internal tooling, and developer productivity will all be part of your day-to-day focus. You’ll spend your time building, operating, and improving the systems that engineering teams rely on while contributing to best practices and operational excellence across the organization.
What you’ll do
Get ready to go beyond order-taking. Your strategic responsibilities include:
Platform and Infrastructure
- Operate production infrastructure across AWS and Azure, including networking, IAM, and cost
- Build and operate Terraform modules and state at scale, keeping our infrastructure as code clean and reviewable
- Run Kubernetes in production: upgrades, scaling, troubleshooting, and platform improvements
- Operate and improve CI/CD pipelines that the entire engineering org depends on
Reliability & Operational Excellence
- Operationalize SLO/SLI frameworks and observability practices alongside the SRE team
- Drive incident response practice, on-call tooling, and incident review follow-through
- Reduce operational toil through automation across secret rotation, access management, and environment provisioning
- Contribute to capacity planning, disaster recovery, and resilience work across critical systems
Developer Experience & Technical Influence
- Build and maintain internal developer tooling that removes friction across engineering
- Lead rollouts of AI-native tooling for code review, testing, and engineering productivity, e.g., CodeRabbit, Copilot-class assistants, and internal AI workflows
- Own migrations and consolidation of internal platforms such as Jira, Confluence, ticketing, and documentation systems
- Partner with engineering and product leadership to identify and remove the biggest DX bottlenecks, and align infrastructure and reliability investments with business goals
- Mentor engineers and technical leads, fostering growth and knowledge-sharing within the organization
- Lead post-mortems and continuous improvement initiatives to strengthen reliability practices
Innovation & Continuous Improvement
- Evaluate and introduce new technologies, tools, and approaches to improve scalability and efficiency
- Drive standardization and modernization efforts across infrastructure and operational practices
- Lead proof-of-concept and experimentation initiatives to validate new reliability solutions
What we’re looking for
We care about what you can do more than where you’ve done it. However, experience in the following areas will help you hit the ground running in this role:
Must-haves
- 5+ years of experience in software engineering, infrastructure, or related technical disciplines, with a focus on Site Reliability Engineering (SRE), DevOps, Platform Engineering, or similar roles.
- Strong expertise in cloud infrastructure, distributed systems, networking, and observability practices
- Experience designing and operating highly available, scalable production systems
- Deep understanding of scripting, automation, infrastructure as code, CI/CD, and operational best practices
- Experience implementing SLO/SLI frameworks and reliability engineering methodologies
- Incident management, troubleshooting, and on-call experience in complex production environments
- Passion for mentoring engineers and improving engineering culture
Nice to Have
- Experience rolling out AI tooling in an engineering organization
- Experience leading tooling and platform migrations such as Jira, Confluence, or observability stacks
- Experience with chaos engineering practices and reliability testing
- Experience optimizing large-scale cloud infrastructure costs
Perks
ScalePad offers our employees a blend of purpose, growth, and genuinely great perks.
- Everyone’s an owner. Share in our success through our Employee Stock Ownership Plan (ESOP) and RRSP matching.
- Support for growing families. Parental leave programs are in place to support you and your family when it matters most.
- Structured mentorship with builders. Join opt-in mentorship programs and learn directly from founders and senior leaders who’ve scaled multiple SaaS ventures and spent decades in the MSP industry.
- Invest in your growth every year. Access an annual professional development budget to level up your skills, your career, and your impact.
- Set yourself up with great tools. Work with brand new, top-of-the-line hardware and equipment so you can do your best work, whether you’re at home or in one of our hubs.
- Modern ways of working. Roles at ScalePad are structured as remote or hybrid, with hub locations in Vancouver, Toronto, Montreal, and Phoenix. Specific work models are outlined in each posting.
- Support for hybrid life. Receive a monthly stipend to help you create an effective hybrid or remote work environment.
- Well-being and time to recharge. Take care of yourself with 100% employer-paid benefits.
Before You Apply
This is a full-time role for those who are eligible to work in Canada. We thank all applicants for taking the time to apply, but only candidates who make it to the next stage will be contacted.
Note on AI Use : ScalePad uses AI technology to support certain administrative aspects of our hiring process, such as transcription, note-taking, and interview documentation. These tools are strictly used to assist our team and have no influence on candidate evaluation or hiring decisions.
No recruiters, please.
$140k - $180k per year
...and online privacy for all. Right now we are looking for a Site Reliability Engineer to help us tame DNS. About the Position Linux system... ...Thank you for considering this opportunity. Funded.club Senior Recruiters partner exclusively with Startups and are in direct...SeniorFull timeDirect hireWork at officeRemote work$123k - $160k per year
...learning, and shaping the future of customer growth, you’ll find your place here. We are seeking a highly experienced Senior Site Reliability Engineer to own the reliability, performance, and operational excellence of our large-scale, distributed infrastructure. You will...SeniorRemote jobLong term contractFull timeRelocation- ...We are seeking a Senior DevOps & Site Reliability Engineer to own the reliability, scalability, performance, and operational excellence of Medeloop’s platform. This role blends deep DevOps engineering—CI/CD pipelines, infrastructure as code, and cloud architecture—with SRE...SeniorHourly payFull time
$110k - $160k per year
...person to join our team working towards this goal, we would love to hear from you! Role Overview We're seeking a Senior Site Reliability Engineer to join our SaaS-Ops team within Shared Services Engineering. The team owns reliability and operational excellence for...SeniorFull timeInternshipWork at officeLocal areaFlexible hours$197.5k - $225k per year
...investors including Silver Lake Waterman, Moody’s, Sequoia Capital, GV and Riverwood Capital. About the Team: As a Senior Site Reliability Engineer, you will be a key technical leader driving the design and optimization of our Kubernetes-based infrastructure and CI/CD...SeniorFull time- ...in 2024 – but we're just getting started. As a Sr. Site Reliability Engineer, you'll be the guardian of our platform's reliability and performance... .... What You Bring to the Team: Design and implement reliable and scalable AWS architecture to meet the needs of the...SeniorFull timeWork at officeLocal areaRemote workWork from homeHome officeWeekend work
$145k - $185k per year
...environments, and the infrastructure underneath that platform is what makes simulation at scale possible. We're hiring a Senior Site Reliability Engineer to help build and operate that infrastructure. This role sits at the core of how we run large-scale, distributed...SeniorRemote jobFull time- ...a part of our journey! About the role We are committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building and leading processes to ensure the reliability,...SeniorFull timeLocal areaRemote workHome officeFlexible hours
$153k - $187k per year
...stress into a clear signal owners can use to run stronger, more resilient businesses. We’re looking for an incredible Senior Site Reliability Engineer to join our SRE team. We aim to make reliability, security, and speed reinforce one another so that the platform becomes...SeniorFull timeInternship- ...environments. We are a modern, IoT-enabled, cloud-based tool for reliability, safety, and operations of physical equipment and... ...valuing the company at $2.5 billion. We’re looking for a Site Reliability Engineer (SRE) to help advance MaintainX’s reliability, observability...Full time
- ...trillion in AUM and 22 global investment banks. For more information, please visit . The Role CMG is looking for a Site Reliability Engineer (SRE) with a strong focus on monitoring, observability, and alerting to ensure the reliability, performance, and...Remote jobFull timeLocal area
$101.2k - $136.9k per year
...strategy across digital banking, core banking, data platforms and member facing services. We are looking for a highly skilled Site Reliability Engineer (SRE) who will help build, operate and continuously improve the reliability, performance, security and automation of our...Permanent employmentFull timeInternshipWork at officeImmediate startHome officeFlexible hours2 days per week$20 per day
...our careers page to see how you can grow with us! As a Site Reliability Engineer at Hiive, you will be responsible for ensuring the... ...performance and system behavior, and ensuring these services are reliable, scalable, and cost-efficient in production. In this role...Full timeSummer holidayRelocation$120k - $200k per year
...and many more. ABOUT THE ROLE At LayerZero, our Site Reliability Engineering (SRE) team is at the intersection of software and systems engineering... ...internal systems to those external users interact with—are reliable, meet the uptime expectations of our users, and...Full time- ...challenges with continuous learning opportunities, then Tescys could be a good fit for you! About the Role We are looking for a Site Reliability Engineer to join our Network and Security Operations Center (NOC), a team at the heart of platform reliability for mission-critical...Long term contractPermanent employmentFull timeWork at officeRemote work
- ...like an environment that you believe could work for you then read on to find out more. The role: We’re looking for a Site Reliability Engineer to manage, maintain, improve and provide support on our platform. You will be curious by nature, always looking for ways to...Full timeRemote work
- ...Canada say it is a great place to work, compared to 60% at a typical company. Role Summary We are seeking an experienced Site Reliability Engineer to help design, build, and operate the infrastructure that underpins the build pipelines that allow our companies to produce...SeniorFull timeContract workLocal area
$107k - $161k per year
...office to meet with your team for events or meetings. Join Our Team GoDaddy is seeking a highly skilled and motivated Senior Site Reliability Engineer to join our Database Infrastructure team. This role focuses on designing, developing, and deploying automated solutions...SeniorFull timeSecond jobWork at officeLocal areaRemote workWork from home- ...might just be in the right place! We’re looking for a Staff Site Reliability Engineer to join our Data team in Canada. As a Staff Data SRE,... .... ~10% – Mentorship & Technical Review: Actively mentor Senior and Intermediate SREs. You set the bar for IaC quality, observability...SeniorFull timeWork at officeRemote workFlexible hoursShift work
- ...matter. Your Impact As a senior contributor in the APX SRE... ...relentless about the high quality, reliability, and security our customers... ...will reach the entire engineering organization to enable product... ...confidence. Exemplify cloud-native site reliability best practices...SeniorLong term contractFull timeRemote workFlexible hours
- ...visualizing relationships between entities in the system. As a Site Reliability Engineer you will be responsible for the availability, latency,... ...managers. Finally we will ask you to meet with a number of our senior leaders to make sure that you are making the most informed...SeniorFull timeFlexible hours
- ...The Site Reliability Engineering organization at Pinterest is accountable for ensuring overall Pinterest availability as well as enhancing Engineering teams’ capability to design, build and operate robust systems at scale. Pinterest’s applications and infrastructure that...SeniorFull timeWork at officeRelocationRelocation package
- ...made, and we are not tiptoeing into it. We are rebuilding our engineering culture around a simple belief: AI changes everything. How teams... ...team builds on — and this role is at the center of keeping it reliable, fast, and scalable. As a Staff SRE, you'll own the infrastructure...Remote jobFull timeInternshipWork at officeLocal areaFlexible hoursShift workWeekend work
$120k - $160k per year
...ABOUT YOU We are looking for a Site Reliability Engineer (Monetization) who is pragmatic, product-minded, and equally comfortable writing... ...developers ship features and what infrastructure needs to stay reliable - and to bring the reliability lens into design decisions...Long term contractFull time- ...client base with operations throughout North America, Central America, Europe, Australia, and Japan. As one of our Lead Site Reliability Engineers, you will combine hands-on technical expertise with strategic technical leadership across infrastructure and software development...Long term contractFull time
- ...Gauss Labs is seeking a highly skilled Site Reliability Engineer to join our team in Vancouver. As an SRE at Gauss Labs, you will play a critical role in ensuring our industrial AI platform's reliability, performance, and scalability. You will be responsible for building and...Full time
$180.4k - $230.4k per year
...Coalition. About the role We are looking for a Staff Site Reliability Engineer to lead AI enablement across our engineering organization.... ...and tooling infrastructure to ensure AI-generated output is reliable, secure, and production-worthy. This role owns that layer....Full timeRemote workHome officeFlexible hoursShift work$69k - $90k per year
...UN-endorsed Zero Project. About the role As a Junior Site Reliability Engineer (SRE) at Fable, you will help support the reliability,... ...build more accessible digital experiences, and maintaining reliable, high-performing systems is critical to delivering that impact...Full timeInternship$260k - $275k per year
...enterprises • Solve complex reliability challenges at scale • Influence architecture and engineering culture at a company level •... ...will focus on creating reusable, reliable, and scalable solutions that abstract... ..., Platform Engineering, or Site Reliability Engineering role,...Full time$150k - $240k per year
...obsessed about achieving the high quality and reliability our customers demand. You will work... ...technical deliverables will reach the entire engineering organization to enable product teams to... ...-effective. Exemplify cloud-native site reliability best practices. Write code...Long term contractFull timeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer intern Remote
- site reliability engineer remote Remote
- site reliability engineer sre Remote
- site reliability engineer Remote
- senior site reliability engineer Remote
- sénior service a la clientèle Remote
- senior network engineer Remote
- senior engineer Remote
- senior graphic designer Remote
- senior mechanical design engineer Remote
