Site Reliability Engineer
TOTEM Recruteur de talent
Employment Status: Permanent
Schedule: 40 hours/week – 100% remote work
We are looking for an experienced Site Reliability Engineer to join a team responsible for the reliability, performance, and resilience of high-availability SaaS platforms.
Working in an AWS and Kubernetes environment, you will help design, automate, monitor, and continuously improve the infrastructure supporting critical cloud-based services. This role combines hands-on engineering with operational leadership, giving you direct ownership of system availability, scalability, and incident response.
You will be involved throughout the service lifecycle, from architecture and launch preparation to production monitoring and continuous improvement.
Responsibilities
- Partner with engineering teams during system design, capacity planning, launch readiness, and production deployment.
- Monitor and improve service availability, latency, performance, and overall system health.
- Identify recurring operational issues and implement sustainable solutions that improve scalability and resilience.
- Define and evolve observability practices, including dashboards, alerts, SLOs, and SLIs.
- Build and maintain automated infrastructure using Terraform and CI/CD pipelines.
- Develop automation and operational tooling to reduce manual intervention and support self-healing systems.
- Coordinate incident response and act as Incident Commander during critical production events.
- Facilitate blameless post-incident reviews and ensure that corrective actions are completed.
- Use AI-assisted engineering tools responsibly to accelerate development and operational workflows.
- Maintain clear technical documentation and contribute to the continuous improvement of SRE practices.
Required profile
- Significant experience in Site Reliability Engineering, Cloud Engineering, DevOps, or a similar infrastructure-focused role.
- Experience supporting complex or large-scale SaaS environments with high availability requirements.
- Strong hands-on knowledge of AWS services and architecture, including multi-account environments, VPC, EC2, and EKS.
- Proven experience operating and troubleshooting Kubernetes environments at scale.
- Strong knowledge of Infrastructure as Code, particularly Terraform.
- Experience building or maintaining CI/CD pipelines using GitLab, Jenkins, or comparable tools.
- Experience with enterprise observability platforms such as Datadog, Prometheus, Grafana, or equivalent solutions.
- Strong scripting skills using Python, Bash, or a similar language.
- Direct experience participating in on-call rotations, coordinating incident response, and conducting post-incident reviews.
- Familiarity with Java or .NET application environments is considered an asset.
- Ability to communicate clearly and collaborate with development, infrastructure, security, and operations teams.
- Must be legally authorized to work in Canada.
What to Expect
- Remote-first work environment within Canada.
- Occasional visits to a local office or participation in in-person meetings may be required, representing less than 10% of the role.
- Participation in a scheduled on-call rotation is required.
Does this opportunity sound like a good fit for you? Apply now through our website or by sending your resume to View email address on emploisti.com.
Thank you for your interest in this position; only candidates who meet our client’s requirements will be contacted.
The masculine gender is used as a neutral form.
#totemtech
$110k - $120k per year
...professional development support, discounts through Perkopolis, and recognition programs that celebrate your impact The Job: Site Reliability Engineer The Site Reliability Engineer is responsible for ensuring the availability, performance, and resilience of the...SuggestedFull timeTemporary workInternshipWork at officeRemote work$100k - $125k per year
...We are seeking an experienced and motivated Software Engineer to join our dynamic Site Reliability Engineering (SRE) team. As a Site Reliability Engineer,... ...culture Tech at Tipalti Our tech teams are the engine behind our business. Tipalti’s tech ecosystem is extremely...SuggestedFull timeWork at officeFlexible hours- ...Docebians around the world and help us reinvent the way people learn, because learning never stops. Role Overview As a Senior Site Reliability Engineer, you'll take a hands-on lead role in high severity incident response while also shaping the underlying infrastructure that...SuggestedFull timeFor contractorsWork at officeWorldwide3 days per week
- ...learning never stops. The Adventure Ahead As the Manager of Site Reliability Engineering (SRE), you will lead a talented team of engineers dedicated... ...Product, Engineering, and Support leadership to integrate reliable SRE practices into early planning and product delivery...SuggestedLong term contractFull timeFor contractorsWork at officeWorldwide3 days per week
$140k - $155k per year
...profiles! This is a hands-on senior engineering role focused on improving production... ...enable engineering teams to ship secure, reliable, and scalable software with confidence. You... ...engineers on cloud-native technologies, site reliability engineering principles, and operational...SuggestedRemote jobPermanent employmentFull timeFlexible hours$140k - $182k per year
...client base with operations throughout North America, Central America, Europe, Australia, and Japan. As one of our Senior Site Reliability Engineers, you will be 100% hands-on across infrastructure and software development. You will support and help inform the evolution of...Full time$154k - $200k per year
...client base with operations throughout North America, Central America, Europe, Australia, and Japan. As one of our Lead Site Reliability Engineers, you will combine hands-on technical expertise with strategic technical leadership across infrastructure and software development...Long term contractFull time$153.82k - $277k per year
...place where you can thrive, we can’t wait to meet you. Site Reliability Engineers (SREs) at Braze are responsible for keeping all internal-facing... ...level feature requirements and helping translate them into reliable, highly scalable technology stacks. You will be responsible...Permanent employmentFull timeInternshipWork at officeLocal areaRemote workFlexible hoursRotating shift- ...of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a systems thinker. You’ll create middleware and platform guardrails...Full timeWork at officeLocal areaRemote workWorldwideMonday to fridayFlexible hours
- ...competitive advantage. Job Description The SRE Role · SREs are engineers with the right mix of knowledge and skills in software... ...experimentation and observation to entire systems to improve reliability, performance and operability). · We constantly evaluate products...Full time
$140k - $180k per year
...the shared infrastructure that all other engineering teams at Tripstack depend on. This spans... ...boundary sits, and the deployment patterns and reliability standards that apply across both. This... ...engineering depth - interconnects, BGP, site-to-site VPN, cross-region peering GDPR...RemplacementFull timeWork at officeImmediate startFlexible hours- ...AI more natural, capable, and useful. We are looking for a Site Reliability Engineer to help build and operate the infrastructure behind that work... ...storage, scheduling, and the operational tooling that keeps them reliable. This is a hands-on role for someone who enjoys taking...Full timeRemote work
$80 - $110 per hour
...Networks is headquartered in Irvine, CA USA with Asia HQ in Singapore and also operating in Denmark, Spain and Vietnam. The Site Reliability Engineer will improve the availability, performance, scalability and recoverability of AXON Networks cloud solutions. You will...Remote jobFull timeContract work$92k - $118k per year
...Build reliable, resilient cloud platforms that keep critical financial services running. The Role We are looking for an experienced Site Reliability and DevOps Engineer to join the SRE team within our growing Corporate Action Processing group. You will bring 3+ years...InternshipImmediate start- ...customers. Cohere is a team of researchers, engineers, designers, and more, who are passionate... ...building high-performance, scalable and reliable machine learning systems? Do you want to... ...NLP applications? We are looking for a Site Reliability Engineer to join the Model...Full timeWork at officeRemote workFlexible hours
- ...visibility, and optimize spend across the enterprise. The Site Reliability Engineer III (SRE III) plays a critical role in ensuring Emburse’s... ...Excellence & Automation Design, develop, and automate reliable cloud infrastructure and platform services. Apply Infrastructure...Full timeManual laborLocal areaFlexible hours
$144k - $200k per year
**The Team** Platform Engineering is the department within SRE that is responsible for a range... ...role in developing and maintaining the reliable and globally connected multi-cloud network... ...Overview** We are seeking a talented Site Reliability Engineer (SRE) with a strong...Full timeWork at officeRemote workWorldwideFlexible hours$130k - $180k per year
...Our Platform is growing and we are looking to hire a Senior Site Reliability Engineer (SRE) / Cloud Engineer Our main Cloud Platform is Azure... ...implement, and maintain CI/CD pipelines, enabling rapid and reliable software releases. Automate and optimize our infrastructure...Full timeRemote workVisa sponsorshipWork visaFlexible hours$110k - $125k per year
...systems bringing #BetterGlobalHealth to patients everyday! Apply today and find plenty of reasons to SMILE! The Cloud Site Reliability Engineer (SRE) is responsible for ensuring the reliability, scalability, and performance of production-grade services deployed across...Full timeRemote workFlexible hours$116k - $235.1k per year
...About the Role: Site Reliability Engineering (SRE) at Tubi is not a traditional operations team. We are a software engineering organization that applies a developer's mindset and toolkit to the challenges of building and running large-scale, distributed systems. Our mission...Long term contractRemplacementFull timeTemporary workWork at officeLocal areaImmediate startFlexible hours2 days per week- ...that are particularly strong in a few areas, and have some interest and capabilities in others. About the Role: As a Site Reliability Engineer , you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS...Full time
- ...To learn more about CIBC please visit What youll be doing As a member of CIBCs Application Reliability Engineering Platform team the Consultant Site Reliability Engineering will play a key role in CIBC Digital and Client Experience Technology to support the...Full timeContract work3 days per week1 day per week
- ...a skill on their LinkedIn profiles! This is a hands-on senior engineering role focused on improving production resilience, strengthening... ...operational practices that enable engineering teams to ship secure, reliable, and scalable software with confidence. You will help...Permanent employmentFull timeRemote work
$140k - $165k per year
...transformation, transitioning from traditional service desk operations to an AI-driven, self-healing operational fabric. As Lead Site Reliability Engineer (SRE), you will lead enterprise observability, AIOps automation, and SRE governance across a global footprint. You will...$78.62 - $89.34 per hour
Our client, is seeking a talented and proactive Site Reliability Engineer (SRE) / Senior Database Platform Engineer to join their core Data Engineering and Operations team. In this engineering-focused role, you will move beyond traditional database administration to act...Long term contractPermanent employmentFull timeContract workWork at office$120k - $170k per year
...We are seeking a highly skilled and motivated Senior DevOps Engineer to join our dynamic team and play a key role in designing, implementing... ..., investigate incidents, and drive improvements that increase reliability and operational efficiency; Write and maintain automation,...Full timeWork at officeLocal areaFlexible hours- ...Description Application Support (Windows, UNIX/Linux, OpenShift, and PostgreSQL) This role is part of the Production Support and Reliability Engineering team, responsible for ensuring the stability, availability, and performance of enterprise applications running across Windows...Permanent employmentFull timeLocal area
$82.2 - $89.34 per hour
We are seeking a highly skilled Site Reliability Specialist IV for an enterprise-level contract opportunity based in Toronto. In this role, you will take on a premier cloud engineering, platform automation, and operational reliability capacity, specializing in designing, building...Long term contractPermanent employmentContract work$141k - $191k per year
...and develop your career. As an SRE Manager, you will lead a team of 10+ engineers, oversee their development and ensure operational excellence. About the Role: In this opportunity as Site Reliability Engineering Manager , you will be responsible for: Team Leadership...Work at officeLocal areaFlexible hours2 days per week3 days per week- ...contributors of all seniorities. Join Tenstorrent as a Staff Reliability Engineer and help define the reliability strategy behind the next... ...takes to run hands-on testing at manufacturing and test partner sites, including in Taiwan. Tenstorrent offers a highly competitive...Permanent employmentInternshipSecond job
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- senior site reliability engineer Toronto, ON
- site reliability engineer Toronto, ON
- site reliability engineer intern Toronto, ON
- site safety Toronto, ON
- site maintenance Toronto, ON
- site carpenter Toronto, ON
- website developer Toronto, ON
- senior site reliability engineer
- site reliability engineer sre
- site reliability engineer
