Intermediate Site Reliability Engineer
$130k - $150k per yearContactMonkey
Hey there! We're ContactMonkey
Our mission is to power measurable employee engagement worldwide, and we're looking for an Intermediate Site Reliability Engineer to join our Engineering team.
About the job
We're looking for someone who enjoys running production systems, understands how applications work, and brings practical application security experience.
You'll work closely with our SRE and development teams to maintain our infrastructure, improve deployments, investigate production issues, and address security risks. You'll also contribute to the technical controls that support SOC 2 audits and GDPR compliance.
Our environment includes AWS, Kubernetes on EKS, Terraform, Terragrunt, GitHub Actions, Prometheus, Grafana, and CloudWatch. Our applications use Ruby on Rails, Vue.js, and Node.js, with MySQL, PostgreSQL, and Sidekiq supporting the backend.
You'll take ownership of defined projects and operational improvements, with senior engineers available for guidance and review. There's room to develop deeper expertise in reliability, infrastructure automation, and application security.
Your impact
- Infrastructure & reliability: Maintain AWS and Kubernetes environments, troubleshoot production issues, and improve availability, performance, and resource usage.
- Terraform & Terragrunt: Build and maintain infrastructure as code, review plans, manage environment configuration, and address infrastructure drift.
- Deployments & developer experience: Improve CI/CD pipelines, deployment automation, release checks, and rollback procedures.
- Monitoring & incidents: Improve monitoring and alerts, join the on-call rotation, and contribute to incident reviews.
- Application security: Work with developers to assess vulnerabilities, review security risks, and validate fixes.
- CI/CD security: Maintain code, dependency, secrets, container, and infrastructure scanning.
- Cloud security: Strengthen IAM, secrets management, network controls, and Kubernetes security.
- SOC 2 & GDPR: Support technical controls, audit evidence, and personal data protection.
- Recovery: Test backups and recovery procedures, maintain runbooks, and support production readiness.
- Collaboration: Participate in code reviews, document changes, and support application, AI, and data engineering teams.
About you
- Around 3–5 years of experience in SRE, DevOps, platform engineering, cloud operations, or a related engineering role. Equivalent practical experience is welcome.
- Hands-on experience supporting production workloads in AWS.
- Experience writing and maintaining Terraform modules and Terragrunt configuration, including reviewing plans and working with remote state.
- Experience with Docker and Kubernetes, including troubleshooting deployments, services, health checks, and resource limits.
- A solid understanding of Linux, networking, DNS, and TLS.
- Ability to write automation in Python, Bash, Ruby, JavaScript, or another suitable language.
- Experience with Git, pull requests, CI/CD pipelines, and deployment workflows.
- Experience using logs, metrics, dashboards, and alerts to investigate production problems.
- Practical application security experience through vulnerability remediation, secure code review, threat modelling, or security tooling.
- Understanding of common web application risks, including broken access control, injection, authentication weaknesses, and sensitive data exposure.
- Familiarity with IAM, least privilege, secrets management, and encryption.
- Working knowledge of SOC 2 controls and GDPR principles relevant to engineering, including access restrictions, data minimization, retention, and deletion.
- Clear communication skills and good judgment about when to work independently, request a review, or escalate an issue.
How you can stand out
- Supported Ruby on Rails or Node.js applications in production
- Worked with MySQL, PostgreSQL, Redis or Valkey, and Sidekiq
- Experience with GitHub Actions, Argo CD, Helm, Karpenter, or KEDA
- Built useful monitoring with Prometheus, Grafana, or CloudWatch
- Supported services across multiple AWS regions
- Helped remediate penetration-test findings or contributed technical evidence to a SOC 2 audit
- Participated in backup restoration or disaster recovery exercises.
- Familiar with securing AI integrations, agent workloads, or MCP services
- You hold relevant AWS, Kubernetes, Terraform, or security certifications.
Working with the team
You'll work closely with SRE and application engineers, and support AI and data initiatives where they depend on shared infrastructure.
We value people who ask questions, explain their reasoning, and leave systems easier for others to understand. You'll take part in technical discussions and code reviews, share what you learn, and help improve how we operate.
Reducing recurring incidents, unnecessary alerts, and repetitive manual work is part of the role. We want reliable systems and sustainable operations for the people supporting them.
What we bring to the table
100% employer-paid benefits and a Health Spending Account from day one
Work from anywhere in the world for up to four weeks
A stock option plan so you can own a piece of our success
An RRSP Group Savings Plan
A generous vacation package
A personal development budget
One personal day and two volunteering days
Your birthday off
Five health days per year
A downtown Toronto office for hybrid work, with plenty of snacks
Compensation and work details
The salary range for this role is $130,000-$150,000 Compensation is based on experience, skills, and our internal compensation framework and equity.
We're happy to discuss compensation throughout the hiring process.
This is a full-time position on our SRE team.
The role includes a shared on-call rotation after onboarding. We'll discuss the schedule, escalation support, and expectations during the interview process.
Who we are
ContactMonkey helps organizations create, send, and measure internal communications directly within Outlook and Gmail.
Our platform brings together email design, employee engagement tools, and analytics so internal communications teams can understand what reaches their people and what gets a response.
As the product grows, we're investing in reliability, security, and tooling that helps our engineering teams deliver changes confidently.
Diversity is our strength
At ContactMonkey, we're building products for diverse organizations, and we need a diverse team to do that. We strongly encourage applications from everyone regardless of race, religion, colour, national origin, gender, sexual orientation, age, marital status, or disability status.
We are committed to an accessible hiring process. If you need accommodations or adjustments during interviews or beyond, please let us know so we can arrange the support you need.
AI Disclosure
We use AI to take notes during our interviews. Applications and interviews are reviewed by our Talent Acquisition team. Our applicant tracking system uses AI for workflows and hiring process efficiencies.
$110k - $120k per year
...professional development support, discounts through Perkopolis, and recognition programs that celebrate your impact The Job: Site Reliability Engineer The Site Reliability Engineer is responsible for ensuring the availability, performance, and resilience of the...SuggestedFull timeTemporary workInternshipWork at officeRemote work$100k - $125k per year
...We are seeking an experienced and motivated Software Engineer to join our dynamic Site Reliability Engineering (SRE) team. As a Site Reliability Engineer,... ...culture Tech at Tipalti Our tech teams are the engine behind our business. Tipalti’s tech ecosystem is extremely...SuggestedFull timeWork at officeFlexible hours$110k - $130k per year
...of metasearch brands. Hospitality is all about taking care of others, and it defines our culture. About the job As a Site Reliability Engineer II on the Serving Platforms team within Infrastructure Engineering, you will design, automate, and manage the core container...SuggestedFull timeWork at officeLocal areaWorldwideFlexible hours2 days per week$140k - $155k per year
...profiles! This is a hands-on senior engineering role focused on improving production... ...enable engineering teams to ship secure, reliable, and scalable software with confidence. You... ...engineers on cloud-native technologies, site reliability engineering principles, and operational...SuggestedRemote jobPermanent employmentFull timeFlexible hours$140k - $182k per year
...client base with operations throughout North America, Central America, Europe, Australia, and Japan. As one of our Senior Site Reliability Engineers, you will be 100% hands-on across infrastructure and software development. You will support and help inform the evolution of...SuggestedFull time$153.82k - $277k per year
...place where you can thrive, we can’t wait to meet you. Site Reliability Engineers (SREs) at Braze are responsible for keeping all internal-facing... ...level feature requirements and helping translate them into reliable, highly scalable technology stacks. You will be responsible...Permanent employmentFull timeInternshipWork at officeLocal areaRemote workFlexible hoursRotating shift$154k - $200k per year
...client base with operations throughout North America, Central America, Europe, Australia, and Japan. As one of our Lead Site Reliability Engineers, you will combine hands-on technical expertise with strategic technical leadership across infrastructure and software development...Long term contractFull time- ...competitive advantage. Job Description The SRE Role · SREs are engineers with the right mix of knowledge and skills in software... ...experimentation and observation to entire systems to improve reliability, performance and operability). · We constantly evaluate products...Full time
- ...of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a systems thinker. You’ll create middleware and platform guardrails...Full timeWork at officeLocal areaRemote workWorldwideMonday to fridayFlexible hours
$80 - $110 per hour
...Networks is headquartered in Irvine, CA USA with Asia HQ in Singapore and also operating in Denmark, Spain and Vietnam. The Site Reliability Engineer will improve the availability, performance, scalability and recoverability of AXON Networks cloud solutions. You will...Remote jobFull timeContract work- ...customers. Cohere is a team of researchers, engineers, designers, and more, who are passionate... ...building high-performance, scalable and reliable machine learning systems? Do you want to... ...NLP applications? We are looking for a Site Reliability Engineer to join the Model...Full timeWork at officeRemote workFlexible hours
$144k - $200k per year
**The Team** Platform Engineering is the department within SRE that is responsible for a range... ...role in developing and maintaining the reliable and globally connected multi-cloud network... ...Overview** We are seeking a talented Site Reliability Engineer (SRE) with a strong...Full timeWork at officeRemote workWorldwideFlexible hours- ...visibility, and optimize spend across the enterprise. The Site Reliability Engineer III (SRE III) plays a critical role in ensuring Emburse’s... ...Excellence & Automation Design, develop, and automate reliable cloud infrastructure and platform services. Apply Infrastructure...Full timeManual laborLocal areaFlexible hours
$130k - $180k per year
...Our Platform is growing and we are looking to hire a Senior Site Reliability Engineer (SRE) / Cloud Engineer Our main Cloud Platform is Azure... ...implement, and maintain CI/CD pipelines, enabling rapid and reliable software releases. Automate and optimize our infrastructure...Full timeRemote workVisa sponsorshipWork visaFlexible hours$110k - $125k per year
...systems bringing #BetterGlobalHealth to patients everyday! Apply today and find plenty of reasons to SMILE! The Cloud Site Reliability Engineer (SRE) is responsible for ensuring the reliability, scalability, and performance of production-grade services deployed across...Full timeRemote workFlexible hours- ...Employment Status: Permanent Schedule: 40 hours/week – 100% remote work Job Description We are looking for an experienced Site Reliability Engineer to join a team responsible for the reliability, performance, and resilience of high-availability SaaS platforms. Working...Permanent employmentWork at officeLocal areaRemote work
- ...daily transactions. The organization is focused on building highly available, data-intensive systems and is seeking a Lead Site Reliability Engineer to help shape the future of its infrastructure and reliability strategy. This role combines hands-on technical leadership...Long term contractPermanent employmentFull timeRemote work
- ...Role Description Capgemini is looking for a Production support Engineer to work for the Commercial Line of Business. An ideal candidate... ..., operations teams, and other stakeholders to ensure system reliability and availability. Documentation: Maintain clear and concise...Permanent employmentFull timeLocal area
$120k - $170k per year
...We are seeking a highly skilled and motivated Senior DevOps Engineer to join our dynamic team and play a key role in designing, implementing... ..., investigate incidents, and drive improvements that increase reliability and operational efficiency; Write and maintain automation,...Full timeWork at officeLocal areaFlexible hours- ...digital partner that combines Strategy, Experience & Design, Engineering and Managed Services. We build digital solutions that deliver... ...real fixes. QUALIFICATIONS ~6+ years in SRE, platform reliability or observability engineering, with strong hands-on AWS experience...Full time
$96k - $130k per year
...career growth, you’ll thrive here. Report to an experienced Engineering Manager who can offer you mentorship, autonomy, ownership, and... ...on major product features end-to-end with a focus on quality, reliability, and scalability. Be hands-on with the codebase - participate...Permanent employmentFull timeWork at officeFlexible hours$65k - $80k per year
...growing and we’re looking for a Quality Engineer who is passionate about mastering the best... ...integration and a powerful no-code customization engine, Method enables users to design workflows... ...and Performance Tests to ensure fast and reliable iterative development. ● Create test...Permanent employmentFull timeInternshipWork at officeFlexible hours2 days per week3 days per week- ...AI more natural, capable, and useful. We are looking for a Site Reliability Engineer to help build and operate the infrastructure behind that work... ...storage, scheduling, and the operational tooling that keeps them reliable. This is a hands-on role for someone who enjoys taking...Full timeRemote work
$90k - $130k per year
...workstreams are carried by a single engineer each, and a significant wave... .... We are hiring an intermediate engineer to pair on those projects... ...Delivery Bar: improve the reliability, observability, and maintainability... ...Participate in exciting on-site team retreats for...Long term contractFull timeWork at officeRemote workHome office2 days per week3 days per week$116k - $235.1k per year
...About the Role: Site Reliability Engineering (SRE) at Tubi is not a traditional operations team. We are a software engineering organization that applies a developer's mindset and toolkit to the challenges of building and running large-scale, distributed systems. Our mission...Long term contractRemplacementFull timeTemporary workWork at officeLocal areaImmediate startFlexible hours2 days per week$110k - $130k per year
...of metasearch brands. Hospitality is all about taking care of others, and it defines our culture. About the job As a Site Reliability Engineer II on the Serving Platforms team within Infrastructure Engineering, you will design, automate, and manage the core container...Work at officeLocal areaWorldwideFlexible hours2 days per week- ...Intermediate Associate Engineer, Electrical Toronto, ON 30 Forensic Engineering is one of Canada’s largest and most respected multi-disciplinary... ...subject matter experts to consult with clients, perform site examinations, testing, and analysis, and prepare clear, concise...Full timeWork at office
- ...contributors of all seniorities. Join Tenstorrent as a Staff Reliability Engineer and help define the reliability strategy behind the next... ...influence the future of AI hardware by helping build highly reliable systems that power tomorrow's largest AI workloads. Tenstorrent...Permanent employmentFull timeInternshipSecond job
$78.62 - $89.34 per hour
Our client, is seeking a talented and proactive Site Reliability Engineer (SRE) / Senior Database Platform Engineer to join their core Data Engineering and Operations team. In this engineering-focused role, you will move beyond traditional database administration to act...Long term contractPermanent employmentFull timeContract workWork at office$115k - $125k per year
...(symbol: KGC). Job Summary The Fixed Plant Asset Reliability Engineer, as part of the Asset Management team, plays a vital role in enhancing... ...improvement. This position requires 50%+ travel to remote sites overseas, ensuring that global operations are supported for...Long term contractTemporary workFor contractorsCasual workLocal areaImmediate startRemote workOverseas
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Intermediate Site Reliability Engineer. Be the first to apply!
- senior site reliability engineer Toronto, ON
- site reliability engineer Toronto, ON
- site reliability engineer intern Toronto, ON
- website developer Toronto, ON
- site maintenance Toronto, ON
- site safety Toronto, ON
- site carpenter Toronto, ON
- senior site reliability engineer
- site reliability engineer
- site reliability engineer intern


