Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Intermediate Site Reliability Engineer

$130k - $150k per year

ContactMonkey

Hey there! We're ContactMonkey

Our mission is to power measurable employee engagement worldwide, and we're looking for an Intermediate Site Reliability Engineer to join our Engineering team.

 

About the job

We're looking for someone who enjoys running production systems, understands how applications work, and brings practical application security experience.

You'll work closely with our SRE and development teams to maintain our infrastructure, improve deployments, investigate production issues, and address security risks. You'll also contribute to the technical controls that support SOC 2 audits and GDPR compliance.

Our environment includes AWS, Kubernetes on EKS, Terraform, Terragrunt, GitHub Actions, Prometheus, Grafana, and CloudWatch. Our applications use Ruby on Rails, Vue.js, and Node.js, with MySQL, PostgreSQL, and Sidekiq supporting the backend.

You'll take ownership of defined projects and operational improvements, with senior engineers available for guidance and review. There's room to develop deeper expertise in reliability, infrastructure automation, and application security.

Your impact

  • Infrastructure & reliability: Maintain AWS and Kubernetes environments, troubleshoot production issues, and improve availability, performance, and resource usage.
  • Terraform & Terragrunt: Build and maintain infrastructure as code, review plans, manage environment configuration, and address infrastructure drift.
  • Deployments & developer experience: Improve CI/CD pipelines, deployment automation, release checks, and rollback procedures.
  • Monitoring & incidents: Improve monitoring and alerts, join the on-call rotation, and contribute to incident reviews.
  • Application security: Work with developers to assess vulnerabilities, review security risks, and validate fixes.
  • CI/CD security: Maintain code, dependency, secrets, container, and infrastructure scanning.
  • Cloud security: Strengthen IAM, secrets management, network controls, and Kubernetes security.
  • SOC 2 & GDPR: Support technical controls, audit evidence, and personal data protection.
  • Recovery: Test backups and recovery procedures, maintain runbooks, and support production readiness.
  • Collaboration: Participate in code reviews, document changes, and support application, AI, and data engineering teams.

About you

  • Around 3–5 years of experience in SRE, DevOps, platform engineering, cloud operations, or a related engineering role. Equivalent practical experience is welcome.
  • Hands-on experience supporting production workloads in AWS.
  • Experience writing and maintaining Terraform modules and Terragrunt configuration, including reviewing plans and working with remote state.
  • Experience with Docker and Kubernetes, including troubleshooting deployments, services, health checks, and resource limits.
  • A solid understanding of Linux, networking, DNS, and TLS.
  • Ability to write automation in Python, Bash, Ruby, JavaScript, or another suitable language.
  • Experience with Git, pull requests, CI/CD pipelines, and deployment workflows.
  • Experience using logs, metrics, dashboards, and alerts to investigate production problems.
  • Practical application security experience through vulnerability remediation, secure code review, threat modelling, or security tooling.
  • Understanding of common web application risks, including broken access control, injection, authentication weaknesses, and sensitive data exposure.
  • Familiarity with IAM, least privilege, secrets management, and encryption.
  • Working knowledge of SOC 2 controls and GDPR principles relevant to engineering, including access restrictions, data minimization, retention, and deletion.
  • Clear communication skills and good judgment about when to work independently, request a review, or escalate an issue.

How you can stand out

  • Supported Ruby on Rails or Node.js applications in production
  • Worked with MySQL, PostgreSQL, Redis or Valkey, and Sidekiq
  • Experience with GitHub Actions, Argo CD, Helm, Karpenter, or KEDA
  • Built useful monitoring with Prometheus, Grafana, or CloudWatch
  • Supported services across multiple AWS regions
  • Helped remediate penetration-test findings or contributed technical evidence to a SOC 2 audit
  • Participated in backup restoration or disaster recovery exercises.
  • Familiar with securing AI integrations, agent workloads, or MCP services
  • You hold relevant AWS, Kubernetes, Terraform, or security certifications.

Working with the team


You'll work closely with SRE and application engineers, and support AI and data initiatives where they depend on shared infrastructure.

We value people who ask questions, explain their reasoning, and leave systems easier for others to understand. You'll take part in technical discussions and code reviews, share what you learn, and help improve how we operate.

Reducing recurring incidents, unnecessary alerts, and repetitive manual work is part of the role. We want reliable systems and sustainable operations for the people supporting them.

What we bring to the table


100% employer-paid benefits and a Health Spending Account from day one
Work from anywhere in the world for up to four weeks
A stock option plan so you can own a piece of our success
An RRSP Group Savings Plan
A generous vacation package
A personal development budget
One personal day and two volunteering days
Your birthday off
Five health days per year
A downtown Toronto office for hybrid work, with plenty of snacks

 

Compensation and work details

The salary range for this role is $130,000-$150,000 Compensation is based on experience, skills, and our internal compensation framework and equity.

We're happy to discuss compensation throughout the hiring process.

This is a full-time position on our SRE team.

The role includes a shared on-call rotation after onboarding. We'll discuss the schedule, escalation support, and expectations during the interview process.

Who we are


ContactMonkey helps organizations create, send, and measure internal communications directly within Outlook and Gmail.

Our platform brings together email design, employee engagement tools, and analytics so internal communications teams can understand what reaches their people and what gets a response.

As the product grows, we're investing in reliability, security, and tooling that helps our engineering teams deliver changes confidently.

Diversity is our strength
At ContactMonkey, we're building products for diverse organizations, and we need a diverse team to do that. We strongly encourage applications from everyone regardless of race, religion, colour, national origin, gender, sexual orientation, age, marital status, or disability status.

We are committed to an accessible hiring process. If you need accommodations or adjustments during interviews or beyond, please let us know so we can arrange the support you need.

AI Disclosure
We use AI to take notes during our interviews. Applications and interviews are reviewed by our Talent Acquisition team. Our applicant tracking system uses AI for workflows and hiring process efficiencies.

Vacancy posted 10 hours ago
Similar jobs that could be interesting for youBased on the Intermediate Site Reliability Engineer in Toronto, ON vacancy
  • $110k - $120k per year

     ...professional development support, discounts through Perkopolis, and recognition programs that celebrate your impact The Job: Site Reliability Engineer The Site Reliability Engineer is responsible for ensuring the availability, performance, and resilience of the... 
    Suggested
    Full time
    Temporary work
    Internship
    Work at office
    Remote work

    Momentum Financial Services Group

    Toronto, ON
    more than 2 months ago
  • $100k - $125k per year

     ...We are seeking an experienced and motivated Software Engineer to join our dynamic Site Reliability Engineering (SRE) team. As a Site Reliability Engineer,...  ...culture Tech at Tipalti  Our tech teams are the engine behind our business. Tipalti’s tech ecosystem is extremely... 
    Suggested
    Full time
    Work at office
    Flexible hours

    Tipalti

    Toronto, ON
    a month ago
  • $110k - $130k per year

     ...of metasearch brands. Hospitality is all about taking care of others, and it defines our culture. About the job As a Site Reliability Engineer II on the Serving Platforms team within Infrastructure Engineering, you will design, automate, and manage the core container... 
    Suggested
    Full time
    Work at office
    Local area
    Worldwide
    Flexible hours
    2 days per week

    Opentable

    Toronto, ON
    6 days ago
  • $140k - $155k per year

     ...profiles! This is a hands-on senior engineering role focused on improving production...  ...enable engineering teams to ship secure, reliable, and scalable software with confidence. You...  ...engineers on cloud-native technologies, site reliability engineering principles, and operational... 
    Suggested
    Remote job
    Permanent employment
    Full time
    Flexible hours

    Caseware

    Toronto, ON
    a month ago
  • $140k - $182k per year

     ...client base with operations throughout North America, Central America, Europe, Australia, and Japan. As one of our Senior Site Reliability Engineers, you will be 100% hands-on across infrastructure and software development. You will support and help inform the evolution of... 
    Suggested
    Full time

    Movable Ink

    Toronto, ON
    more than 2 months ago
  • $153.82k - $277k per year

     ...place where you can thrive, we can’t wait to meet you. Site Reliability Engineers (SREs) at Braze are responsible for keeping all internal-facing...  ...level feature requirements and helping translate them into reliable, highly scalable technology stacks. You will be responsible... 
    Permanent employment
    Full time
    Internship
    Work at office
    Local area
    Remote work
    Flexible hours
    Rotating shift

    Braze

    Toronto, ON
    a month ago
  • $154k - $200k per year

     ...client base with operations throughout North America, Central America, Europe, Australia, and Japan. As one of our Lead Site Reliability Engineers, you will combine hands-on technical expertise with strategic technical leadership across infrastructure and software development... 
    Long term contract
    Full time

    Movable Ink

    Toronto, ON
    more than 2 months ago
  •  ...competitive advantage. Job Description The SRE Role · SREs are engineers with the right mix of knowledge and skills in software...  ...experimentation and observation to entire systems to improve reliability, performance and operability). · We constantly evaluate products... 
    Full time

    Serigor Inc

    Toronto, ON
    more than 2 months ago
  •  ...of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage Means…  You are an engineer, a builder, and a systems thinker. You’ll create middleware and platform guardrails... 
    Full time
    Work at office
    Local area
    Remote work
    Worldwide
    Monday to friday
    Flexible hours

    Imanage

    Toronto, ON
    more than 2 months ago
  • $80 - $110 per hour

     ...Networks is headquartered in Irvine, CA USA with Asia HQ in Singapore and also operating in Denmark, Spain and Vietnam. The Site Reliability Engineer  will improve the availability, performance, scalability and recoverability of AXON Networks cloud solutions. You will... 
    Remote job
    Full time
    Contract work

    Axon-networks

    Toronto, ON
    a month ago
  •  ...customers. Cohere is a team of researchers, engineers, designers, and more, who are passionate...  ...building high-performance, scalable and reliable machine learning systems? Do you want to...  ...NLP applications? We are looking for a Site Reliability Engineer to join the Model... 
    Full time
    Work at office
    Remote work
    Flexible hours

    Cohere

    Toronto, ON
    more than 2 months ago
  • $144k - $200k per year

    **The Team** Platform Engineering is the department within SRE that is responsible for a range...  ...role in developing and maintaining the reliable and globally connected multi-cloud network...  ...Overview** We are seeking a talented Site Reliability Engineer (SRE) with a strong... 
    Full time
    Work at office
    Remote work
    Worldwide
    Flexible hours

    Mongodb

    Toronto, ON
    more than 2 months ago
  •  ...visibility, and optimize spend across the enterprise. The Site Reliability Engineer III (SRE III) plays a critical role in ensuring Emburse’s...  ...Excellence & Automation Design, develop, and automate reliable cloud infrastructure and platform services. Apply Infrastructure... 
    Full time
    Manual labor
    Local area
    Flexible hours

    Emburse

    Toronto, ON
    more than 2 months ago
  • $130k - $180k per year

     ...Our Platform is growing and we are looking to hire a Senior Site Reliability Engineer (SRE) / Cloud Engineer Our main Cloud Platform is Azure...  ...implement, and maintain CI/CD pipelines, enabling rapid and reliable software releases. Automate and optimize our infrastructure... 
    Full time
    Remote work
    Visa sponsorship
    Work visa
    Flexible hours

    Acquird.io

    Toronto, ON
    more than 2 months ago
  • $110k - $125k per year

     ...systems bringing #BetterGlobalHealth to patients everyday! Apply today and find plenty of reasons to SMILE! The Cloud Site Reliability Engineer (SRE) is responsible for ensuring the reliability, scalability, and performance of production-grade services deployed across... 
    Full time
    Remote work
    Flexible hours

    Smile Digital Health

    Toronto, ON
    more than 2 months ago
  •  ...Employment Status: Permanent Schedule: 40 hours/week – 100% remote work Job Description We are looking for an experienced Site Reliability Engineer to join a team responsible for the reliability, performance, and resilience of high-availability SaaS platforms. Working... 
    Permanent employment
    Work at office
    Local area
    Remote work

    TOTEM Recruteur de talent

    Toronto, ON
    7 days ago
  •  ...daily transactions. The organization is focused on building highly available, data-intensive systems and is seeking a Lead Site Reliability Engineer to help shape the future of its infrastructure and reliability strategy. This role combines hands-on technical leadership... 
    Long term contract
    Permanent employment
    Full time
    Remote work

    GuruLink

    Toronto, ON
    1 day ago
  •  ...Role Description Capgemini is looking for a Production support Engineer to work for the Commercial Line of Business. An ideal candidate...  ..., operations teams, and other stakeholders to ensure system reliability and availability. Documentation: Maintain clear and concise... 
    Permanent employment
    Full time
    Local area

    Capgemini

    Toronto, ON
    15 days ago
  • $120k - $170k per year

     ...We are seeking a highly skilled and motivated Senior DevOps Engineer to join our dynamic team and play a key role in designing, implementing...  ..., investigate incidents, and drive improvements that increase reliability and operational efficiency; Write and maintain automation,... 
    Full time
    Work at office
    Local area
    Flexible hours

    Magnet Forensics

    Toronto, ON
    a month ago
  •  ...digital partner that combines Strategy, Experience & Design, Engineering and Managed Services. We build digital solutions that deliver...  ...real fixes. QUALIFICATIONS ~6+ years in SRE, platform reliability or observability engineering, with strong hands-on AWS experience... 
    Full time

    Appnovation Technologies

    Toronto, ON
    5 days ago
  • $96k - $130k per year

     ...career growth, you’ll thrive here. Report to an experienced Engineering Manager who can offer you mentorship, autonomy, ownership, and...  ...on major product features end-to-end with a focus on quality, reliability, and scalability. Be hands-on with the codebase - participate... 
    Permanent employment
    Full time
    Work at office
    Flexible hours

    Achievers

    Toronto, ON
    more than 2 months ago
  • $65k - $80k per year

     ...growing and we’re looking for a Quality Engineer who is passionate about mastering the best...  ...integration and a powerful no-code customization engine, Method enables users to design workflows...  ...and Performance Tests to ensure fast and reliable iterative development. ● Create test... 
    Permanent employment
    Full time
    Internship
    Work at office
    Flexible hours
    2 days per week
    3 days per week

    Method Integration Inc.

    Toronto, ON
    more than 2 months ago
  •  ...AI more natural, capable, and useful. We are looking for a Site Reliability Engineer to help build and operate the infrastructure behind that work...  ...storage, scheduling, and the operational tooling that keeps them reliable. This is a hands-on role for someone who enjoys taking... 
    Full time
    Remote work

    bosonai

    Toronto, ON
    a month ago
  • $90k - $130k per year

     ...workstreams are carried by a single engineer each, and a significant wave...  .... We are hiring an intermediate engineer to pair on those projects...  ...Delivery Bar: improve the reliability, observability, and maintainability...  ...Participate in exciting on-site team retreats for... 
    Long term contract
    Full time
    Work at office
    Remote work
    Home office
    2 days per week
    3 days per week

    Medme Health

    Toronto, ON
    a month ago
  • $116k - $235.1k per year

     ...About the Role: Site Reliability Engineering (SRE) at Tubi is not a traditional operations team. We are a software engineering organization that applies a developer's mindset and toolkit to the challenges of building and running large-scale, distributed systems. Our mission... 
    Long term contract
    Remplacement
    Full time
    Temporary work
    Work at office
    Local area
    Immediate start
    Flexible hours
    2 days per week

    Tubi - Canada

    Toronto, ON
    22 days ago
  • $110k - $130k per year

     ...of metasearch brands. Hospitality is all about taking care of others, and it defines our culture. About the job As a Site Reliability Engineer II on the Serving Platforms team within Infrastructure Engineering, you will design, automate, and manage the core container... 
    Work at office
    Local area
    Worldwide
    Flexible hours
    2 days per week

    OpenTable

    Toronto, ON
    9 days ago
  •  ...Intermediate Associate Engineer, Electrical Toronto, ON  30 Forensic Engineering is one of Canada’s largest and most respected multi-disciplinary...  ...subject matter experts to consult with clients, perform site examinations, testing, and analysis, and prepare clear, concise... 
    Full time
    Work at office

    30 Forensic Engineering

    Toronto, ON
    more than 2 months ago
  •  ...contributors of all seniorities. Join Tenstorrent as a Staff Reliability Engineer and help define the reliability strategy behind the next...  ...influence the future of AI hardware by helping build highly reliable systems that power tomorrow's largest AI workloads.   Tenstorrent... 
    Permanent employment
    Full time
    Internship
    Second job

    Tenstorrent

    Toronto, ON
    more than 2 months ago
  • $78.62 - $89.34 per hour

    Our client, is seeking a talented and proactive Site Reliability Engineer (SRE) / Senior Database Platform Engineer to join their core Data Engineering and Operations team. In this engineering-focused role, you will move beyond traditional database administration to act... 
    Long term contract
    Permanent employment
    Full time
    Contract work
    Work at office

    Randstad

    Toronto, ON
    more than 2 months ago
  • $115k - $125k per year

     ...(symbol: KGC).   Job Summary   The Fixed Plant Asset Reliability Engineer, as part of the Asset Management team, plays a vital role in enhancing...  ...improvement. This position requires 50%+ travel to remote sites overseas, ensuring that global operations are supported for... 
    Long term contract
    Temporary work
    For contractors
    Casual work
    Local area
    Immediate start
    Remote work
    Overseas

    Kinross Gold Corporation

    Toronto, ON
    2 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Intermediate Site Reliability Engineer. Be the first to apply!