Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Site Reliability Engineer

$140k - $155k per year
Full-time

Caseware

Toronto, ON
  • Remote job

Caseware is one of Canada's original Fintech companies, having led the global audit and accounting software industry for over 30 years, with more than 500,000 users across 130 countries and available in 16 different languages. While you might not have heard of us (yet) over 36,000 accounting and audit professionals list Caseware as a skill on their LinkedIn profiles!



This is a hands-on senior engineering role focused on improving production resilience, strengthening security, driving operational excellence, and enhancing the developer experience across the organization.

In this role, you will design, build, and evolve the foundational systems, tooling, and operational practices that enable engineering teams to ship secure, reliable, and scalable software with confidence. You will help establish reliability standards, define service level objectives (SLOs), improve observability, automate operational processes, and drive incident management and post-incident learning practices that strengthen platform stability over time.

Partnering closely with Engineering, Security, Platform, and Product teams, you will architect scalable distributed systems, optimize Kubernetes and AWS-based infrastructure, and build automated delivery pipelines that support rapid and safe software releases. You will play a key role in reducing operational toil, improving system performance, increasing platform reliability, and ensuring that our infrastructure can support continued business growth.

 


❗ This is a full-time permanent position  

❗ This is an existing vacancy 


  Location:   This is a remote location open to candidates legally authorized to work in Canada.

What you will be doing:

 



  • Drive reliability engineering initiatives and operational excellence for mission-critical services running on AWS and Kubernetes.

  • Design, implement, and continuously improve deployment, release, and rollback strategies across complex distributed systems.

  • Establish secure-by-default CI/CD pipelines with robust automation, governance, and policy-driven controls.

  • Enhance platform observability through metrics, logs, tracing, and actionable alerting to improve system visibility and operational efficiency.

  • Define, implement, and mature Service Level Indicators (SLIs), Service Level Objectives (SLOs), and reliability standards across the organization.

  • Lead response efforts for high-severity incidents, ensuring timely resolution, effective communication, and meaningful post-incident reviews that drive continuous improvement.

  • Partner closely with engineering teams to strengthen platform standards, improve service resilience, optimize runtime performance, and embed reliability best practices.

  • Mentor and guide engineers on cloud-native technologies, site reliability engineering principles, and operational excellence practices, fostering a culture of continuous learning and accountability.

 

What you will bring:


  • 8+ years of experience in Site Reliability Engineering (SRE), Platform Engineering, DevOps, or related cloud-native engineering roles.

  • Deep expertise in AWS services, including EKS, IAM, VPC, Lambda, CloudFront, S3, and cloud networking/security best practices.

  • Advanced experience operating and scaling production Kubernetes environments.

  • Strong hands-on experience with Istio service mesh, including traffic management, security, observability, and resiliency.

  • Proven expertise with Infrastructure as Code (IaC), preferably using AWS CDK.

  • Experience building and managing CI/CD pipelines using GitHub Actions or similar platforms.

  • Strong troubleshooting, performance optimization, and incident management experience in distributed systems.

  • Excellent communication, collaboration, and technical leadership skills.

Observability & Reliability



  • Experience designing and operating monitoring, logging, tracing, and alerting solutions for cloud-native platforms.

  • Strong knowledge of AWS CloudWatch, OpenTelemetry, AWS X-Ray, and Kubernetes observability tooling.

  • Experience defining and operationalizing SLIs, SLOs, alerting strategies, runbooks, and reliability metrics.

  • Proven ability to leverage observability data to improve service reliability, reduce incident impact, and optimize operational performance.

Software Engineering & Platform Development



  • Strong proficiency in TypeScript and Node.js for platform engineering, automation, and operational tooling.

  • Experience building and maintaining scalable backend services, APIs, and event-driven systems.

  • Deep understanding of Kubernetes architecture, controllers, Gateway API, ingress management, and service networking.

  • Experience implementing zero-trust architectures, mTLS, and service-to-service security controls.

  • Commitment to high-quality engineering practices, including automated testing, code reviews, and observability-driven development.

  • Strong understanding of resilience engineering, including autoscaling, disruption management, failure testing, and safe deployment strategies.

Nice to Have



  • Experience with progressive delivery practices such as canary, blue/green, and feature-flag-based deployments.

  • Experience working in regulated, compliance-driven, or security-sensitive SaaS environments.

  • Familiarity with FinOps principles and cost optimization strategies for cloud platforms.

  • Experience building internal developer platforms and self-service engineering tooling.

  • Cloud-native certifications such as CKA, CKAD, CKS, KCSA, or KCNA.

  • Kubestronaut certification or equivalent advanced Kubernetes expertise is highly regarded.



Salary Range:

The annual base salary for this position is between $140,000 CAD and $155,000 CAD per year. 

This role is also eligible for discretionary bonus and/or commission, as well as other benefits. Actual pay within the listed range will be determined based on factors such as transferable skills, relevant experience, market conditions, and primary work location. The posted range is subject to change and may be updated periodically.

 

What's in it for you:

▪️Innovation is at our core. We work with cutting-edge technology in accounting and financial reporting, constantly pushing the boundaries to create impactful software solutions. 

▪️We are committed to a collaborative culture, where your ideas are valued, and knowledge sharing is encouraged within a supportive, inclusive team. 

▪️Work-life balance is important to us. We offer flexible work options, remote opportunities, and generous time-off policies to ensure a healthy work-life balance. 

▪️We offer competitive compensation, including a competitive salary and comprehensive benefits such as health insurance and retirement plans. 

▪️We are driven by impactful work . Your contributions directly affect how our clients manage financial processes and drive their success. 

▪️Recognition and rewards matter to us . We celebrate hard work through recognition programs, performance bonuses, and opportunities for career growth. 

▪️We embrace global opportunities . Work on international projects and collaborate with a diverse, global team. 

About Caseware:

Caseware's cutting-edge software products are meticulously designed for accounting firms, corporations, and governments. Our teams are continually collaborating, innovating, and building upon our existing suite of products. With a customer-focused mindset, we are building technology that is shaping what the future of audits, financial reporting, and financial data analytics will look like.

With a recent strategic investment from Hg Capital in 2020, Caseware is now in its next major growth phase as we double down on the people and products that have made Caseware so successful to date.

One of Caseware's core values is Many Voices, One Team and with that in mind, we're dedicated to building teams as diverse as our customers in an equitable and inclusive way. We welcome and encourage candidates of all backgrounds to apply. Should you require accommodations or have any questions at any point during the application or interview process, please e-mail our People Operations team at [email protected] .

AI Usage:

The recruitment process may use AI assisted tools but not for candidate screening or assessment. All final hiring decisions are made by humans to ensure fairness, transparency, and oversight.  

Background Check:

Any candidates successful in obtaining an offer for a position will need to successfully complete a background check through Certn.co which typically includes an Identity Verification and Criminal Record Check. Executives and Senior Managers will undergo a Soft Credit Check as well. Candidates residing in the Netherlands and Germany are excluded from undergoing background checks via Certn.co 

Security and Fraud:

Caseware takes the security of candidates seriously. All legitimate communication from us will come from email addresses ending in @caseware.com and our open positions are always listed on reputable job boards and on our website We will NEVER ask for payment or financial information from you. If you receive an unsolicited job offer, proceed with extreme caution.

We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.

Vacancy posted a month ago
Similar jobs that could be interesting for youBased on the Staff Site Reliability Engineer in Toronto, ON vacancy
  • $144k - $200k per year

    **The Team** Platform Engineering is the department within SRE that is responsible for a range...  ...role in developing and maintaining the reliable and globally connected multi-cloud network...  ...Overview** We are seeking a talented Site Reliability Engineer (SRE) with a strong... 
    Suggested
    Full time
    Work at office
    Remote work
    Worldwide
    Flexible hours

    Mongodb

    Toronto, ON
    more than 2 months ago
  •  ...contributors of all seniorities. Join Tenstorrent as a Staff Reliability Engineer and help define the reliability strategy behind the next generation...  ...the future of AI hardware by helping build highly reliable systems that power tomorrow's largest AI workloads.   Tenstorrent... 
    Suggested
    Permanent employment
    Full time
    Internship
    Second job

    Tenstorrent

    Toronto, ON
    more than 2 months ago
  • $100k - $125k per year

     ...We are seeking an experienced and motivated Software Engineer to join our dynamic Site Reliability Engineering (SRE) team. As a Site Reliability Engineer,...  ...culture Tech at Tipalti  Our tech teams are the engine behind our business. Tipalti’s tech ecosystem is extremely... 
    Suggested
    Full time
    Work at office
    Flexible hours

    Tipalti

    Toronto, ON
    a month ago
  • $110k - $120k per year

     ...professional development support, discounts through Perkopolis, and recognition programs that celebrate your impact The Job: Site Reliability Engineer The Site Reliability Engineer is responsible for ensuring the availability, performance, and resilience of the... 
    Suggested
    Full time
    Temporary work
    Internship
    Work at office
    Remote work

    Momentum Financial Services Group

    Toronto, ON
    more than 2 months ago
  • $110k - $130k per year

     ...of metasearch brands. Hospitality is all about taking care of others, and it defines our culture. About the job As a Site Reliability Engineer II on the Serving Platforms team within Infrastructure Engineering, you will design, automate, and manage the core container... 
    Suggested
    Full time
    Work at office
    Local area
    Worldwide
    Flexible hours
    2 days per week

    Opentable

    Toronto, ON
    3 days ago
  •  ...of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage Means…  You are an engineer, a builder, and a systems thinker. You’ll create middleware and platform guardrails... 
    Full time
    Work at office
    Local area
    Remote work
    Worldwide
    Monday to friday
    Flexible hours

    Imanage

    Toronto, ON
    more than 2 months ago
  • $154k - $200k per year

     ...client base with operations throughout North America, Central America, Europe, Australia, and Japan. As one of our Lead Site Reliability Engineers, you will combine hands-on technical expertise with strategic technical leadership across infrastructure and software development... 
    Long term contract
    Full time

    Movable Ink

    Toronto, ON
    more than 2 months ago
  • $153.82k - $277k per year

     ...place where you can thrive, we can’t wait to meet you. Site Reliability Engineers (SREs) at Braze are responsible for keeping all internal-facing...  ...level feature requirements and helping translate them into reliable, highly scalable technology stacks. You will be responsible... 
    Permanent employment
    Full time
    Internship
    Work at office
    Local area
    Remote work
    Flexible hours
    Rotating shift

    Braze

    Toronto, ON
    a month ago
  •  ...competitive advantage. Job Description The SRE Role · SREs are engineers with the right mix of knowledge and skills in software...  ...experimentation and observation to entire systems to improve reliability, performance and operability). · We constantly evaluate products... 
    Full time

    Serigor Inc

    Toronto, ON
    more than 2 months ago
  •  ...learning never stops. The Adventure Ahead As the Manager of Site Reliability Engineering (SRE), you will lead a talented team of engineers dedicated...  ...Product, Engineering, and Support leadership to integrate reliable SRE practices into early planning and product delivery... 
    Long term contract
    Full time
    For contractors
    Work at office
    Worldwide
    3 days per week

    Docebo

    Toronto, ON
    more than 2 months ago
  • $140k - $182k per year

     ...client base with operations throughout North America, Central America, Europe, Australia, and Japan. As one of our Senior Site Reliability Engineers, you will be 100% hands-on across infrastructure and software development. You will support and help inform the evolution of... 
    Full time

    Movable Ink

    Toronto, ON
    more than 2 months ago
  • $130k - $180k per year

     ...Our Platform is growing and we are looking to hire a Senior Site Reliability Engineer (SRE) / Cloud Engineer Our main Cloud Platform is Azure...  ...with Equity & Health Benefits ~*Comp range is higher for Staff level Benefits: Meritocracy - Leadership opportunities... 
    Full time
    Remote work
    Visa sponsorship
    Work visa
    Flexible hours

    Acquird.io

    Toronto, ON
    more than 2 months ago
  • $80 - $110 per hour

     ...Networks is headquartered in Irvine, CA USA with Asia HQ in Singapore and also operating in Denmark, Spain and Vietnam. The Site Reliability Engineer  will improve the availability, performance, scalability and recoverability of AXON Networks cloud solutions. You will... 
    Remote job
    Full time
    Contract work

    Axon-networks

    Toronto, ON
    a month ago
  • $110k - $125k per year

     ...systems bringing #BetterGlobalHealth to patients everyday! Apply today and find plenty of reasons to SMILE! The Cloud Site Reliability Engineer (SRE) is responsible for ensuring the reliability, scalability, and performance of production-grade services deployed across... 
    Full time
    Remote work
    Flexible hours

    Smile Digital Health

    Toronto, ON
    more than 2 months ago
  •  ...visibility, and optimize spend across the enterprise. The Site Reliability Engineer III (SRE III) plays a critical role in ensuring Emburse’s...  ...Excellence & Automation Design, develop, and automate reliable cloud infrastructure and platform services. Apply Infrastructure... 
    Full time
    Manual labor
    Local area
    Flexible hours

    Emburse

    Toronto, ON
    more than 2 months ago
  • $172k - $229k per year

     ...BuildOps is looking for a Staff Software Engineer to set and drive our company-wide technical strategy for building, shipping, and operating reliable software. This is a high-impact, cross-functional role for an engineer who sees quality and reliability as properties of... 
    Long term contract
    Permanent employment
    Full time
    For contractors
    Work at office
    Local area
    Work from home
    Flexible hours

    Buildops

    Toronto, ON
    more than 2 months ago
  •  ...customers. Cohere is a team of researchers, engineers, designers, and more, who are passionate...  ...building high-performance, scalable and reliable machine learning systems? Do you want to...  ...NLP applications? We are looking for a Site Reliability Engineer to join the Model... 
    Full time
    Work at office
    Remote work
    Flexible hours

    Cohere

    Toronto, ON
    more than 2 months ago
  • $160k - $220k per year

     ...excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Staff Software Reliability Engineer - Data Platform  About the Team The Data Platform team is responsible for the foundational data services, systems,... 
    Full time
    Local area
    Worldwide
    Flexible hours

    Okta

    Toronto, ON
    more than 2 months ago
  • $121k - $171k per year

     ...Company: Qualcomm Canada ULC Job Area: Engineering Group, Engineering Group Software...  ...and developing test cases which enhance reliability and release quality to internal and external...  ...and Recruiting Agencies :   Our Careers Site is only for individuals seeking a job at... 
    Remplacement
    Full time
    Work from home

    Qualcomm

    Toronto, ON
    11 days ago
  •  ...digital partner that combines Strategy, Experience & Design, Engineering and Managed Services. We build digital solutions that deliver...  ...real fixes. QUALIFICATIONS ~6+ years in SRE, platform reliability or observability engineering, with strong hands-on AWS experience... 
    Full time

    Appnovation Technologies

    Toronto, ON
    2 days ago
  • $120k - $170k per year

     ...We are seeking a highly skilled and motivated Senior DevOps Engineer to join our dynamic team and play a key role in designing, implementing...  ..., investigate incidents, and drive improvements that increase reliability and operational efficiency; Write and maintain automation,... 
    Full time
    Work at office
    Local area
    Flexible hours

    Magnet Forensics

    Toronto, ON
    a month ago
  • $170k - $190k per year

     ...company. We’re looking for a Staff Developer to help shape and build...  ..., office hours, and engineering discussions. This is how we...  ...and user workflows. Own the reliability, scalability, observability,...  ...DevOps experience operating reliable systems in AWS or a comparable... 
    Remote job
    Full time
    Work at office
    Flexible hours

    Fullscript

    Toronto, ON
    29 days ago
  • $216k - $297k per year

     ...come join ours. About this role: Our Engineering organization owns the software that...  ...plumbing underneath. We are hiring a Staff Engineer to own that plumbing. Concretely...  ...buying versus what we should build. Own reliability for the platform: SLOs, on-call,... 
    Full time
    Work at office
    Local area
    Remote work
    Monday to friday
    3 days per week

    Faire

    Toronto, ON
    1 day ago
  • $160k - $220k per year

     .... We are a diverse team of engineers, product managers, and designers...  .... About the role As a Staff Fullstack Engineer on the Okta...  ...the request-time authorization engine at the Agent Gateway and at...  ...observability, documentation, reliability, and support processes. Partner... 
    Full time
    Local area
    Worldwide
    Shift work

    Okta

    Toronto, ON
    9 days ago
  •  ...About The Role Vantage is hiring a Staff Engineer. You help set technical strategy across teams and deliver it through direct implementation, broad technical leadership, or both. As a Staff Engineer, you drive the technical roadmap multiple teams build against, and you'... 
    Remote job
    Long term contract
    Full time
    Contract work
    Temporary work
    Internship
    Work at office
    Local area
    Home office
    Flexible hours

    Ag Vantage

    Toronto, ON
    22 days ago
  •  ...the Forbes World's Best Banks list since 2021.  As a  Staff Cloud Engineer , you will serve as a senior technical expert supporting...  ...~8+ years of experience in Cloud Engineering, DevOps, or Site Reliability Engineering roles in enterprise environments.   ~ Advanced... 
    Hourly pay
    Permanent employment
    Full time
    Work at office

    Eq Bank

    Toronto, ON
    more than 2 months ago
  •  ...of our journey. The Role As Staff Engineer, you will own the technical strategy across...  ....   ~ Engineering quality and reliability — Set the bar for engineering...  ...’s platform is secure, performant, and reliable. Own the response to the most critical... 
    Full time

    Enable

    Toronto, ON
    more than 2 months ago
  •  ...the Role We're looking for a Senior/Staff DevOps Engineer who has spent the last several years building...  ...lets AI and industrial systems run reliably at scale. You understand what it takes...  ...~6+ years of experience in DevOps, Site Reliability Engineering, Platform Engineering... 
    Full time

    Nexxa.ai

    Toronto, ON
    a month ago
  •  ...Employment Status: Permanent Schedule: 40 hours/week – 100% remote work Job Description We are looking for an experienced Site Reliability Engineer to join a team responsible for the reliability, performance, and resilience of high-availability SaaS platforms. Working... 
    Permanent employment
    Work at office
    Local area
    Remote work

    TOTEM Recruteur de talent

    Toronto, ON
    4 days ago
  • $220k - $359.21k per year

     ...WHAT YOU’LL DO We’re looking for a Staff Software Engineer to join the User Targeting team as a hands...  ...end to end, from the segmentation engine and data pipelines through the APIs and...  ...improving their performance, reliability, and cost-effectiveness Mentor & Influence... 
    Full time
    Part time
    Internship
    Manual labor
    Work at office
    Local area
    Flexible hours

    Braze

    Toronto, ON
    8 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Site Reliability Engineer. Be the first to apply!