Site Reliability Engineer
Tecsys Inc.
Having recognized the advantages of remote work, including employee morale, productivity, reduced commuting on employee wellbeing and the environment, we are proud to be a digital-first company. The technologies and programs in which we invested have provided a fantastic foundation to this end. Our digital-first work environment, together with our conveniently located offices and collaborative workspaces, provide our team with the freedom and flexibility to work in the way that makes our employees most productive.
About us
Tecsys is a fast-growing innovator offering supply chain solutions to industry leading healthcare systems, hospitals, and pharmacy businesses to distributors, retailers, and 3PLs. We work with industry leaders to transform their supply chains through technology. If you thrive on tackling interesting challenges with continuous learning opportunities, then Tescys could be a good fit for you!
About the Role
We are looking for a Site Reliability Engineer to join our Network and Security Operations Center (NOC), a team at the heart of platform reliability for mission-critical SaaS environments. You will help maintain, optimize, and ensure the reliability and performance of the systems that power our cloud infrastructure across AWS and Kubernetes , with a strong focus on automation , observability , and continuous improvement . This role blends reliability engineering with incident command, giving you real ownership over uptime, performance, and innovation. You will be part of a highly skilled team that values creative problem-solving, operational excellence, and continuous improvement through automation and resilience engineering.
Responsibilities
- Collaborate with engineering teams to support services from design through launch, including system design consulting, capacity planning, and launch reviews
- Maintain service reliability post-deployment by monitoring availability, latency, and overall system health
- Identify pain points and drive continuous improvements to enhance scalability, simplicity, and platform resilience
- Own observability by developing and improving monitoring, alerting, dashboards, and defining SLOs/SLIs (Datadog)
- Build and enhance automation, internal tooling, and IaC frameworks (Terraform, CI/CD) to enable scalable and self-healing systems
- Scale systems sustainably through automation and reliability-focused improvements
- Leverage AI tools (e.g., Amazon Kiro) to accelerate execution while validating outputs
- Lead incident response , act as Incident Commander when required, and drive blameless postmortems with long-term fixes
- Implement and maintain logging , monitoring , alerting , and SLA reporting practices
- Create and maintain technical documentation and contribute to SRE best practices
- Partner with platform engineering, deployment teams, and cross-functional stakeholders to support growth and system stability
- Collaborate with internal teams and vendors globally to ensure high performance, availability, and reliability across environments
Qualifications
- Strong relevant experience in Site Reliability, Cloud, or DevOps Engineering in SaaS or large-scale production environments
- Strong experience with AWS (multi-account, VPC, EC2, EKS) and Kubernetes at scale
- Hands-on expertise with Infrastructure as Code and automation tools (Terraform, Ansible, or similar)
- Experience with CI/CD pipelines and release automation (GitLab preferred, Jenkins acceptable)
- Proficiency in monitoring and observability tools (Datadog or equivalent), including metrics, logging, alerting, and dashboards
- Experience designing, deploying, and operating large-scale, distributed systems and multi-vendor platforms
- Solid incident management experience, including on-call rotations , escalations , and postmortems
- Strong scripting skills in Python, Bash, Java, or similar for automation and diagnostics
- Familiarity with AI-assisted engineering tools (e.g., Amazon Kiro) and ability to validate outputs effectively
- Basic knowledge of Java or .NET-based development environments
- Proactive mindset with strong ownership, problem-solving, and knowledge-sharing habits
- Willingness to participate in on-call rotations and occasional travel to the office(less than 10%)
We understand that experience comes in many forms and that careers are not always linear. If you don't meet every requirement in this posting, we still encourage you to apply.
At Tecsys, we are committed to fostering a diverse and inclusive workplace where all employees feel valued, respected, and empowered. We believe that diversity drives innovation and strengthens our ability to deliver exceptional solutions. We welcome and encourage applicants from all backgrounds, experiences, and perspectives to join our team.
Tecsys is an equal opportunity employer. Accommodation is available for applicants selected for an interview.
NB: if you are applying to this position, you must be a Canadian Citizen or a Permanent Resident of Canada, OR , have a valid Canadian work permit.
***
A Note on Our Hiring Process: We do not use AI to automatically screen or reject candidates. However, we do use specific screening questions to prioritize the most relevant applications for human review.
At Tecsys, we welcome the thoughtful use of AI tools to help you prepare your application, for example, to improve clarity, organize your resume, or practice interview responses. However, we ask that all information you provide reflects your real experience, and that any assessments or written submissions represent your own work and thinking.During interviews, we expect candidates to engage without the use of AI tools, scripts, or real-time assistance. Authentic, direct conversation helps us get to know how you think, collaborate, and communicate. AI can support your preparation, but it shouldn’t speak or act on your behalf. We genuinely want to meet you .
$120k - $200k per year
..., and many more. ABOUT THE ROLE At LayerZero, our Site Reliability Engineering (SRE) team is at the intersection of software and systems engineering... ...internal systems to those external users interact with—are reliable, meet the uptime expectations of our users, and...SuggestedFull time$100k - $125k per year
...We are seeking an experienced and motivated Software Engineer to join our dynamic Site Reliability Engineering (SRE) team. As a Site Reliability Engineer,... ...culture Tech at Tipalti Our tech teams are the engine behind our business. Tipalti’s tech ecosystem is...SuggestedFull timeWork at officeFlexible hours$20 per day
...our careers page to see how you can grow with us! As a Site Reliability Engineer at Hiive, you will be responsible for ensuring the... ...performance and system behavior, and ensuring these services are reliable, scalable, and cost-efficient in production. In this role...SuggestedFull timeSummer holidayRelocation- ...like an environment that you believe could work for you then read on to find out more. The role: We’re looking for a Site Reliability Engineer to manage, maintain, improve and provide support on our platform. You will be curious by nature, always looking for ways to...SuggestedFull timeRemote work
$110k - $120k per year
...Employment type: Full-Time, Permanent Reports to: Site Reliability Lead, Platform Engineering We're looking for Site Reliability Engineers to help... ...reliability practices (logging standards, use of reliable libraries, SLA/SLO goals) Participate in on-call responsibilities...SuggestedPermanent employmentFull time- ...trillion in AUM and 22 global investment banks. For more information, please visit . The Role CMG is looking for a Site Reliability Engineer (SRE) with a strong focus on monitoring, observability, and alerting to ensure the reliability, performance, and...Remote jobFull timeLocal area
- ...environments. We are a modern, IoT-enabled, cloud-based tool for reliability, safety, and operations of physical equipment and... ...valuing the company at $2.5 billion. We’re looking for a Site Reliability Engineer to help advance MaintainX’s reliability, observability, and...Full timeImmediate start
$101.2k - $136.9k per year
...strategy across digital banking, core banking, data platforms and member facing services. We are looking for a highly skilled Site Reliability Engineer (SRE) who will help build, operate and continuously improve the reliability, performance, security and automation of our...Permanent employmentFull timeInternshipWork at officeImmediate startHome officeFlexible hours2 days per week- ...you’re relentless about the high quality, reliability, and security our customers demand. You... ...technical deliverables will reach the entire engineering organization to enable product teams to... ...confidence. Exemplify cloud-native site reliability best practices with a strong...Long term contractFull timeRemote workFlexible hours
$145k - $185k per year
...environments, and the infrastructure underneath that platform is what makes simulation at scale possible. We're hiring a Senior Site Reliability Engineer to help build and operate that infrastructure. This role sits at the core of how we run large-scale, distributed simulation...Remote jobFull time$197.5k - $225k per year
...investors including Silver Lake Waterman, Moody’s, Sequoia Capital, GV and Riverwood Capital. About the Team: As a Senior Site Reliability Engineer, you will be a key technical leader driving the design and optimization of our Kubernetes-based infrastructure and CI/CD...Full time$145.9k per year
...Are you a seasoned production engineering professional who wants to work on a team that goes beyond the status quo? Then Jobber might be the place for you! We're looking for a Senior Site Reliability Engineer to be part of our Product Software Engineering team. Jobber...Long term contractFull timeInternshipWork at officeLocal areaWorldwide$110k - $160k per year
...to join our team working towards this goal, we would love to hear from you! Role Overview We're seeking a Senior Site Reliability Engineer to join our SaaS-Ops team within Shared Services Engineering. The team owns reliability and operational excellence for our highly...Full timeInternshipWork at officeLocal areaFlexible hours$140k - $180k per year
...the world and our goal is preserving uncensored Internet access and online privacy for all. Right now we are looking for a Site Reliability Engineer to help us tame DNS. About the Position Linux system administration and troubleshooting Network configuration and...Full timeDirect hireWork at officeRemote work$260k - $275k per year
...enterprises • Solve complex reliability challenges at scale • Influence architecture and engineering culture at a company level •... ...will focus on creating reusable, reliable, and scalable solutions that abstract... ..., Platform Engineering, or Site Reliability Engineering role,...Full time- ...a part of our journey! About the role We are committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building and leading processes to ensure the reliability,...Full timeLocal areaRemote workHome officeFlexible hours
$153k - $187k per year
...resilient businesses. We’re looking for an incredible Senior Site Reliability Engineer to join our SRE team. We aim to make reliability, security,... ...reinforce one another so that the platform becomes the engine of Relay’s growth. Your love of making high-impact decisions...Full timeInternship- ...corporate culture by MSP Today, G2, and Great Place to Work™, to name a few. About the role We’re looking for a Senior Site Reliability Engineer (SRE) to help strengthen and scale our multi-cloud platform and developer experience. This is a hands-on senior individual...Full timeInternshipRemote workWork from home
- ..., and many more. ABOUT THE ROLE At LayerZero, our Site Reliability Engineering (SRE) team is at the intersection of software and systems engineering... ...systems to those external users interact with — are reliable, meet the uptime expectations of our users, and...Full time
$120k - $160k per year
...ABOUT YOU We are looking for a Site Reliability Engineer (Monetization) who is pragmatic, product-minded, and equally comfortable writing... ...developers ship features and what infrastructure needs to stay reliable - and to bring the reliability lens into design decisions...Long term contractFull time- ...made, and we are not tiptoeing into it. We are rebuilding our engineering culture around a simple belief: AI changes everything. How teams... ...team builds on — and this role is at the center of keeping it reliable, fast, and scalable. As a Staff SRE, you'll own the infrastructure...Remote jobFull timeInternshipWork at officeLocal areaFlexible hoursShift workWeekend work
- ...We are seeking a Senior DevOps & Site Reliability Engineer to own the reliability, scalability, performance, and operational excellence of Medeloop’s platform. This role blends deep DevOps engineering—CI/CD pipelines, infrastructure as code, and cloud architecture—with SRE...Hourly payFull time
- ...for a new opportunity? Or just checking the market? Well… you might just be in the right place! We’re looking for a Staff Site Reliability Engineer to join our Data team in Canada. As a Staff Data SRE, you are the technical backbone of the Data Office's...Full timeWork at officeRemote workFlexible hoursShift work
$123k - $160k per year
...your place here. We are seeking a highly experienced Senior Site Reliability Engineer to own the reliability, performance, and operational... ...Kotlin, Bash, or similar languages, with an emphasis on building reliable automation and tooling. ~ Hands-on experience with modern...Remote jobLong term contractFull timeRelocation- ...une équipe dynamique en tant qu’ Ingénieur·e Fiabilité de Site (Site Reliability Engineer) pour l’un de nos clients. Le Site Reliability Engineering... ...systems and working with high scale scalable and reliable services. Like to work in a fast-moving environment and...Full timeApprenticeshipLocal areaWorldwide
$107k - $161k per year
...to meet with your team for events or meetings. Join Our Team GoDaddy is seeking a highly skilled and motivated Senior Site Reliability Engineer to join our Database Infrastructure team. This role focuses on designing, developing, and deploying automated solutions for...Full timeSecond jobWork at officeLocal areaRemote workWork from home- ...overview of this role You'll join the Dedicated team as a Site Reliability Engineer focused on Environment Automation , where your work will... .... In this role, you'll help keep these environments reliable, scalable, secure, and consistent by treating everything as...Full timeRemote workHome office
- ...a skill on their LinkedIn profiles! This is a hands-on senior engineering role focused on improving production resilience, strengthening... ...operational practices that enable engineering teams to ship secure, reliable, and scalable software with confidence. You will help...Permanent employmentFull timeRemote work
$135k - $203k per year
...world where Identity belongs to you. As a Staff SRE Engineer, you will champion all things pertaining to reliability at Okta on our Customer Identity (CIC) product.... ...if you: Have 6+ years industry experience as a Site Reliability Engineer or adjust disciplines (DevOps/...Long term contractFull timeRemote workFlexible hours$163k - $194k per year
...daily. We run the platforms that every engineering team at Life360 depends on, including AWS... ...Life360 is hiring an AI-Native Site Reliability Engineer — a senior engineer who doesn’... ...infrastructure in the future. As a Senior Site Reliable Engineer II - Infrastructure (AI Native...Full timeSummer workRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
