Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Tech Lead Site Reliability Engineering

$140k - $180k per year

Etraveli Group

It is a B2B Flights as a Service provider and a world leader in virtual interlining.
Operating from offices in Canada, India, and Poland, Tripstack is the gateway into Etraveli Group’s world leading tech platform - giving partners access to global flight content, virtual interlining, and a full suite of services including payments, fraud prevention, pricing, and customer support. As a world leader in virtual interlining technology, Tripstack connects non-partner, low-cost and full-service carriers, enabling the creation of unique and flexible itineraries through a simple, cost-effective API.

Its technology ingest over 30B price points and handles over 240 million searches daily.
Through partnerships with airlines, OTAs, and other distribution channels across the globe, Tripstack expands networks, drives new revenue streams, and offers more choice at competitive prices, all backed by robust technology and traveler protection.

Our SRE team runs the shared infrastructure that all other engineering teams at Tripstack depend on. This spans three Kubernetes environments today - GKE on GCP as the primary platform for booking and search workloads, a multi-tenant Kubernetes-on-OpenStack cluster running Talos Linux for content acquisition and SRE tooling, and an isolated PCI-compliant environment for payment processing - plus a substantial VM footprint, self-hosted Concourse CI, and a Prometheus, Thanos, and Grafana observability stack.
Our Toronto data centre closes in September 2026, and the first phase of migrating the Tripstack Platform to a new OpenStack and Talos-based Kubernetes environment in Gothenburg, Sweden - built in partnership with our parent company Etraveli Group under a shared responsibility model - is in progress. The most immediate priority for this role is moving our GCP footprint from the US to the EU by the end of 2026 with a zero downtime objective.
First, lead the Toronto SRE team day-to-day: on-call, incident management, delivery, and developing engineers who can communicate risk and progress clearly to leadership. Second, shape the overall strategy for application deployments across GCP and Gothenburg - which workloads move, which stay, where the PII boundary sits, and the deployment patterns and reliability standards that apply across both.
This is a hands-on leadership role on a distributed team spanning Toronto, Pune, and Kraków, working closely with engineering counterparts in Stockholm and Gothenburg.
Provide day-to-day technical and people leadership for SRE engineers based in Canada: priorities, delivery, code and change review, career development
Own on-call and incident management for shared infrastructure: sustainable rotations, current runbooks, and post-incident reviews that result in concrete improvements
Develop engineers who communicate well with leadership: your team should be able to present migration risks and reliability trade-offs to senior stakeholders directly
Coordinate closely with SRE and engineering colleagues in Pune and Kraków so that ownership and handoffs across time zones are clearly defined
Shape deployment strategy across GCP and Gothenburg
Own the GCP US-to-EU region migration end to end: planning, sequencing, cutover, and validation for live production traffic, with completion targeted by the end of 2026
Establish a consistent deployment model across both platforms: Concourse pipelines, Helm, and GitOps patterns that work the same way whether the target is GCP or Gothenburg
Define the standard for how application teams onboard workloads: node pools, namespaces, quotas, network policy, and secrets management, all documented and consistent
Deliver the data centre migration
Manager, SRE and ETG ITOPS counterparts, within the shared responsibility model - Tripstack owning the Kubernetes control plane and everything above the hypervisor, adopting ETG standards below it
Complete phase one, then plan and execute the subsequent migration waves - moving business-critical services from single-homed to fully redundant, with failover scenarios enabled and tested regularly
Own the network architecture of a cross-Atlantic platform: peering between Canada, Sweden, and GCP EU, and latency requirements for critical paths
Decommission legacy infrastructure as part of the migration: legacy Terraform, Puppet-managed VMs, and VM-based tooling that has a Kubernetes-native replacement
SLIs, error budgets, and dashboards for our critical APIs, so that reliability decisions are based on data
Make post-incident follow-through, capacity planning, and change safety standard practice across the engineering organization
Partner with our Security team on infrastructure security: least-privilege access across both identity systems, secrets management through Vault, network segmentation, and a consistent patching and vulnerability-management cadence
Enforce the data-residency requirement that PII stays in the EU region, through both policy and technical controls, with compliance treated as an integral part of migration planning
Ensure infrastructure leaving service is decommissioned securely: credentials rotated, access revoked, and data destruction documented
Proven experience leading an SRE, platform, or infrastructure team, including running on-call rotations, acting as incident commander, managing performance, and developing engineers into senior roles

  • Deep production Kubernetes experience including self-managed or bare-metal clusters: cluster lifecycle, upgrades, networking (CNI, ingress, load balancing), and multi-tenant isolation
  • Strong GCP experience - GKE, IAM, VPC networking, Cloud SQL - ideally including a region or cross-region migration with production traffic
  • Senior-level Infrastructure as Code and GitOps experience - Terraform, Helm, and pipeline-as-code CI/CD; experience migrating away from legacy configuration management (Puppet, Ansible) is directly relevant
  • Security-minded operations - least-privilege access design, secrets management (Vault or equivalent), network policy, and patching discipline, with security treated as a core part of reliability
  • Experience with a major infrastructure transition - a data centre migration, cloud migration, or platform rebuild with production traffic
  • Clear written and verbal English; Additional Experience That Would Be Considered An Asset

OpenStack operations experience - Neutron networking, Cinder/Ceph storage, Octavia load balancing, Keystone identity
Talos Linux or another immutable, API-managed Kubernetes OS in production
Concourse CI or comparable pipelines-as-code platforms at organizational scale
Experience operating PCI-scoped or similarly regulated environments, including change control and audit discipline
Regular use of agentic coding tools (Claude Code, Gemini, or equivalent) in your workflow, with sound judgement about validating AI-generated configuration before production use
Operational exposure to data systems our SRE team supports - Druid, Redpanda, Elasticsearch, Airflow
Network engineering depth - interconnects, BGP, site-to-site VPN, cross-region peering
GDPR data-residency, SOC 2, or ISO 27001 experience - the EU migration makes this increasingly relevant
Exposure to travel, flights, or large-scale search and cache systems
Canada - Toronto Office : 140, 000 - 180,000 CAD / Annual Our pay ranges reflect the minimum and maximum target for new hire pay for the full-time position determined by role, level, and location.The pay range shown is based on our compensation structure in place at the time of posting and may be updated periodically based on business needs. Individual pay is based on additional factors including job-related skills, experience, and relevant education and/or training.
The targeted pay range listed reflects the base pay only and does not include bonus, or other benefits.
We use AI in our hiring process.
At Tripstack, we proudly believe in embracing diversity. This is true for our team, clients, communities and stakeholders. We encourage applications regardless of race, colour, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or any other grounds protected by law.

Vacancy posted 20 hours ago
Similar jobs that could be interesting for youBased on the Tech Lead Site Reliability Engineering in Toronto, ON vacancy
  • $104.24k - $143.3k per year

     ...Join some of the most innovative thinkers in FinTech as we lead the evolution of financial technology. If you are an innovative...  ...at an innovative and growing company. Our Lead Site Reliability Engineers will provide a stable infrastructure platform throughout our... 
    Suggested
    Full time
    Work at office
    Immediate start
    Remote work
    Shift work
    Weekend work
    2 days per week

    SimCorp

    Toronto, ON
    8 days ago
  • $100k - $125k per year

     ...We are seeking an experienced and motivated Software Engineer to join our dynamic Site Reliability Engineering (SRE) team. As a Site Reliability Engineer,...  ...management, tax compliance, and treasury. Tipalti partners with leading financial institutions such as Citi, Wells Fargo, J.P.... 
    Suggested
    Full time
    Work at office
    Flexible hours

    Tipalti

    Toronto, ON
    1 day ago
  •  ...to reimagine what’s possible. Join us and help the world’s leading organizations unlock the value of technology and build a more...  ...PostgreSQL) This role is part of the Production Support and Reliability Engineering team, responsible for ensuring the stability, availability,... 
    Suggested
    Permanent employment
    Full time
    Local area

    Capgemini

    Toronto, ON
    7 hours ago
  •  ...Senior Site Reliability Engineer - Edge Location : Ottawa/Toronto, On-Site Reports to: Head of Security The Role We build rugged...  .... This is a build-from-scratch role. You'll define what "reliable" means for hardware that has to keep working in harsh, sometimes... 
    Suggested
    Full time

    dominion%20dynamics

    Toronto, ON
    1 day ago
  • $110k - $120k per year

     ...behind Money Mart—Canada’s largest non-bank branch network—and a leader in financial solutions for underserved communities. From...  ...recognition programs that celebrate your impact The Job: Site Reliability Engineer The Site Reliability Engineer is responsible for ensuring... 
    Suggested
    Temporary work
    Internship
    Work at office
    Remote work

    Momentum Financial Services Group

    Toronto, ON
    a month ago
  •  ...Platform Engineer – DevOps, Site Reliability Engineering (SRE) & Dynatrace Required Skills • Strong experience as a Platform Engineer with expertise...  ...when required. Observability & Monitoring • Lead the implementation and administration of enterprise... 
    Permanent employment

    Astra North Infoteck Inc.

    Toronto, ON
    27 days ago
  • $141k - $191k per year

     ...Reuters and develop your career. As an SRE Manager, you will lead a team of 10+ engineers, oversee their development and ensure operational excellence. About the Role: In this opportunity as Site Reliability Engineering Manager , you will be responsible for: Team... 
    Work at office
    Local area
    Flexible hours
    2 days per week
    3 days per week

    Thomson Reuters

    Toronto, ON
    more than 2 months ago
  • $78.62 - $89.34 per hour

    Our client, is seeking a talented and proactive Site Reliability Specialist / Senior Cloud Platform Engineer to join their core Cloud Engineering division. In this engineering-focused role, you will move far beyond basic operational support to act as a principal architect... 
    Long term contract
    Permanent employment
    Contract work
    Work at office

    Randstad

    Toronto, ON
    27 days ago
  •  ...Site Reliability Engineer  Location: Toronto, ON Work Model: Hybrid (2 days per week in-person at the Toronto office preferred) Required...  ...Description Dynatrace & AI-Driven Observability • Lead the implementation and optimization of the Dynatrace platform... 
    Contract work
    Work at office
    2 days per week

    Astra North Infoteck Inc.

    Toronto, ON
    a month ago
  •  ...Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance...  .... Join Tenstorrent as a Staff Reliability Engineer and help define the reliability strategy...  ...testing at manufacturing and test partner sites, including in Taiwan.   Tenstorrent... 
    Permanent employment
    Internship
    Second job

    Tenstorrent

    Toronto, ON
    8 days ago
  • $172k - $229k per year

     ...BuildOps is looking for a Staff Software Engineer to set and drive our company-wide technical...  ...strategy for building, shipping, and operating reliable software. This is a high-impact, cross-...  ...of customer-impacting failures and lead cross-team initiatives that address root causes... 
    Long term contract
    Permanent employment
    For contractors
    Work at office
    Local area
    Work from home
    Flexible hours

    BuildOps

    Toronto, ON
    1 hour ago
  • $130k per year

     ...maintain a Canadian Security Clearance (Reliability / Controlled Good or higher). Serco...  ...and other floating structures. The Lead Marine Engineer leads teams of less experienced technical...  ...to make an impact every day across 100+ sites in the areas of Defense, Citizen Services... 
    Full time
    Contract work
    Internship
    Local area
    Immediate start

    Serco North America

    Toronto, ON
    8 days ago
  •  ...Join us and help the world’s leading organizations unlock the value...  ...supporting and managing data engineering pipelines in production...  ...SLAs • Serve as technical lead for AMS support, coordinating...  ...improvements for system stability and reliability" The base compensation... 
    Permanent employment
    Full time
    Internship
    Local area

    Capgemini

    Toronto, ON
    7 hours ago
  • $140.6k - $190.6k per year

     ...Job Description Job Description We’re looking for a Lead Engineer to join the CoCounsel Applications team and play a hands-on technical...  ...~ Drive Execution: Break down complex initiatives, provide reliable estimates, and help the team deliver work from development through... 
    Full time
    Work at office
    Local area
    Flexible hours

    Thomson Reuters

    Toronto, ON
    1 day ago
  • $140.6k - $190.6k per year

     ...AI agent ecosystem. Our Platform Engineering team owns the backend services, cloud...  ...AI-powered experiences to operate reliably, securely, and at scale. As a Lead Platform Engineer , you'll...  ...agent workflows, ensuring scalable, reliable, and maintainable solutions across... 
    Long term contract
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours

    Thomson Reuters

    Toronto, ON
    1 day ago
  • $130k - $150k per year

     ...you develop a professional roadmap that takes your career to new heights. Job Summary  Codal is searching for a  hands-on Engineering Lead to perform various architectural tasks and provide technical leadership. To ensure success, you should have extensive experience... 
    Long term contract
    Work at office

    Codal

    Toronto, ON
    1 hour ago
  • $160k - $220k per year

     ...on this mission. If you are too, let's talk. Staff Software Reliability Engineer - Data Platform  About the Team The Data Platform team is...  ...has a directive from engineering leadership to make OKTA a leader in the use of data and machine learning to improve end-user security... 
    Local area
    Worldwide
    Flexible hours

    Okta

    Toronto, ON
    15 days ago
  •  ...standard project delivery. We are currently looking for a Production Lead Engineer to join our team and support our client as they transition from...  ...production reports, forecasts, and performance metrics for site leadership. Identify opportunities to improve mine... 
    Long term contract
    Permanent employment
    Full time
    Temporary work
    Work at office
    Relocation
    Relocation package

    Peter Lucas Project Management Inc.

    Toronto, ON
    15 days ago
  • $110k - $135k per year

     ..., Canada, Zafin partners with leading financial institutions across...  ...Operations, the Lead Cloud Operations Engineer is a senior-level technical...  ...will be building scalable, reliable, and efficient infrastructure...  ...recovery solutions (Azure Backup, Site Recovery) FinOps and cost... 
    Full time

    Zafin

    Toronto, ON
    6 days ago
  • $140.6k - $190.6k per year

     ...The Lead AI Forward Engineer designs and guides the delivery of AI-powered solutions that reduce operational toil and accelerate technology teams...  ...from prototype to production, ensuring solutions meet reliability, security, and compliance expectations. Define reusable... 
    Full time
    Manual labor
    Work at office
    Local area
    Flexible hours

    Thomson Reuters

    Toronto, ON
    1 day ago
  • $140.6k - $190.6k per year

     ...We are seeking a highly skilled and experienced Lead IAM Engineer to join our team. This role will be responsible for the design, implementation...  ..., and standardization improve team efficiency and reliability. Identity platforms are secure, stable, and highly available... 
    Long term contract
    Remplacement
    Work at office
    Local area
    Flexible hours
    2 days per week
    3 days per week

    Thomson Reuters

    Toronto, ON
    1 day ago
  •  ...build digital experiences that connect people and brands - seamlessly, meaningfully, and for impact. We are looking for a Lead Frontend Engineer to steward the development of a new web application built with Next.js 16 (App Router), React 19, and TypeScript. As a cornerstone... 
    Flexible hours

    Vaimo

    Toronto, ON
    12 hours ago
  • $140k - $175k per year

     ...Lead Software Engineer, AI What if the AI agents you architect empower legal professionals to act with sharper judgment and greater leverage...  ..., tooling, and observability our AI agents need to run reliably at scale, often several efforts at once. We're also in the middle... 
    Full time
    Internship
    Work at office
    Local area
    Flexible hours
    Shift work

    Thomson Reuters

    Toronto, ON
    1 day ago
  •  ...experiencing substantial growth in North America, now made up of over 1,000 engineers, architects and planners across Canada and the USA. This...  ...our expansion to new heights. Job Description Lead Electrical Engineer is a senior-level engineering role based in... 
    Full time
    Local area

    Egis Group

    Toronto, ON
    9 days ago
  •  ...Job Title: Mechanical Engineer – Offshore Reliability Experience: Minimum 12 Years Qualification: Bachelor’s Degree in Mechanical Engineering Industry: Oil & Gas / Refinery (Offshore) Work Location : Saudi Arab Job Description: The Mechanical Engineer... 
    Permanent employment
    Full time

    Hudson Manpower

    Toronto, ON
    4 days ago
  • $135k - $165k per year

     ...Job Title: Lead Supply Chain Simulation Engineer Primary Job Location: Toronto, ON, Canada Location Flexibility: Hybrid, Toronto based Employment...  ...design, build and own Workbench’s warehouse simulation engine, wherein a warehouse design is simulated and evaluated... 
    Permanent employment
    Full time
    Contract work
    Part time
    Internship
    Local area
    Immediate start
    Remote work
    Flexible hours

    Fulfillment IQ

    Toronto, ON
    3 hours ago
  •  ...Job Title: Rotating Engineer – Offshore Reliability Experience: Minimum 12 Years Qualification: Bachelor’s Degree in Mechanical Engineering Industry: Oil & Gas / Refinery (Offshore) Work Location : Saudi Arab Job Description: The Rotating Engineer – Offshore... 
    Permanent employment
    Full time

    Hudson Manpower

    Toronto, ON
    4 days ago
  •  ...Job Title: Rotating Engineer – Onshore Reliability Experience: Minimum 12 Years Qualification: Bachelor’s Degree in Mechanical Engineering...  ...ensure safe and efficient plant operations. # Ensure reliable operation and optimal performance of rotating equipment including... 
    Permanent employment
    Full time

    Hudson Manpower

    Toronto, ON
    4 days ago
  •  ...Job Title: Mechanical Engineer – Onshore Reliability Experience: Minimum 12 Years Qualification: Bachelor’s Degree in Mechanical Engineering...  ..., and optimize maintenance activities to ensure safe, reliable, and efficient plant operations. # Develop and implement... 
    Permanent employment
    Full time

    Hudson Manpower

    Toronto, ON
    4 days ago
  • $140k - $190k per year

     ...Ontario Faster, more frequent, and reliable access to rapid transit with more than...  .... Job Description The MEP Site Superintendent leads and supervises all MEP activities on-site...  .... Liaise with consultants, engineers, and clients regarding technical issues... 
    Full time
    Contract work
    For subcontractor
    Local area

    Ontario Transit Group

    Toronto, ON
    16 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Tech Lead Site Reliability Engineering. Be the first to apply!