Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer, Kong Konnect

Full-time

Kong Company

Are you ready to unlock intelligence?

If you don’t think you meet all of the criteria below but are still interested in the job, please apply. Nobody checks every box - we’re looking for candidates that are particularly strong in a few areas, and have some interest and capabilities in others.

About the Role:

As a Site Reliability Engineer , you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS platform that powers the world’s API connectivity.

You’ll design, automate, and run production systems serving thousands of customers across AWS, GCP, and Azure. You’ll work on everything from multi-region Kubernetes clusters to service mesh and gateway architectures , ensuring the reliability, scalability, and security of Kong’s SaaS offerings.

This is a hands-on role ideal for engineers who thrive on running production SaaS systems at scale , automating operations, and continuously improving performance, resilience, and deployment pipelines.

What You’ll Do:

  • Operate and scale Kong’s global SaaS platform (Konnect) , ensuring reliability, availability, and performance across regions and clouds.

  • Build, automate, and maintain Kubernetes-based infrastructure and deployment workflows using Terraform/Terragrunt , Helm , and ArgoCD .

  • Design, maintain, and optimize multi-region data and caching layers — including PostgreSQL, Redis, ClickHouse, and Druid — for high availability and low latency.

  • Operate and improve Kong Gateway and Kong Mesh environments supporting hybrid and distributed architectures.

  • Develop and maintain CI/CD pipelines and GitOps workflows to automate service delivery and ensure consistent infrastructure changes.

  • Enhance observability and incident response readiness through systems like Datadog, Prometheus, Grafana, and Thanos , defining and tracking SLOs.

  • Collaborate closely with development and security teams to ensure smooth operation of SaaS services in compliance with reliability, security, and regulatory standards.

  • Participate in a global 24/7 on-call rotation and drive continuous improvement of operational playbooks and postmortem practices.

  • Lead and contribute to scaling initiatives that improve elasticity, reliability, and cost-efficiency across the SaaS platform.

What You’ll Bring:

  • BS in Computer Science or equivalent practical experience.

  • Proven experience managing SaaS or PaaS systems at enterprise scale (multi-region, multi-tenant, secure environments).

  • Deep expertise in Kubernetes , including debugging cluster/networking issues and designing for fault tolerance and scalability.

  • Strong proficiency with Infrastructure as Code tools like Terraform or Terragrunt .

  • Experience with CI/CD pipelines and GitOps workflows (ArgoCD, Atlantis, Helm).

  • Proficiency in one or more programming languages ( Go, Python, Bash ) for automation and tooling.

  • Solid understanding of Linux/Unix systems , networking (DNS, TLS/SSL, , load balancers and distributed systems .

  • Experiencing working with API gateway and service mesh technologies

  • Familiarity with streaming systems like Kafka and observability platforms (Datadog, Prometheus, Grafana).

  • Experience working in a 24/7/365 production support environment.

Bonus Points:

  • Hands-on experience with Kong Gateway , Kong Mesh , or similar service connectivity technologies.

  • Experience operating ClickHouse , Druid , or other time-series and analytics databases .

  • Experience managing PostgreSQL and Redis in multi-region configurations.

  • Working knowledge of AWS networking (PrivateLink, Transit Gateway, VPC Peering, Firewalls), Azure VNet , or GCP NCC .

  • Strong understanding of disaster recovery , resiliency testing , and compliance-driven reliability practices.

#LI-KC1

About Kong:

Kong Inc., a leading developer of API and AI connectivity technologies, is building the infrastructure that powers the agentic era. Trusted by the Fortune 500 and startups alike, Kong's unified API and AI platform, Kong Konnect, enables organizations to secure, manage, accelerate, govern, and monetize the flow of intelligence across APIs and AI models. For more information, visit .

Vacancy posted 16 hours ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer, Kong Konnect in Toronto, ON vacancy
  • $153.82k - $277k per year

     ...a place where you can thrive, we can’t wait to meet you. Site Reliability Engineers (SREs) at Braze are responsible for keeping all internal-facing...  ...of messages to our customers' end-users daily. This Senior SRE role is specifically focused on supporting the Braze Ruby... 
    Senior
    Permanent employment
    Full time
    Internship
    Work at office
    Local area
    Remote work
    Flexible hours
    Rotating shift

    Braze

    Toronto, ON
    8 hours ago
  •  ...Docebians around the world and help us reinvent the way people learn, because learning never stops. Role Overview As a Senior Site Reliability Engineer, you'll take a hands-on lead role in high severity incident response while also shaping the underlying infrastructure... 
    Senior
    Full time
    For contractors
    Work at office
    Worldwide
    3 days per week

    Docebo

    Toronto, ON
    16 hours ago
  • $140k - $182k per year

     ...global client base with operations throughout North America, Central America, Europe, Australia, and Japan. As one of our Senior Site Reliability Engineers, you will be 100% hands-on across infrastructure and software development. You will support and help inform the... 
    Senior
    Full time

    Movable Ink

    Toronto, ON
    16 hours ago
  •  ...best of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage Means…  You are an engineer, a builder, and a systems thinker. You’ll create middleware and platform... 
    Senior
    Full time
    Work at office
    Local area
    Remote work
    Worldwide
    Monday to friday
    Flexible hours

    Imanage

    Toronto, ON
    16 hours ago
  •  ...capabilities in others. About the role: The Senior Software Engineer will be a core member of the Billing Platform...  ...the microservices and integrations that power Kong's commercial infrastructure. You'll build scalable, reliable TypeScript services that handle the full... 
    Senior
    Full time
    Contract work

    Kong Company

    Toronto, ON
    16 hours ago
  • $243k - $297k per year

     ...businesses. As Relay continues to scale, the reliability, performance, and resilience of our...  ...and business success. This is a senior leadership role responsible not only for guiding a strong team of Site Reliability Engineers, but for shaping how reliability strategy... 
    Senior
    Long term contract
    Full time
    Internship
    Work at office
    Trial period
    Flexible hours

    Relay

    Toronto, ON
    16 hours ago
  • $130k - $180k per year

     ...to legally work in Canada (visa or sponsorship won't be provided) Our Platform is growing and we are looking to hire a Senior Site Reliability Engineer (SRE) / Cloud Engineer Our main Cloud Platform is Azure (those with Azure will be prioritized first) About Us:... 
    Senior
    Full time
    Remote work
    Visa sponsorship
    Work visa
    Flexible hours

    Acquird.io

    Toronto, ON
    16 hours ago
  • $110k - $120k per year

     ...recognition programs that celebrate your impact The Job: Site Reliability Engineer The Site Reliability Engineer is responsible for ensuring...  ...reporting, SLO performance metrics, and incident trends to senior management. What You Bring : Technical Proficiency:... 
    Senior
    Full time
    Temporary work
    Internship
    Work at office
    Remote work

    Momentum Financial Services Group

    Toronto, ON
    16 hours ago
  • $100k - $125k per year

     ...We are seeking an experienced and motivated Software Engineer to join our dynamic Site Reliability Engineering (SRE) team. As a Site Reliability Engineer,...  ...culture Tech at Tipalti  Our tech teams are the engine behind our business. Tipalti’s tech ecosystem is extremely... 
    Full time
    Work at office
    Flexible hours

    Tipalti

    Toronto, ON
    8 hours ago
  • $140k - $180k per year

     ...infrastructure that all other engineering teams at Tripstack depend on....  ...the deployment patterns and reliability standards that apply across both...  ...reliability trade-offs to senior stakeholders directly Coordinate...  ...depth - interconnects, BGP, site-to-site VPN, cross-region... 
    Senior
    Remplacement
    Full time
    Work at office
    Immediate start
    Flexible hours

    Tripstack

    Toronto, ON
    16 hours ago
  •  ...Senior Site Reliability Engineer - Edge Location : Ottawa/Toronto, On-Site Reports to: Head of Security The Role We build rugged edge-compute devices for extreme environments — the small, hardened computers that run alongside our mesh radios and sensors in the... 
    Senior
    Full time

    dominion%20dynamics

    Toronto, ON
    13 hours ago
  • $140k - $155k per year

     ...LinkedIn profiles! This is a hands-on senior engineering role focused on improving production...  ...enable engineering teams to ship secure, reliable, and scalable software with confidence....  ...engineers on cloud-native technologies, site reliability engineering principles, and... 
    Senior
    Remote job
    Permanent employment
    Full time
    Flexible hours

    Caseware

    Toronto, ON
    16 hours ago
  •  ...competitive advantage. Job Description The SRE Role · SREs are engineers with the right mix of knowledge and skills in software...  ...experimentation and observation to entire systems to improve reliability, performance and operability). · We constantly evaluate products... 
    Full time

    Serigor Inc

    Toronto, ON
    16 hours ago
  •  ...learning never stops. The Adventure Ahead As the Manager of Site Reliability Engineering (SRE), you will lead a talented team of engineers dedicated...  ...Product, Engineering, and Support leadership to integrate reliable SRE practices into early planning and product delivery... 
    Long term contract
    Full time
    For contractors
    Work at office
    Worldwide
    3 days per week

    Docebo

    Toronto, ON
    16 hours ago
  • $154k - $200k per year

     ...client base with operations throughout North America, Central America, Europe, Australia, and Japan. As one of our Lead Site Reliability Engineers, you will combine hands-on technical expertise with strategic technical leadership across infrastructure and software development... 
    Long term contract
    Full time

    Movable Ink

    Toronto, ON
    16 hours ago
  •  ...the API Management group at Kong, you will build tools and services...  ...APIs. Working closely with Engineering and Product teams, you will...  ...correctness and operational reliability. - Monitor service health and...  ...from and seek guidance from senior engineers to grow your technical... 
    Senior
    Full time

    Kong Company

    Toronto, ON
    16 hours ago
  •  ...AI more natural, capable, and useful. We are looking for a Site Reliability Engineer to help build and operate the infrastructure behind that work...  ...storage, scheduling, and the operational tooling that keeps them reliable. This is a hands-on role for someone who enjoys taking... 
    Full time
    Remote work

    bosonai

    Toronto, ON
    3 days ago
  •  ...Employment Status: Permanent Schedule: 40 hours/week – 100% remote work Job Description We are looking for an experienced Site Reliability Engineer to join a team responsible for the reliability, performance, and resilience of high-availability SaaS platforms. Working... 
    Permanent employment
    Work at office
    Local area
    Remote work

    TOTEM Recruteur de talent

    Toronto, ON
    7 days ago
  • $144k - $200k per year

    **The Team** Platform Engineering is the department within SRE that is responsible for a range...  ...role in developing and maintaining the reliable and globally connected multi-cloud network...  ...Overview** We are seeking a talented Site Reliability Engineer (SRE) with a strong... 
    Full time
    Work at office
    Remote work
    Worldwide
    Flexible hours

    Mongodb

    Toronto, ON
    16 hours ago
  •  ...visibility, and optimize spend across the enterprise. The Site Reliability Engineer III (SRE III) plays a critical role in ensuring Emburse’s...  ...Excellence & Automation Design, develop, and automate reliable cloud infrastructure and platform services. Apply Infrastructure... 
    Full time
    Manual labor
    Local area
    Flexible hours

    Emburse

    Toronto, ON
    16 hours ago
  •  ...customers. Cohere is a team of researchers, engineers, designers, and more, who are passionate...  ...building high-performance, scalable and reliable machine learning systems? Do you want to...  ...NLP applications? We are looking for a Site Reliability Engineer to join the Model... 
    Full time
    Work at office
    Remote work
    Flexible hours

    Cohere

    Toronto, ON
    16 hours ago
  • $110k - $125k per year

     ...systems bringing #BetterGlobalHealth to patients everyday! Apply today and find plenty of reasons to SMILE! The Cloud Site Reliability Engineer (SRE) is responsible for ensuring the reliability, scalability, and performance of production-grade services deployed across... 
    Full time
    Remote work
    Flexible hours

    Smile Digital Health

    Toronto, ON
    16 hours ago
  • $120k - $170k per year

     ...Overview  We are seeking a highly skilled and motivated Senior DevOps Engineer to join our dynamic team and play a key role in designing,...  ...investigate incidents, and drive improvements that increase reliability and operational efficiency; Write and maintain automation... 
    Senior
    Full time
    Work at office
    Local area
    Flexible hours

    Magnet Forensics

    Toronto, ON
    16 hours ago
  • $150k - $250k per year

     ...About The Role We're looking for a Senior Site Reliability Engineer to help us run one of the most exciting GPU clusters around—our Toronto datacenter packed with NVIDIA H100 and A100 GPUs, over 20PB of Ceph storage, terabit networking, and hundreds of servers. You'll... 
    Senior
    Full time

    Boson Ai

    Toronto, ON
    16 hours ago
  •  ...SLOs, shape capacity plans, and ensure the reliability, durability, and operational safety of...  ...that underpins Atlas. You’ll join a small, senior team of SREs as founding members of this...  ...processes. We are a small team of software engineers with a strong bias towards software... 
    Senior
    Long term contract
    Full time
    Work at office
    Local area
    Immediate start
    Remote work
    Worldwide
    Shift work

    Mongodb

    Toronto, ON
    16 hours ago
  •  ...PostgreSQL) This role is part of the Production Support and Reliability Engineering team, responsible for ensuring the stability, availability,...  ...and licenses, Relevant experience and skills, Seniority and performance, Market and business consideration, Internal... 
    Permanent employment
    Full time
    Local area

    Capgemini

    Toronto, ON
    19 days ago
  • $78.62 - $89.34 per hour

    Our client, is seeking a talented and proactive Site Reliability Engineer (SRE) / Senior Database Platform Engineer to join their core Data Engineering and Operations team. In this engineering-focused role, you will move beyond traditional database administration to act as... 
    Senior
    Long term contract
    Permanent employment
    Full time
    Contract work
    Work at office

    Randstad

    Toronto, ON
    a month ago
  • $115k - $125k per year

     ...Founded in 1993, Kinross is a Canadian-based senior gold mining company with operations and...  ...symbol: KGC). Job Summary The Reliability Engineer, as part of the Asset Management team,...  ...position requires 50%+ travel to remote sites overseas, ensuring that global... 
    Senior
    Long term contract
    Temporary work
    For contractors
    Casual work
    Local area
    Immediate start
    Remote work
    Overseas

    Kinross Gold Corporation

    Toronto, ON
    5 hours ago
  • $82.2 - $89.34 per hour

    We are seeking a highly skilled Site Reliability Specialist IV for an enterprise-level contract opportunity based in Toronto. In this role, you will take on a premier cloud engineering, platform automation, and operational reliability capacity, specializing in designing, building... 
    Long term contract
    Permanent employment
    Contract work

    Randstad

    Toronto, ON
    1 day ago
  •  ...role: In the API Management group at Kong, you will build tools & services that enable...  ...documentation. Working closely with Engineering and Product teams, you will develop API-focused...  ...Kong's unified API and AI platform, Kong Konnect, enables organizations to secure, manage,... 
    Senior
    Full time

    Kong Company

    Toronto, ON
    16 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer, Kong Konnect. Be the first to apply!