Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer II (AI Platform)

$110k - $130k per year
Full-time

Opentable

This hybrid role requires working in the office two days per week.

With millions of diners, 70,000+ restaurant partners and 25+ years of experience, OpenTable, part of Booking Holdings, Inc. (NASDAQ: BKNG), is an industry leader with a passion for helping restaurants thrive. Our world-class technology empowers restaurants to focus on what matters most – their team, their guests, and their bottom line – while enabling diners to discover and book the perfect restaurant for every occasion. 

Every employee at OpenTable has a tangible impact on what we do and how we do it. You’ll also be part of a global team and its portfolio of metasearch brands. Hospitality is all about taking care of others, and it defines our culture.

About the job

As a Site Reliability Engineer II on the Serving Platforms team within Infrastructure Engineering, you will design, automate, and manage the core container stack and infrastructure powering our global business applications. Operating in a high-scale, self-hosted environment, you will serve as a subject matter expert for Kubernetes, Linux systems, and cloud-native automation, directly driving the reliability, security, and efficiency of our platform. In this role, you will collaborate with cross-functional engineering teams worldwide, lead greenfield infrastructure projects, resolve complex incidents, and build self-service capabilities that empower application developers across the organization.

Responsibilities


  • Maintain, tune, and ensure high availability for the low-level Linux operating system and Kubernetes control plane across our self-hosted bare-metal infrastructure.

  • Architect, build, and maintain scalable container management, configuration management, and automation tools across global environments.

  • Investigate, resolve, and conduct root-cause analysis for complex infrastructure disruptions and performance bottlenecks at the system call level.

  • Participate in high-impact platform engineering projects and collaborate with globally distributed engineering teams to drive infrastructure standardization.

  • Participate in the team's on-call rotation to support critical production systems and ensure operational resilience.

  • Develop and maintain self-service tools, automation pipelines, and robust infrastructure monitoring to eliminate manual operational overhead.

Minimum Qualifications


  • 5+ years of hands-on Linux experience (e.g., Ubuntu, CentOS) with expertise in kernel tuning (sysctl), process management (cgroups/namespaces), system calls, and performance optimization.

  • 3+ years of experience using configuration management systems such as Puppet, Chef, Ansible, or SaltStack in production environments.

  • Proven experience building, operating, and troubleshooting bare-metal Kubernetes clusters from the ground up, including control plane, etcd, and CNI plugin management.

  • Proficiency with continuous system automation and scripting in languages such as Go, Python, Ruby, Perl, or Bash.

  • Demonstrated experience responding to live service disruptions, leading root-cause analysis, and operating messaging systems (e.g., Kafka or RabbitMQ) in production.

Preferred Qualifications


  • Experience operating, scaling, and monitoring AI/ML or LLM-powered services and workloads in high-concurrency production environments.

  • Hands-on expertise with public cloud providers (AWS, GCE, or Azure) and containerized CI/CD pipelines (e.g., GitHub, Jenkins, CircleCI, Docker).

  • Experience with distributed key-value stores (e.g., Consul, etcd, Zookeeper, Redis) and enterprise observability/alerting tools (e.g., Prometheus, Sensu).

  • Familiarity with server virtualization infrastructure (e.g., Proxmox, VMware, Xen, OpenStack) and low-level networking concepts (IPtables/NFTables, routing, load balancing).

  • Experience developing and maintaining OS-level software packaging (RPM/DEB) and participating in globally distributed software engineering teams.

This posting is for an existing vacancy.

Benefits and Perks



  • Work from (almost) anywhere for up to 20 days per year

  • Focus on mental health and well-being:


    • Company-paid therapy sessions through SpringHealth

    • Company-paid subscription to Headspace

    • Annual company-wide week off a year - the whole team fully recharges (and returns without a pile-up of work!)

  • Paid parental leave

  • Generous paid vacation + time off for your birthday

  • Paid volunteer time

  • Focus on your career growth:


    • Development Dollars

    • Leadership development

    • Access to thousands of on-demand e-learnings

  • Travel Discounts

  • Employee Resource Groups

  • 20 days of paid time off

  • Private health and dental insurance

  • Life and Disability insurance

The best connections happen face-to-face, whether you’re sitting down to dinner or having coffee with a coworker. That’s why OpenTable has adopted a hybrid workplace model. This role aligns with that approach, with an expectation of coming into the office two days a week—giving employees the best of both worlds: in-person collaboration and flexibility.

The best connections happen face-to-face, whether you’re sitting down to dinner or having coffee with a coworker. That’s why OpenTable has adopted a hybrid workplace model. This role aligns with that approach, with an expectation of coming into the office two days a week—giving employees the best of both worlds: in-person collaboration and flexibility.

The expected range of compensation for this position based in Toronto, Canada, including commission and/or bonuses, is $110,000-$130,000 CAD. There are a variety of factors that go into determining a compensation range, including but not limited to external market benchmark data, geographic location, and years of experience sought/required.

We offer a competitive base salary and benefits including: health benefits; flexible spending account; retirement benefits; life insurance; paid time off (including PTO, paid sick leave, medical leave, bereavement leave, floating holidays and paid holidays); and parental leave benefits. This role is eligible to be considered for an annual bonus and equity grant.

Work Environment & Flexibility

At OpenTable, we pride ourselves on fostering a global and dynamic work environment. As a team member with us, you will benefit from a schedule tailored to accommodate a global workforce operating across multiple time zones. While the majority of your responsibilities may align with conventional business hours, there will be instances where you are expected to manage communications - via calls, Slack messages, or emails - outside of regular working hours to effectively collaborate with international colleagues, respond to restaurant partners, and/or address urgent matters. OpenTable will always abide by and consider local laws and regulations.

Inclusion

We’re committed to creating a workplace where everyone feels they belong and can thrive. We know the best ideas come when we bring different voices to the table, so we're building a team as dynamic as the diners and restaurants we serve—and fostering a culture where everyone feels welcome to be themselves.

If you need accommodations during the application or interview process, or on the job, we’re here to support you. Please reach out to your recruiter to request any accommodations.

#LI-Hybrid

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer II (AI Platform) in Toronto, ON vacancy
  •  ...combines Strategy, Experience & Design, Engineering and Managed Services. We build digital solutions...  ...pod to build and run an internal agent platform on AWS Bedrock AgentCore for a global...  ...~6+ years in SRE, platform reliability or observability engineering, with strong... 
    Suggested
    Full time

    Appnovation Technologies

    Toronto, ON
    1 day ago
  • $110k - $130k per year

     ...of metasearch brands. Hospitality is all about taking care of others, and it defines our culture. About the job As a Site Reliability Engineer II on the Serving Platforms team within Infrastructure Engineering, you will design, automate, and manage the core container stack... 
    Suggested
    Work at office
    Local area
    Worldwide
    Flexible hours
    2 days per week

    OpenTable

    Toronto, ON
    5 days ago
  • $110k - $120k per year

     ...recognition programs that celebrate your impact The Job: Site Reliability Engineer The Site Reliability Engineer is responsible for...  ...the organization's digital banking and financial services platforms. This role focuses on automating operational processes, defining... 
    Suggested
    Full time
    Temporary work
    Internship
    Work at office
    Remote work

    Momentum Financial Services Group

    Toronto, ON
    more than 2 months ago
  • $100k - $125k per year

     ...We are seeking an experienced and motivated Software Engineer to join our dynamic Site Reliability Engineering (SRE) team. As a Site Reliability Engineer,...  ...in real time.      Why join Tipalti? Tipalti is the AI-powered platform for finance automation, elevating how finance... 
    Suggested
    Full time
    Work at office
    Flexible hours

    Tipalti

    Toronto, ON
    a month ago
  • $154k - $200k per year

     ...for marketers through data-activated content generation and AI decisioning. The world’s most innovative brands rely on Movable...  ...America, Europe, Australia, and Japan. As one of our Lead Site Reliability Engineers, you will combine hands-on technical expertise with strategic... 
    Suggested
    Long term contract
    Full time

    Movable Ink

    Toronto, ON
    more than 2 months ago
  •  ...drives business value and inspires extraordinary results.  Our AI-powered platform helps organizations modernize financial operations,...  ...visibility, and optimize spend across the enterprise. The Site Reliability Engineer III (SRE III) plays a critical role in ensuring Emburse... 
    Full time
    Manual labor
    Local area
    Flexible hours

    Emburse

    Toronto, ON
    more than 2 months ago
  • $140k - $155k per year

     ...This is a hands-on senior engineering role focused on improving production...  ...teams to ship secure, reliable, and scalable software with confidence...  ...practices that strengthen platform stability over time. Partnering...  ...on cloud-native technologies, site reliability engineering... 
    Remote job
    Permanent employment
    Full time
    Flexible hours

    Caseware

    Toronto, ON
    a month ago
  • $140k - $182k per year

     ...for marketers through data-activated content generation and AI decisioning. The world’s most innovative brands rely on Movable...  ..., Europe, Australia, and Japan. As one of our Senior Site Reliability Engineers, you will be 100% hands-on across infrastructure and software... 
    Full time

    Movable Ink

    Toronto, ON
    more than 2 months ago
  •  .... Actual Impact. At Docebo, we’re using AI to change how people learn at work—and we...  ...change it. We’re an AI-powered learning platform that helps organizations create, deliver,...  ...The Adventure Ahead As the Manager of Site Reliability Engineering (SRE), you will lead a talented... 
    Long term contract
    Full time
    For contractors
    Work at office
    Worldwide
    3 days per week

    Docebo

    Toronto, ON
    more than 2 months ago
  • $153.82k - $277k per year

     ...If Braze sounds like a place where you can thrive, we can’t wait to meet you. Site Reliability Engineers (SREs) at Braze are responsible for keeping all internal-facing services and platforms running smoothly, ensuring consistent system reliability and maximum... 
    Permanent employment
    Full time
    Internship
    Work at office
    Local area
    Remote work
    Flexible hours
    Rotating shift

    Braze

    Toronto, ON
    a month ago
  •  ...about belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage Means…  You are an engineer, a builder, and a systems thinker. You’ll create middleware and platform guardrails that empower developers to innovate quickly and reliably... 
    Full time
    Work at office
    Local area
    Remote work
    Worldwide
    Monday to friday
    Flexible hours

    Imanage

    Toronto, ON
    more than 2 months ago
  • $80 - $110 per hour

     ...AXON Networks delivers a robust AI-driven, analytics-based orchestration platform and a wide portfolio of next-gen high-speed routers that leverage...  ...also operating in Denmark, Spain and Vietnam. The Site Reliability Engineer  will improve the availability, performance... 
    Remote job
    Full time
    Contract work

    Axon-networks

    Toronto, ON
    a month ago
  •  ...Description The SRE Role · SREs are engineers with the right mix of knowledge and...  ...observation to entire systems to improve reliability, performance and operability). · We constantly...  ...or ansible. · Cloud technologies and platforms such as AWS or Azure using API or... 
    Full time

    Serigor Inc

    Toronto, ON
    more than 2 months ago
  • $110k - $125k per year

     ...results can be seen in the impact of our innovative health data platform and data management solutions, which are used in over 20...  ...Apply today and find plenty of reasons to SMILE! The Cloud Site Reliability Engineer (SRE) is responsible for ensuring the reliability,... 
    Full time
    Remote work
    Flexible hours

    Smile Digital Health

    Toronto, ON
    more than 2 months ago
  •  ...enterprises who are building AI systems to power magical experiences...  ...is a team of researchers, engineers, designers, and more, who are...  ...high-performance, scalable and reliable machine learning systems? Do...  ...build the next generation of AI platforms powering advanced NLP... 
    Full time
    Work at office
    Remote work
    Flexible hours

    Cohere

    Toronto, ON
    more than 2 months ago
  • $144k - $200k per year

    **The Team** Platform Engineering is the department within SRE that is responsible for a range of critical...  ...role in developing and maintaining the reliable and globally connected multi-cloud...  ...Overview** We are seeking a talented Site Reliability Engineer (SRE) with a strong... 
    Full time
    Work at office
    Remote work
    Worldwide
    Flexible hours

    Mongodb

    Toronto, ON
    more than 2 months ago
  • $170k - $185k per year

     ...s top-performing asset managers? We are seeking a Head of Platform Engineering, Reliability & Control to lead our horizontal engineering function and...  ...~ Drive enterprise-wide automation by implementing AI-enabled monitoring and process orchestration capabilities that... 
    Full time
    Work at office
    3 days per week

    Connor, Clark

    Toronto, ON
    more than 2 months ago
  • $130k - $180k per year

     ...month). Must be able to legally work in Canada (visa or sponsorship won't be provided) Our Platform is growing and we are looking to hire a Senior Site Reliability Engineer (SRE) / Cloud Engineer Our main Cloud Platform is Azure (those with Azure will be prioritized... 
    Full time
    Remote work
    Visa sponsorship
    Work visa
    Flexible hours

    Acquird.io

    Toronto, ON
    more than 2 months ago
  •  ...expertise across connectivity, AI, security and more, we’ll map...  ...Summary As a Senior Software Engineer specializing in agentic...  ...influential voice in shaping our GenAI platform's architecture and strategy....  ...tools that are scalable, reliable, and maintainable across the organization... 
    Long term contract
    Full time
    Contract work
    Local area

    Rivian and Volkswagen Group Technologies

    Toronto, ON
    more than 2 months ago
  •  ...Purpose of Job The Senior AI Platform Operations Engineer is accountable for the reliability, operability, and controlled enablement of the organization’s AI platform.  This role ensures that AI Platform services and solutions are production-ready, secure, observable... 
    Full time

    Eq Bank

    Toronto, ON
    more than 2 months ago
  • $120k - $170k per year

     ...About the Role We're looking for a Senior / Principal AI Platform Engineer to build the secure, scalable platform capabilities that enable...  ..., and policy enforcement capabilities. Build highly reliable platform services with clear Service Level Objectives (SLOs),... 
    Long term contract
    Permanent employment
    Full time
    Flexible hours

    RAVL

    Toronto, ON
    more than 2 months ago
  • $160k - $220k per year

     ...Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building...  ...this mission. If you are too, let's talk. Staff Software Reliability Engineer - Data Platform  About the Team The Data Platform team is responsible... 
    Full time
    Local area
    Worldwide
    Flexible hours

    Okta

    Toronto, ON
    more than 2 months ago
  •  ...About the Role As a Staff Platform Engineer - AI Infrastructure, you will build and scale the infrastructure behind Paytm's AI inference...  ...directly affect how fast we ship agents and AI features, how reliably they run, and how efficiently we use our hardware across... 
    Full time

    Paytm

    Toronto, ON
    more than 2 months ago
  •  ...itself — keep reading. About the Role The Agentic AI Foundations team is building the core platform, systems, and primitives that enable Socure to...  ...workflows to agent-native operations. As a Software Engineer II on the team, you will help design, build, and harden... 
    Full time
    Internship

    Socure

    Toronto, ON
    a month ago
  •  ...Engineering Manager, AI Conversation Platform Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their... 
    Full time

    Stripe, Inc.

    Toronto, ON
    more than 2 months ago
  • $140k - $160k per year

     ...managers? We are looking for a  Manager, Platform Engineering for an existing vacancy in our IS team...  ...standards, and step into the hardest reliability, integration, data workflow and production...  ...hands-on depth in at least two of: site reliability engineering, observability,... 
    Full time
    Internship
    Work at office
    Remote work

    Connor, Clark

    Toronto, ON
    3 days ago
  •  ...Purpose of the Job We are looking for a Staff AI Platform & Agent Runtime Engineer to build the foundation for enterprise-scale AI and agent execution...  ..., and scale intelligent agentic workloads securely and reliably.  You will sit at the intersection of platform engineering... 
    Full time

    Eq Bank

    Toronto, ON
    more than 2 months ago
  • $135k - $150k per year

     ...to offer at  . The Opportunity: We’re looking for an AI Engineer, AI Platform & ML Engineering to join our Data & AI Technology Partners...  ...production workflows Help monitor AI cost, performance, reliability, usage, and operational risk, while contributing to reusable... 
    Full time
    Internship

    Aviso Wealth

    Toronto, ON
    more than 2 months ago
  •  ...leading mobile-first work execution platform for industrial and frontline...  ...We’re looking for a Senior AI Platform Developer to build...  ...MaintainX to ship AI-powered products reliably at scale. You will design and...  ...and AI tool execution into reliable, production-ready workflows.... 
    Full time

    Maintainx

    Toronto, ON
    10 days ago
  • $125.28k - $210.6k per year

     ...YOU'LL DO We’re looking for a Software Engineer II to join our WhatsApp team. Our team is...  ...compose messages and and a highly parallelized platform for sending and processing messaging...  ...allows marketers to combine and activate AI agents, models, and features at every touchpoint... 
    Full time
    Internship
    Work at office
    Local area
    Flexible hours

    Braze

    Toronto, ON
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer II (AI Platform). Be the first to apply!