Site Reliability Engineer II (AI Platform)
$110k - $130k per yearOpentable
This hybrid role requires working in the office two days per week.
With millions of diners, 70,000+ restaurant partners and 25+ years of experience, OpenTable, part of Booking Holdings, Inc. (NASDAQ: BKNG), is an industry leader with a passion for helping restaurants thrive. Our world-class technology empowers restaurants to focus on what matters most – their team, their guests, and their bottom line – while enabling diners to discover and book the perfect restaurant for every occasion.
Every employee at OpenTable has a tangible impact on what we do and how we do it. You’ll also be part of a global team and its portfolio of metasearch brands. Hospitality is all about taking care of others, and it defines our culture.
About the job
As a Site Reliability Engineer II on the Serving Platforms team within Infrastructure Engineering, you will design, automate, and manage the core container stack and infrastructure powering our global business applications. Operating in a high-scale, self-hosted environment, you will serve as a subject matter expert for Kubernetes, Linux systems, and cloud-native automation, directly driving the reliability, security, and efficiency of our platform. In this role, you will collaborate with cross-functional engineering teams worldwide, lead greenfield infrastructure projects, resolve complex incidents, and build self-service capabilities that empower application developers across the organization.
Responsibilities
- Maintain, tune, and ensure high availability for the low-level Linux operating system and Kubernetes control plane across our self-hosted bare-metal infrastructure.
- Architect, build, and maintain scalable container management, configuration management, and automation tools across global environments.
- Investigate, resolve, and conduct root-cause analysis for complex infrastructure disruptions and performance bottlenecks at the system call level.
- Participate in high-impact platform engineering projects and collaborate with globally distributed engineering teams to drive infrastructure standardization.
- Participate in the team's on-call rotation to support critical production systems and ensure operational resilience.
- Develop and maintain self-service tools, automation pipelines, and robust infrastructure monitoring to eliminate manual operational overhead.
Minimum Qualifications
- 5+ years of hands-on Linux experience (e.g., Ubuntu, CentOS) with expertise in kernel tuning (sysctl), process management (cgroups/namespaces), system calls, and performance optimization.
- 3+ years of experience using configuration management systems such as Puppet, Chef, Ansible, or SaltStack in production environments.
- Proven experience building, operating, and troubleshooting bare-metal Kubernetes clusters from the ground up, including control plane, etcd, and CNI plugin management.
- Proficiency with continuous system automation and scripting in languages such as Go, Python, Ruby, Perl, or Bash.
- Demonstrated experience responding to live service disruptions, leading root-cause analysis, and operating messaging systems (e.g., Kafka or RabbitMQ) in production.
Preferred Qualifications
- Experience operating, scaling, and monitoring AI/ML or LLM-powered services and workloads in high-concurrency production environments.
- Hands-on expertise with public cloud providers (AWS, GCE, or Azure) and containerized CI/CD pipelines (e.g., GitHub, Jenkins, CircleCI, Docker).
- Experience with distributed key-value stores (e.g., Consul, etcd, Zookeeper, Redis) and enterprise observability/alerting tools (e.g., Prometheus, Sensu).
- Familiarity with server virtualization infrastructure (e.g., Proxmox, VMware, Xen, OpenStack) and low-level networking concepts (IPtables/NFTables, routing, load balancing).
- Experience developing and maintaining OS-level software packaging (RPM/DEB) and participating in globally distributed software engineering teams.
This posting is for an existing vacancy.
Benefits and Perks
- Work from (almost) anywhere for up to 20 days per year
- Focus on mental health and well-being:
- Company-paid therapy sessions through SpringHealth
- Company-paid subscription to Headspace
- Annual company-wide week off a year - the whole team fully recharges (and returns without a pile-up of work!)
- Paid parental leave
- Generous paid vacation + time off for your birthday
- Paid volunteer time
- Focus on your career growth:
- Development Dollars
- Leadership development
- Access to thousands of on-demand e-learnings
- Travel Discounts
- Employee Resource Groups
- 20 days of paid time off
- Private health and dental insurance
- Life and Disability insurance
The best connections happen face-to-face, whether you’re sitting down to dinner or having coffee with a coworker. That’s why OpenTable has adopted a hybrid workplace model. This role aligns with that approach, with an expectation of coming into the office two days a week—giving employees the best of both worlds: in-person collaboration and flexibility.
The best connections happen face-to-face, whether you’re sitting down to dinner or having coffee with a coworker. That’s why OpenTable has adopted a hybrid workplace model. This role aligns with that approach, with an expectation of coming into the office two days a week—giving employees the best of both worlds: in-person collaboration and flexibility.
The expected range of compensation for this position based in Toronto, Canada, including commission and/or bonuses, is $110,000-$130,000 CAD. There are a variety of factors that go into determining a compensation range, including but not limited to external market benchmark data, geographic location, and years of experience sought/required.
We offer a competitive base salary and benefits including: health benefits; flexible spending account; retirement benefits; life insurance; paid time off (including PTO, paid sick leave, medical leave, bereavement leave, floating holidays and paid holidays); and parental leave benefits. This role is eligible to be considered for an annual bonus and equity grant.
Work Environment & Flexibility
At OpenTable, we pride ourselves on fostering a global and dynamic work environment. As a team member with us, you will benefit from a schedule tailored to accommodate a global workforce operating across multiple time zones. While the majority of your responsibilities may align with conventional business hours, there will be instances where you are expected to manage communications - via calls, Slack messages, or emails - outside of regular working hours to effectively collaborate with international colleagues, respond to restaurant partners, and/or address urgent matters. OpenTable will always abide by and consider local laws and regulations.
Inclusion
We’re committed to creating a workplace where everyone feels they belong and can thrive. We know the best ideas come when we bring different voices to the table, so we're building a team as dynamic as the diners and restaurants we serve—and fostering a culture where everyone feels welcome to be themselves.
If you need accommodations during the application or interview process, or on the job, we’re here to support you. Please reach out to your recruiter to request any accommodations.
#LI-Hybrid
- ...combines Strategy, Experience & Design, Engineering and Managed Services. We build digital solutions... ...pod to build and run an internal agent platform on AWS Bedrock AgentCore for a global... ...~6+ years in SRE, platform reliability or observability engineering, with strong...SuggestedFull time
$110k - $130k per year
...of metasearch brands. Hospitality is all about taking care of others, and it defines our culture. About the job As a Site Reliability Engineer II on the Serving Platforms team within Infrastructure Engineering, you will design, automate, and manage the core container stack...SuggestedWork at officeLocal areaWorldwideFlexible hours2 days per week$110k - $120k per year
...recognition programs that celebrate your impact The Job: Site Reliability Engineer The Site Reliability Engineer is responsible for... ...the organization's digital banking and financial services platforms. This role focuses on automating operational processes, defining...SuggestedFull timeTemporary workInternshipWork at officeRemote work$100k - $125k per year
...We are seeking an experienced and motivated Software Engineer to join our dynamic Site Reliability Engineering (SRE) team. As a Site Reliability Engineer,... ...in real time. Why join Tipalti? Tipalti is the AI-powered platform for finance automation, elevating how finance...SuggestedFull timeWork at officeFlexible hours$154k - $200k per year
...for marketers through data-activated content generation and AI decisioning. The world’s most innovative brands rely on Movable... ...America, Europe, Australia, and Japan. As one of our Lead Site Reliability Engineers, you will combine hands-on technical expertise with strategic...SuggestedLong term contractFull time- ...drives business value and inspires extraordinary results. Our AI-powered platform helps organizations modernize financial operations,... ...visibility, and optimize spend across the enterprise. The Site Reliability Engineer III (SRE III) plays a critical role in ensuring Emburse...Full timeManual laborLocal areaFlexible hours
$140k - $155k per year
...This is a hands-on senior engineering role focused on improving production... ...teams to ship secure, reliable, and scalable software with confidence... ...practices that strengthen platform stability over time. Partnering... ...on cloud-native technologies, site reliability engineering...Remote jobPermanent employmentFull timeFlexible hours$140k - $182k per year
...for marketers through data-activated content generation and AI decisioning. The world’s most innovative brands rely on Movable... ..., Europe, Australia, and Japan. As one of our Senior Site Reliability Engineers, you will be 100% hands-on across infrastructure and software...Full time- .... Actual Impact. At Docebo, we’re using AI to change how people learn at work—and we... ...change it. We’re an AI-powered learning platform that helps organizations create, deliver,... ...The Adventure Ahead As the Manager of Site Reliability Engineering (SRE), you will lead a talented...Long term contractFull timeFor contractorsWork at officeWorldwide3 days per week
$153.82k - $277k per year
...If Braze sounds like a place where you can thrive, we can’t wait to meet you. Site Reliability Engineers (SREs) at Braze are responsible for keeping all internal-facing services and platforms running smoothly, ensuring consistent system reliability and maximum...Permanent employmentFull timeInternshipWork at officeLocal areaRemote workFlexible hoursRotating shift- ...about belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a systems thinker. You’ll create middleware and platform guardrails that empower developers to innovate quickly and reliably...Full timeWork at officeLocal areaRemote workWorldwideMonday to fridayFlexible hours
$80 - $110 per hour
...AXON Networks delivers a robust AI-driven, analytics-based orchestration platform and a wide portfolio of next-gen high-speed routers that leverage... ...also operating in Denmark, Spain and Vietnam. The Site Reliability Engineer will improve the availability, performance...Remote jobFull timeContract work- ...Description The SRE Role · SREs are engineers with the right mix of knowledge and... ...observation to entire systems to improve reliability, performance and operability). · We constantly... ...or ansible. · Cloud technologies and platforms such as AWS or Azure using API or...Full time
$110k - $125k per year
...results can be seen in the impact of our innovative health data platform and data management solutions, which are used in over 20... ...Apply today and find plenty of reasons to SMILE! The Cloud Site Reliability Engineer (SRE) is responsible for ensuring the reliability,...Full timeRemote workFlexible hours- ...enterprises who are building AI systems to power magical experiences... ...is a team of researchers, engineers, designers, and more, who are... ...high-performance, scalable and reliable machine learning systems? Do... ...build the next generation of AI platforms powering advanced NLP...Full timeWork at officeRemote workFlexible hours
$144k - $200k per year
**The Team** Platform Engineering is the department within SRE that is responsible for a range of critical... ...role in developing and maintaining the reliable and globally connected multi-cloud... ...Overview** We are seeking a talented Site Reliability Engineer (SRE) with a strong...Full timeWork at officeRemote workWorldwideFlexible hours$170k - $185k per year
...s top-performing asset managers? We are seeking a Head of Platform Engineering, Reliability & Control to lead our horizontal engineering function and... ...~ Drive enterprise-wide automation by implementing AI-enabled monitoring and process orchestration capabilities that...Full timeWork at office3 days per week$130k - $180k per year
...month). Must be able to legally work in Canada (visa or sponsorship won't be provided) Our Platform is growing and we are looking to hire a Senior Site Reliability Engineer (SRE) / Cloud Engineer Our main Cloud Platform is Azure (those with Azure will be prioritized...Full timeRemote workVisa sponsorshipWork visaFlexible hours- ...expertise across connectivity, AI, security and more, we’ll map... ...Summary As a Senior Software Engineer specializing in agentic... ...influential voice in shaping our GenAI platform's architecture and strategy.... ...tools that are scalable, reliable, and maintainable across the organization...Long term contractFull timeContract workLocal area
- ...Purpose of Job The Senior AI Platform Operations Engineer is accountable for the reliability, operability, and controlled enablement of the organization’s AI platform. This role ensures that AI Platform services and solutions are production-ready, secure, observable...Full time
$120k - $170k per year
...About the Role We're looking for a Senior / Principal AI Platform Engineer to build the secure, scalable platform capabilities that enable... ..., and policy enforcement capabilities. Build highly reliable platform services with clear Service Level Objectives (SLOs),...Long term contractPermanent employmentFull timeFlexible hours$160k - $220k per year
...Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building... ...this mission. If you are too, let's talk. Staff Software Reliability Engineer - Data Platform About the Team The Data Platform team is responsible...Full timeLocal areaWorldwideFlexible hours- ...About the Role As a Staff Platform Engineer - AI Infrastructure, you will build and scale the infrastructure behind Paytm's AI inference... ...directly affect how fast we ship agents and AI features, how reliably they run, and how efficiently we use our hardware across...Full time
- ...itself — keep reading. About the Role The Agentic AI Foundations team is building the core platform, systems, and primitives that enable Socure to... ...workflows to agent-native operations. As a Software Engineer II on the team, you will help design, build, and harden...Full timeInternship
- ...Engineering Manager, AI Conversation Platform Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their...Full time
$140k - $160k per year
...managers? We are looking for a Manager, Platform Engineering for an existing vacancy in our IS team... ...standards, and step into the hardest reliability, integration, data workflow and production... ...hands-on depth in at least two of: site reliability engineering, observability,...Full timeInternshipWork at officeRemote work- ...Purpose of the Job We are looking for a Staff AI Platform & Agent Runtime Engineer to build the foundation for enterprise-scale AI and agent execution... ..., and scale intelligent agentic workloads securely and reliably. You will sit at the intersection of platform engineering...Full time
$135k - $150k per year
...to offer at . The Opportunity: We’re looking for an AI Engineer, AI Platform & ML Engineering to join our Data & AI Technology Partners... ...production workflows Help monitor AI cost, performance, reliability, usage, and operational risk, while contributing to reusable...Full timeInternship- ...leading mobile-first work execution platform for industrial and frontline... ...We’re looking for a Senior AI Platform Developer to build... ...MaintainX to ship AI-powered products reliably at scale. You will design and... ...and AI tool execution into reliable, production-ready workflows....Full time
$125.28k - $210.6k per year
...YOU'LL DO We’re looking for a Software Engineer II to join our WhatsApp team. Our team is... ...compose messages and and a highly parallelized platform for sending and processing messaging... ...allows marketers to combine and activate AI agents, models, and features at every touchpoint...Full timeInternshipWork at officeLocal areaFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer II (AI Platform). Be the first to apply!
- site reliability engineer Toronto, ON
- site reliability engineer intern Toronto, ON
- senior site reliability engineer Toronto, ON
- website developer Toronto, ON
- site maintenance Toronto, ON
- site safety Toronto, ON
- site carpenter Toronto, ON
- site reliability engineer
- site reliability engineer intern
- senior site reliability engineer

