Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI DevOps & Reliability Engineer

$123k - $160k per year
Full-time

Branch Metrics

Remote
  • Remote job

At Branch, we power every touchpoint with links that work and insights that prove it. From click to conversion, we make growth measurable. Our unparalleled attribution, backed by AI-enhanced linking, is trusted to deliver seamless experiences that increase ROI, decrease wasted spend, and eliminate siloed attribution.

We bring the same rigor to how we build our team, by empowering our people to move fast, own outcomes, and build something that matters. We take pride in making meaningful investments in our team’s health, wealth, and growth so individuals can thrive as we scale. Our culture values smart, humble, and collaborative teammates who take accountability and drive results in an environment where their work truly moves the business forward.

We are innovative, scaling with purpose, and led by seasoned leaders who know how to build enduring companies. Trusted by brands like Instacart, Western Union, NBCUniversal, ZocDoc, and Sephora, we’re big enough to matter, small enough for you to make a real impact. If you’re excited by the grit of building, rapid learning, and shaping the future of customer growth, you’ll find your place here.

About The Group


We're hiring an AI DevOps & Reliability Engineer to own how software ships and runs at Branch. The role has two areas: half central platform and standards work, half embedded with an engineering team. Centrally, you'll build and operate the delivery platform (CI/CD pipelines, deployment automation, environments) so teams can release safely, frequently, and on demand. Embedded, you'll work hands-on with an engineering team day-to-day on their infrastructure, deployment, and operational practices, mentoring them and building their capability over time.

You'll also lead the adoption of AI in DevOps and SRE work at Branch. Bringing modern AI tooling (Claude Code, agentic workflows) into runbook generation, alerting, incident response, and operational tooling is a core part of this role, not a side project. It's a strategic direction we're committed to.

As a lead, you'll work directly with engineering leadership to shape the operations and delivery roadmap across multiple milestones.

What You'll Do


Delivery & Release Engineering



  • Design and expand deployment automation, advancing the org toward on-demand and continuous production releases.

  • Establish release practices and standards: progressive delivery, rollback, release tracking, deployment inventory teams can trust.

  • Extend automation deeper into production paths, reducing manual steps and release toil.

  • Enable verification through automation: quality gates as code, build engineering supports our efforts.

Pipelines & Guardrails



  • Own CI/CD standards across teams: quality gates, automated checks, guardrails that catch problems before production.

  • Build pipeline tooling that makes the safe path the easy path for engineers.

Environments



  • Design and build out dev, staging, and on-demand (ephemeral) environments that mirror production and spin up on request.

  • Treat environment provisioning as a product: fast, reproducible, self-service.

AI-Embedded Ops



  • Bring AI tooling into operations: automated runbook generation, intelligent alerting, AI-assisted incident response, operational tooling.

  • Help build an org-wide, AI-augmented ops practice and share patterns across teams.

  • This is a core part of the role, aligned with Branch's broader AI direction.

Infrastructure & GitOps



  • Champion Infrastructure as Code (Terraform / CloudFormation) for provisioning, configuration, and lifecycle management.

  • Drive GitOps-based delivery with Argo CD for secure, repeatable, scalable deployments across Kubernetes.

Operational Reliability



  • Bring a strong reliability foundation: alerting practices, on-call, runbooks, SLI/SLO definition, incident response.

  • Partner with engineering teams on the operational practices that keep their services healthy at high volume.

  • Operate and tune high-volume data infrastructure: streaming pipelines (Kafka) and SQL/NoSQL datastores under heavy production load.

  • Strengthen team-level runbooks, operational readiness, and production hygiene; feed improvements back into the platform.

Embedded Team Work



  • Embed with an assigned engineering team day-to-day, working hands-on with them on infrastructure, deployment, and reliability work.

  • Mentor team engineers on operational best practices, observability, and reliability.

  • Help build the team's capability over time so good practices stick.

Engineering Metrics



  • Stand up DORA metrics (lead time, deployment frequency, change failure rate, MTTR) and use them to target real improvements.

  • Make delivery and reliability health visible to teams and leadership.

Leadership & Partnership



  • Work with engineering leadership on the operations and delivery roadmap.

  • Drive cross-team adoption of standards and tooling through collaboration and influence.

What We're Looking For



  • Hands-on experience adopting AI into DevOps and SRE practices (Claude Code, Cursor, agents, or similar) to improve automation, debugging, and operational efficiency.

  • 7+ years in DevOps, platform, infrastructure, or related engineering roles, ideally in fast-scaling environments.

  • Strong hands-on Kubernetes and AWS experience.

  • Deep IaC experience (Terraform and/or CloudFormation) and the ability to set IaC standards for other teams.

  • Proven CI/CD architecture experience: pipelines, quality gates, release automation.

  • GitOps experience with Argo CD (or Flux) for Kubernetes delivery.

  • Hands-on experience operating streaming infrastructure (Kafka) in production.

  • Experience managing SQL and NoSQL datastores at high volume: performance, scaling, operational health.

  • Solid scripting/automation skills (Python, Bash, or similar).

  • Working knowledge of observability stacks: Prometheus, Grafana, PagerDuty (Loki / Alertmanager a plus).

  • Familiarity with on-call, incident response, SLI/SLO definition, and runbooks, and the operational practices that support them.

  • Strong collaborator and communicator. Comfortable working across teams, mentoring engineers, and driving alignment without authority.

Nice to Have



  • Progressive delivery (canary, blue/green) and feature-flag-driven release experience.

  • Cost / efficiency awareness in cloud infrastructure.

  • Broader data / streaming ecosystem exposure (Spark, schema management, CDC, etc.).

What Success Looks Like



  • Teams ship on demand — merge to prod in hours, no tickets, no waiting on you. Deploy frequency up a tier.

  • Faster without breaking — lead time and MTTR down while change-failure rate holds flat.

  • Platform does the work — safe path is the easy path; manual release steps trending to zero; envs self-service in minutes.

  • AI is in the ops loop — runbooks, alerting, incident response AI-assisted; patterns other teams reuse unprompted.

  • Capability sticks — embedded team owns its own deploy and reliability work after you rotate off.

  • Health is visible — DORA metrics instrumented for teams and leadership; roadmap driven by data.

This role is 100% remote in Canada. This role does not qualify for relocation or visa sponsorship. 

In accordance with applicable law, the following represents a reasonable estimated compensation range for this role: the estimated pay range for this role, if based in Canada is 123,000 CAD to 160,000 CAD. Please note that this information is provided for those hired in Canada only. Compensation for candidates outside of Canada will be based on the candidate’s specific work location. Actual compensation will be determined based on skills, experience, and geographic location and may be more or less than the amount shown above. This role additionally includes a 10% annual bonus tied to company goals. 

The salary range provided represents base compensation and does not include potential equity, which is available for qualifying positions. At Branch, we are committed to the well-being of our team by offering a comprehensive benefits package. From health and wellness programs to paid time off and retirement planning options, we provide a range of benefits for qualified employees. For detailed information on the benefits specific to your position, please consult with your recruiter.

Branch is an equal opportunity employer. All applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran or disability status.

If you think you'd be a good fit for this role, we'd love for you to apply! At Branch, we strive to create an inclusive culture that encourages people from all walks of life to bring their unique, diverse perspectives to work. We aim every day to build an environment that empowers us all to do the best work of our careers, and we can't wait to show you what we have to offer!

A little bit about us:  


Branch is the leading provider of engagement and performance mobile SaaS solutions for growth-focused teams, trusted to maximize the value of their evolving digital strategies. The Branch platform provides a seamless experience across paid and organic, on all channels and platforms, online and offline, to eliminate friction and drive valuable action at the moments of highest intent. With Branch, businesses gain accurate mobile measurement and insights into user interactions, enabling them to drive conversions, engagement, and more intelligent marketing spend.

Branch is an award-winning employer headquartered in Mountain View, CA. World-class brands like Instacart, Western Union, NBCUniversal, Zocdoc and Sephora acquire users, retain customers and drive more conversions with Branch.

Candidate Privacy Information:
For more information on the data that Branch will collect through your application, and how we use, share, delete, and retain that information as part of our recruitment and employment efforts, please see our  HR Privacy Policy .

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the AI DevOps & Reliability Engineer in Remote vacancy
  •  ...Staffinity is currently seeking an AI Platform Engineer - DevOps for a premier corporate client located in Halifax. This is a full-time position offering a competitive salary and comprehensive enterprise benefits package. The base salary range is 100-120 + benefits, RRSP... 
    Suggested
    Full time
    Work at office
    3 days per week

    Staffinity Inc.

    Remote
    1 day ago
  •  ...About the Role We are looking for a Site Reliability Engineer to join our Network and Security...  ...reliability-focused improvements Leverage AI tools (e.g., Amazon Kiro) to accelerate...  ...in Site Reliability, Cloud, or DevOps Engineering in SaaS or large-scale production... 
    Suggested
    Long term contract
    Permanent employment
    Full time
    Work at office
    Remote work

    Tecsys Inc.

    Remote
    1 day ago
  • $102k - $134k per year

     ...helping artists use the latest AI tools and make thoughtful...  ...invaluable role in our success. The engineering team at Warner Music Group...  ...are seeking a highly motivated DevOps Engineer to join our DevOps team...  ...to ensure repeatable and reliable deployments. Experience with Cloudformation... 
    Suggested
    Full time
    Local area
    Worldwide

    Warner Music Group Corp.

    Remote
    1 day ago
  • $120k - $140k per year

     ...visible investments. Built on a native, AI-powered platform and more than a decade...  ...About this position: We’re looking for a DevOps Engineer to help support and grow Later’s cloud...  ...GitOps, and MLOps. You’ll help maintain reliable infrastructure, improve deployment... 
    Suggested
    Long term contract
    Permanent employment
    Full time
    Local area
    Remote work

    Later

    Remote
    1 day ago
  •  ...MaintainX is the world's leading AI-powered maintenance and asset...  ...enabled, cloud-based tool for reliability, safety, and operations of...  ...looking for a Site Reliability Engineer to help advance MaintainX’s reliability...  ...in software development, SRE, DevOps, or production development... 
    Suggested
    Full time
    Immediate start

    Maintainx

    Remote
    1 day ago
  • $110k - $120k per year

     ...Time, Permanent Reports to: Site Reliability Lead, Platform Engineering We're looking for Site...  ...execute the Platform Reliability / DevOps and SRE roadmap Support PayByPhone...  ...practices (logging standards, use of reliable libraries, SLA/SLO goals) Participate... 
    Permanent employment
    Full time

    Paybyphone

    Remote
    1 day ago
  •  ...Our mission is to deliver reliable, secure, and scalable data infrastructure...  ...lakehouse capabilities that engineering teams can depend on to build...  ...closely with Cloud Platform, DevOps, and product engineering so...  ...data infrastructure enters its AI-native era, we are investing in... 
    Full time
    Worldwide

    Appdirect

    Remote
    1 day ago
  • $120k - $200k per year

     ...ABOUT THE ROLE At LayerZero, our Site Reliability Engineering (SRE) team is at the intersection of...  ...those external users interact with—are reliable, meet the uptime expectations of our users...  ...as a Site Reliability Engineer, DevOps Engineer, or similar role. ~ Deep familiarity... 
    Full time

    Layer Zero Labs Llc

    Remote
    1 day ago
  •  ...& Digital Marketing | Enterprise AI | Customer Care AI & Technology |...  ...Canada, and we need exceptional engineers to help us scale. We're seeking a Senior DevOps Engineer with deep expertise in AWS...  ..., build, and maintain scalable, reliable, and secure infrastructure that powers... 
    Long term contract
    Full time

    Telus Digital

    Remote
    1 day ago
  • $145k - $185k per year

     ...place in a warehouse, before any Physical AI system is trusted in the real world, it...  ...possible. We're hiring a Senior Site Reliability Engineer to help build and operate that...  ...Qualifications Experience. 5+ years in SRE, DevOps, or infrastructure engineering roles,... 
    Remote job
    Full time

    Parallel Domain

    Remote
    1 day ago
  • $197.5k - $225k per year

     ...About the Team: As a Senior Site Reliability Engineer, you will be a key technical leader driving...  ...also own the infrastructure behind our AI tooling — building MCP servers and defining...  ...: ~6+ years in SRE , DevOps , or Infrastructure roles , with significant... 
    Full time

    Securityscorecard

    Remote
    1 day ago
  • $65k - $130k per year

     ...encourages employees to collaborate and learn from each other, completely free of barriers. Your role: As a DevOps/SRE, you will be responsible for the reliability and smooth operation of your service in both production and test environments. At Global Relay we use... 
    Full time

    Global Relay

    Remote
    1 day ago
  •  ...Ton rôle en tant que Ingénieur DevOps   Tu auras pour...  ...enhanced with advanced integrated AI. It transforms real-time data...  ...partner network, Nectari delivers a reliable, modern, and accessible analytics...  ....   Your role as DevOps engineer   You will be responsible... 
    Daily paid
    Full time
    Work at office
    Work from home
    Worldwide
    Flexible hours
    Weekend work
    Afternoon shift

    Nectari Software Inc.

    Remote
    1 day ago
  • $100k - $125k per year

     ...experienced and motivated Software Engineer to join our dynamic Site Reliability Engineering (SRE) team. As a Site Reliability...  ...Why join Tipalti? Tipalti is the AI-powered platform for finance...  ...at Tipalti  Our tech teams are the engine behind our business. Tipalti’s tech... 
    Full time
    Work at office
    Flexible hours

    Tipalti

    Remote
    1 day ago
  • $101.2k - $136.9k per year

     ...member facing services. We are looking for a highly skilled Site Reliability Engineer (SRE) who will help build, operate and continuously improve...  ...~ Design and maintain CI/CD pipelines (GitHub Actions, Azure DevOps)  ~ Automate deployments, scaling, configuration and... 
    Permanent employment
    Full time
    Internship
    Work at office
    Immediate start
    Home office
    Flexible hours
    2 days per week

    Vancity

    Remote
    1 day ago
  •  ...role We’re looking for a Senior Site Reliability Engineer (SRE) to help strengthen and scale our multi...  ...across engineering Lead rollouts of AI-native tooling for code review, testing,...  ...on Site Reliability Engineering (SRE), DevOps, Platform Engineering, or similar roles.... 
    Full time
    Internship
    Remote work
    Work from home

    Scalepad

    Remote
    1 day ago
  • $163k - $194k per year

     ...specified location above.  We are AI Native We are building...  ...run the platforms that every engineering team at Life360 depends on,...  ...is hiring an AI-Native Site Reliability Engineer — a senior engineer...  ...future. As a Senior Site Reliable Engineer II - Infrastructure... 
    Full time
    Summer work
    Remote work
    Flexible hours

    Life360

    Remote
    1 day ago
  •  ...Description Prioritize candidates with medical device or regulated hardware experience, strong background in reliability engineering, HALT/HASS/ALT, and statistical modeling (Weibull, lognormal). Technical Evaluation – Assess expertise in DFMEA/PFMEA, fault tree analysis... 
    Full time

    Sapsol Technologies Inc

    Remote
    1 day ago
  •  ...group includes our scientific research & engineering division (Skynet Software) and Canadian...  ...DESCRIPTION: Chelsea Avondale is looking for a Reliability Engineer with a background in...  ...maintaining high-performance, scalable, and reliable web systems. ~ We also encourage... 
    Full time

    Chelsea Avondale

    Remote
    1 day ago
  • $120k - $160k per year

     ...ABOUT YOU We are looking for a Site Reliability Engineer (Monetization) who is pragmatic,...  ...Qualifications & Skills ~3+ years of proven SRE, DevOps, or platform engineering experience: on-...  ...We may use artificial intelligence (AI) tools to support parts of the hiring... 
    Long term contract
    Full time

    Xsolla

    Remote
    1 day ago
  •  ...good fit for you! We're looking for a  Data Platform Engineer (DevOps)to join our TecsysIQ Data & AI team. This is a hybrid role that builds and operates...  ...the data pipelines that keep the platform running reliably at scale. If you are equally comfortable writing Terraform... 
    Permanent employment
    Full time
    Remote work

    Tecsys Inc.

    Remote
    1 day ago
  •  ...where you come in. About the Team Our Platform & Reliability crew at Flinks brings together DevOps Engineers, SREs, and Support Engineers to create a proactive...  ...building product, not fighting fires—supported by emerging AI-assisted automation that streamlines intake, triage,... 
    Full time

    Flinks

    Remote
    9 days ago
  •  ...Position: DevOps Engineer - Senior Location: Toronto, ON (Hybrid) Job ID#: RQ00768 Duration: 6 Months Role Overview...  .... Monitor and assess cloud-hosted applications to ensure reliability, availability, and optimal performance. Identify, analyze,... 
    Full time
    3 days per week

    Symbiotic Digital

    Remote
    1 day ago
  •  ...ABOUT THE ROLE At LayerZero, our Site Reliability Engineering (SRE) team is at the intersection of...  ...those external users interact with — are reliable, meet the uptime expectations of our...  ...practical experience.   ~6+ years in SRE, DevOps, or infrastructure engineering,... 
    Full time

    Layer Zero Labs Llc

    Remote
    1 day ago
  •  ...please visit .     The Role     CMG is looking for a Site Reliability Engineer (SRE) with a strong focus on monitoring, observability, and...  ...to explore solutions using bleeding edge technologies such as AI and bring recommendations to the table. We are in a period of... 
    Remote job
    Full time
    Local area

    Capital Markets Gateway

    Remote
    1 day ago
  • $80k - $150k per year

     ...Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our...  ...built and running on Google App Engine (GAE) and microservices....  ...with demanding availability, reliability, and performance requirements.... 
    Full time
    Contract work
    Fixed term contract
    Internship
    Flexible hours

    Innodata Inc.

    Remote
    1 day ago
  • $100k - $150k per year

     ...currencies. Longevity Opportunity Vision Enjoy the game! Qualifications & Skills: Major: 4 to 7 years experience as a DevOps engineer Extensive experience with Linux Extensive experience with container orchestration using Kubernetes Experience with Self-... 
    Full time
    Flexible hours

    Xsolla

    Remote
    1 day ago
  •  ...like an environment that you believe could work for you then read on to find out more. The role: We’re looking for a Site Reliability Engineer to manage, maintain, improve and provide support on our platform. You will be curious by nature, always looking for ways to... 
    Full time
    Remote work

    Tyk Technologies Limited

    Remote
    1 day ago
  •  ...The Role We’re looking for a Senior DevOps Engineer to help scale and operate the infrastructure...  ...cloud infrastructure, automation, and reliability. You’ll partner closely with...  ...strengthen security, and enable fast, reliable deployments. What You’ll Do Design... 
    Full time
    Remote work
    Flexible hours

    Sureify

    Remote
    1 day ago
  • $110k - $125k per year

     ...on technology that directly impacts the reliability, performance, and security of global...  ...enterprises. We are seeking a talented Senior DevOps Engineer to take ownership of our Secure Access...  ...our build processes while ensuring reliable, efficient software delivery. This... 
    Permanent employment
    Full time
    Flexible hours
    Shift work

    Absolute Software Corporation

    Remote
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI DevOps & Reliability Engineer. Be the first to apply!