Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Ingénieur·e SRE / Site Reliability Engineer

Full-time

Mthree Recruiting Portal

**English version below**

Doit être local à Montréal

Vous souhaitez travailler dans le domaine de la technologie au sein d'une banque d'investissement?


Nous recherchons une personne pour rejoindre une équipe dynamique en tant qu’ Ingénieur·e Fiabilité de Site (Site Reliability Engineer) pour l’un de nos clients. Le Site Reliability Engineering (SRE) est une discipline orientée production, axée sur l’amélioration de la disponibilité des services systèmes, de l’observabilité, de l’évolutivité, de la performance et de la fiabilité des produits technologiques, en appliquant de solides principes d’ingénierie logicielle et en adoptant les technologies et outils les plus récents.



Nous serions ravis de vous rencontrer si vous :


  • Vous intéressez aux systèmes distribués et au travail sur des services hautement évolutifs, fiables et à grande échelle.

  • Aimez évoluer dans un environnement dynamique et n’avez pas peur de changer les choses pour les améliorer.

  • Appréciez les nouveaux défis technologiques et la résolution de problèmes complexes.

  • Croyez qu’une équipe qui collabore efficacement est véritablement plus intelligente que la personne la plus brillante qui la compose.

  • Aspirez à évoluer en tant que personne, coéquipier·e et ingénieur·e.

  • Faites preuve de détermination, de motivation et d’un profond sens des responsabilités.

À propos de mtrois :

Depuis 2010, mtrois aide ses clients à résoudre leurs défis commerciaux et technologiques. Nous sommes une société de conseil en technologie et en affaires avec une main-d'œuvre mondiale qui réalise des projets commerciaux et informatiques significatifs dans certaines des plus grandes organisations de services financiers du monde.

Services principaux


  • Consulting et Conseil

  • Services gérés

  • Programme de diplômés Alumni

  • Programme Alumni Pro

Nous avons une présence mondiale et sommes experts dans la fourniture d'une qualité exceptionnelle à notre base de clients, offrant des services de conseil dans les domaines du risque, de la réglementation et de la conformité ; Produits des fournisseurs ; Support d'application ; Développement d'application ; Cyber et sécurité de l'information ; Science des données et DevOps.

Notre programme Expert offre aux professionnels expérimentés l'accès à des rôles de premier plan dans la technologie, la finance, l'aviation et l'assurance. Rejoignez-nous pour travailler sur des projets technologiques révolutionnaires, des plateformes de trading internationales aux applications critiques pour les principales compagnies aériennes. Nous recrutons des professionnels désireux de faire progresser rapidement leur carrière dans la technologie ou les opérations au sein d'organisations mondiales prestigieuses.


Responsabilités :


  • Travailler en étroite collaboration avec les équipes d’ingénierie et de développement pour concevoir, construire et maintenir des systèmes, tout en les conseillant sur le choix des produits, la conception des schémas et l’optimisation des requêtes.

  • Diagnostiquer et résoudre des problèmes à travers l’ensemble de la stack : matériel, logiciel, application et réseau.

  • Identifier et piloter les opportunités d’amélioration de l’automatisation de nos plateformes ; définir et créer des automatisations pour le déploiement, la gestion et la visibilité de nos services.

  • Identifier de manière proactive les risques liés à la fiabilité des systèmes et y remédier.

  • Représenter l’organisation RPE lors des revues de conception et des exercices de préparation opérationnelle pour les services nouveaux et existants.

  • Travailler aux côtés des membres des équipes mondiales et régionales existantes selon un modèle de suivi en continu (“follow‑the‑sun”).

  • Participer à la rotation d’astreintes ainsi qu’aux appels périodiques avec des spécialistes situés dans d’autres fuseaux horaires.


Compétences Requises :


  • Formation en informatique équivalente à un diplôme de niveau licence (B.Sc.), ou expérience pratique équivalente.


  • Doit impérativement avoir une expérience avec Kubernetes et la gestion d’applications conteneurisées.


  • Expérience avérée en automatisation, notamment via des langages de script tels que Python, Bash ou Perl. La maîtrise d’au moins un langage de plus haut niveau est souhaitée.

  • Expérience du support d’architectures trois tiers, incluant une exposition aux plateformes UNIX/Linux et aux bases de données telles qu’IBM DB2, Sybase, MongoDB, GreenPlum, etc.

  • Expérience avec les dépôts de code source et binaires, les outils de build et CI/CD (Git, Artifactory, Jenkins, Docker), ainsi qu’avec des technologies de streaming de données comme Spark ou Kafka.

  • Maîtrise des outils d’entreprise tels que Grafana, Dynatrace ou AppDynamics.

  • Connaissance et compréhension des architectures logicielles et systèmes modernes : load balancing, file d’attente, caching, modes de défaillance des systèmes distribués, microservices, etc.

  • Solide compréhension des concepts liés au système d’exploitation (processus, allocation mémoire, stack réseau), de leurs impacts sur les applications, et capacité à en effectuer le débogage.

  • De manière générale, une expérience pratique dans l’exploitation de systèmes en ligne à grande échelle constitue un avantage certain.

Chez mtrois, nos valeurs soutiennent des coéquipiers courageux, des moteurs d'aiguilles et des champions de l'apprentissage tout en s'efforçant de soutenir la santé et le bien-être de tous les employés. Nous sommes très fiers de célébrer la diversité de chaque individu qui contribue à faire de mtrois l'entreprise qu'elle est aujourd'hui et sera à l'avenir. Nous valorisons la diversité tant au sein de mtrois qu'avec nos entreprises partenaires, et nous sommes fiers de fournir un environnement où tous nos collègues peuvent s'épanouir. Cela signifie promouvoir une forte culture d'égalité mais, surtout, d'inclusion.

Les candidats doivent être actuellement autorisés à travailler au Canada à temps plein. L'entreprise ne parrainera pas les candidats pour des visas de travail

 

**English Translation**

**Must be local to Montreal**

Want to work in technology at an investment bank?

We are looking for someone to be a part of a dynamic team as a Site Reliability Engineer for one of our clients. Systems Reliability Engineering (SRE) is a production-oriented discipline focused on improving system service availability, observability, scalability, performance, and reliability for technology products by applying sound software engineering principles and adopting the latest technology and tooling.

 

We would like to talk to you if you:


  • Are interested in distributed systems and working with high scale scalable and reliable services.

  • Like to work in a fast-moving environment and you aren't afraid to change things to make them better.

  • Enjoy new technological challenges and solving hard problems.

  • Believe that a team working well together is truly smarter than the single smartest person on that team.

  • Aspire to grow as a person, as a teammate, and as an engineer.

  • Have Grit, drive and a deep feeling of ownership.

 

 

About mthree:

Since 2010, mthree has been helping clients solve their business and technological challenges. We are a technology and business consultancy with a global workforce delivering significant business and IT projects in some of the largest financial services organizations worldwide.


  • Core Services

  • Consulting and Advisory

  • Managed Services

  • Alumni Graduate Program

  • Alumni Pro Program

We have a global presence and are experts in delivering exceptional quality to our client base, providing consulting services across Risk, Regulation & Compliance; Vendor Products; Application Support; Application Development; Cyber & Information Security; Data Science and DevOps areas.

Our Expert program offers experienced professionals access to top roles in tech, finance, aviation and insurance. Join us to work on groundbreaking technology projects, from international trading platforms to critical applications for leading airlines. We recruit professionals who are eager to fast-track their careers in technology or operations within prestigious global organizations.

Responsibilities:


  • Working closely with engineering/development teams to design, build, and maintain systems and help them decide on products to use, schema design and query tuning.

  • Troubleshoot issues across the entire stack: hardware, software, application and network.

  • Identifying and drive opportunities to improve automation for our platforms; scope and create automation for deployment, management and visibility of our services.

  • Proactively identifying and addressing systems reliability risks.

  • Represent the RPE organization in design reviews and operational readiness exercises for new and existing services.

  • Working alongside existing global and regional team members on a follow-the-sun basis.

  • Participate in on-call rotation and periodic conference calls with other specialists from other time zones.

Skills Required:


  • Background in Computer Science equivalent to a B.Sc. Equivalent practical experience is a reasonable substitute.

  • Must have experience with kubernetes and management of containerized applications

  • Automation-related experience is particularly valued using scripting languages such as python, bash, Perl. One higher level language is desired.

  • Experience on supporting three tier architecture which includes exposure to UNIX, Linux platforms and databases such IBM DB2, Sybase, Mongo, GreenPlum etc.

  • Experience with source code and binary repositories, build tools, and CI/CD (Git, Artifactory, Jenkins, Docker) etc and data streaming technologies like Spark, Kafka etc.

  • Hands on experience on enterprise tools set such as Grafana, Dynatrace, AppDynamics etc.

  • Awareness of, and ability to reason about modern software & systems architectures, including load-balancing, queueing, caching, distributed systems failure modes, micro services etc

  • Deep understanding of operating system level concepts such as processes, memory allocation, and the network stack; understanding of how applications are affected by the above, and ability to debug same.

  • Generally speaking, practical experience running large scale online systems is always an advantage.

 

At mthree, our values support courageous teammates, needle movers, and learning champions all while striving to support the health and well-being of all employees.  We take great pride in celebrating the diversity of each individual who contributes to making mthree the company it is today and will be in the future. We value diversity both within mthree and with our partner companies, and we're proud to provide an environment where all our colleagues can flourish. That means promoting a strong culture of equality but, most importantly, inclusion.

Applicants must be currently authorized to work in Canada on a full-time basis. The Company will not sponsor applicants for work visas.

Vacancy posted 6 hours ago
Similar jobs that could be interesting for youBased on the Ingénieur·e SRE / Site Reliability Engineer in Remote vacancy
  •  ...in AUM and 22 global investment banks. For more information, please visit .     The Role     CMG is looking for a Site Reliability Engineer (SRE) with a strong focus on monitoring, observability, and alerting to ensure the reliability, performance, and scalability... 
    Suggested
    Remote job
    Full time
    Local area

    Capital Markets Gateway

    Remote
    6 hours ago
  •  ...environments. We are a modern, IoT-enabled, cloud-based tool for reliability, safety, and operations of physical equipment and...  ...valuing the company at $2.5 billion. We’re looking for a Site Reliability Engineer (SRE) to help advance MaintainX’s reliability, observability,... 
    Suggested
    Full time

    Maintainx

    Remote
    6 hours ago
  • $101.2k - $136.9k per year

     ...strategy across digital banking, core banking, data platforms and member facing services. We are looking for a highly skilled Site Reliability Engineer (SRE) who will help build, operate and continuously improve the reliability, performance, security and automation of our cloud... 
    Suggested
    Permanent employment
    Full time
    Internship
    Work at office
    Immediate start
    Home office
    Flexible hours
    2 days per week

    Vancity

    Remote
    6 hours ago
  •  ...read on to find out more. The role: We’re looking for a Site Reliability Engineer to manage, maintain, improve and provide support on our...  ...operations and support Writing and maintaining documentation on SRE processes and policies Recommending and implementing ways... 
    Suggested
    Full time
    Remote work

    Tyk Technologies Limited

    Remote
    6 hours ago
  • $120k - $200k per year

     ...Coinbase Ventures, Uniswap Labs, Circle Ventures, Delphi Digital, and many more.   ABOUT THE ROLE At LayerZero, our Site Reliability Engineering (SRE) team is at the intersection of software and systems engineering, dedicated to crafting and maintaining large-scale,... 
    Suggested
    Full time

    Layer Zero Labs Llc

    Remote
    6 hours ago
  •  ...a good fit for you! About the Role We are looking for a Site Reliability Engineer to join our Network and Security Operations Center (NOC), a team...  ...and maintain technical documentation and contribute to SRE best practices Partner with platform engineering, deployment... 
    Long term contract
    Permanent employment
    Full time
    Work at office
    Remote work

    Tecsys Inc.

    Remote
    6 hours ago
  • $110k - $160k per year

     ...hear from you!  Role Overview  We're seeking a Senior Site Reliability Engineer to join our SaaS-Ops team within Shared Services Engineering....  ...that demonstrates growing depth in cloud infrastructure and SRE practices; ~ Managed production Kubernetes environments at... 
    Full time
    Internship
    Work at office
    Local area
    Flexible hours

    Magnet Forensics

    Remote
    6 hours ago
  •  ...We are seeking a Senior DevOps & Site Reliability Engineer to own the reliability, scalability, performance, and operational excellence of Medeloop...  ...pipelines, infrastructure as code, and cloud architecture—with SRE discipline: SLOs, incident management, capacity planning,... 
    Hourly pay
    Full time

    Medeloop

    Remote
    6 hours ago
  •  ...and we are not tiptoeing into it. We are rebuilding our engineering culture around a simple belief: AI changes everything....  ...on — and this role is at the center of keeping it reliable, fast, and scalable. As a Staff SRE, you'll own the infrastructure and reliability practices... 
    Remote job
    Full time
    Internship
    Work at office
    Local area
    Flexible hours
    Shift work
    Weekend work

    Babylist, Inc

    Remote
    6 hours ago
  • $120k - $160k per year

     ...ABOUT YOU We are looking for a Site Reliability Engineer (Monetization) who is pragmatic, product-minded, and equally comfortable writing code...  ...production systems to join our Infrastructure department's SRE team . The best candidate will be someone who thrives in a... 
    Long term contract
    Full time

    Xsolla

    Remote
    6 hours ago
  •  ...America, Europe, Australia, and Japan. As one of our Lead Site Reliability Engineers, you will combine hands-on technical expertise with strategic...  ...scenarios Lead cross-functional reliability initiatives with SRE and service engineering teams, influencing architectural... 
    Long term contract
    Full time

    Movable Ink

    Remote
    6 hours ago
  •  ...checking the market? Well… you might just be in the right place! We’re looking for a Staff Site Reliability Engineer to join our Data team in Canada. As a Staff Data SRE, you are the technical backbone of the Data Office's infrastructure platform. Your scope spans... 
    Full time
    Work at office
    Remote work
    Flexible hours
    Shift work

    Lightspeed Commerce

    Remote
    6 hours ago
  •  ...Gauss Labs is seeking a highly skilled Site Reliability Engineer to join our team in Vancouver. As an SRE at Gauss Labs, you will play a critical role in ensuring our industrial AI platform's reliability, performance, and scalability. You will be responsible for building and... 
    Full time

    Gauss Labs

    Remote
    6 hours ago
  • $123k - $160k per year

     ...your place here. We are seeking a highly experienced Senior Site Reliability Engineer to own the reliability, performance, and operational...  ...environments. You’ll be a good fit if you have: ~6+ years in SRE, systems engineering, or software engineering roles, ideally... 
    Remote job
    Long term contract
    Full time
    Relocation

    Branch Metrics

    Remote
    6 hours ago
  •  ...The Site Reliability Engineering organization at Pinterest is accountable for ensuring overall Pinterest availability as well as enhancing Engineering...  ...as Pinterest continues to grow and scale. As a Pinterest SRE, you will design and build systems, platforms, tools, frameworks... 
    Full time
    Work at office
    Relocation
    Relocation package

    Pinterest

    Remote
    6 hours ago
  •  ...une grande organisation mondiale. Les ingénieurs prospèrent chez Tower tout en...  ...the world’s best systematic trading and engineering talent. We empower portfolio managers to...  ...the operational burden low and to enhance reliability Monitoring daily reports and run of scheduled... 
    Daily paid
    Apprenticeship
    Casual work
    Work at office
    Worldwide
    Rotating shift

    Tower Research Capital

    Remote
    6 hours ago
  • $180.4k - $230.4k per year

     ...Coalition. About the role We are looking for a Staff Site Reliability Engineer to lead AI enablement across our engineering organization. As...  ...foundations trustworthy. This role sits within our Platform SRE team, and you'll participate in the team's ad-hoc support... 
    Full time
    Remote work
    Home office
    Flexible hours
    Shift work

    Coalition

    Remote
    6 hours ago
  • $145k - $185k per year

     ...makes simulation at scale possible. We're hiring a Senior Site Reliability Engineer to help build and operate that infrastructure. This role sits...  ...Required Qualifications Experience. 5+ years in SRE, DevOps, or infrastructure engineering roles, with a track record... 
    Remote job
    Full time

    Parallel Domain

    Remote
    6 hours ago
  • $197.5k - $225k per year

     ...GV and Riverwood Capital. About the Team: As a Senior Site Reliability Engineer, you will be a key technical leader driving the design and optimization...  .... Required Qualifications: ~6+ years in SRE , DevOps , or Infrastructure roles , with significant production... 
    Full time

    Securityscorecard

    Remote
    6 hours ago
  •  ...in 2024 – but we're just getting started.      As a Sr. Site Reliability Engineer, you'll be the guardian of our platform's reliability and performance...  ...continuous improvement across our engineering teams. Our SRE Team: We're a bottom-up, collaborative team that thrives... 
    Full time
    Work at office
    Local area
    Remote work
    Work from home
    Home office
    Weekend work

    Third-Party Job Posts

    Remote
    6 hours ago
  •  ...corporate culture by MSP Today, G2, and Great Place to Work™, to name a few.  About the role We’re looking for a Senior Site Reliability Engineer (SRE) to help strengthen and scale our multi-cloud platform and developer experience. This is a hands-on senior individual... 
    Full time
    Internship
    Remote work
    Work from home

    Scalepad

    Remote
    6 hours ago
  • $150k - $240k per year

     ...As a contributor in the SRE organization, you are passionate...  ...achieving the high quality and reliability our customers demand. You will...  ...deliverables will reach the entire engineering organization to enable product...  .... Exemplify cloud-native site reliability best practices.... 
    Long term contract
    Full time
    Remote work

    Axon

    Remote
    6 hours ago
  •  ...visualizing relationships between entities in the system. As a Site Reliability Engineer you will be responsible for the availability, latency,...  ...DevOps, Product and Engineering teams to design and implement SRE practice at Behavox to build foundational infrastructure allowing... 
    Full time
    Flexible hours

    Behavox

    Remote
    6 hours ago
  • $153k - $187k per year

     ...clear signal owners can use to run stronger, more resilient businesses. We’re looking for an incredible Senior Site Reliability Engineer to join our SRE team. We aim to make reliability, security, and speed reinforce one another so that the platform becomes the engine... 
    Full time
    Internship

    Relay

    Remote
    6 hours ago
  •  ...senior contributor in the APX SRE organization, you are passionate...  ...about the high quality, reliability, and security our customers demand...  ...deliverables will reach the entire engineering organization to enable product...  .... Exemplify cloud-native site reliability best practices... 
    Long term contract
    Full time
    Remote work
    Flexible hours

    Axon

    Remote
    6 hours ago
  • $69k - $90k per year

     ...accolades from global entities like the World Summit Awards and the UN-endorsed Zero Project.  About the role  As a Junior Site Reliability Engineer (SRE) at Fable, you will help support the reliability, performance, and scalability of the systems that power our products.... 
    Full time
    Internship

    Fable

    Remote
    6 hours ago
  • $20 per day

     ...our careers page to see how you can grow with us! As a Site Reliability Engineer at Hiive, you will be responsible for ensuring the...  ...performance and system behavior, and ensuring these services are reliable, scalable, and cost-efficient in production. In this role... 
    Full time
    Summer holiday
    Relocation

    Hiive

    Remote
    6 hours ago
  • $140k - $180k per year

     ...the world and our goal is preserving uncensored Internet access and online privacy for all. Right now we are looking for a Site Reliability Engineer to help us tame DNS. About the Position Linux system administration and troubleshooting Network configuration and... 
    Full time
    Direct hire
    Work at office
    Remote work

    Funded.club

    Remote
    6 hours ago
  •  ...a part of our journey! About the role We are committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building and leading processes to ensure the reliability,... 
    Full time
    Local area
    Remote work
    Home office
    Flexible hours

    Clickhouse

    Remote
    6 hours ago
  • $260k - $275k per year

     ...enterprises • Solve complex reliability challenges at scale • Influence architecture and engineering culture at a company level •...  ...will focus on creating reusable, reliable, and scalable solutions that abstract...  ..., Platform Engineering, or Site Reliability Engineering role,... 
    Full time

    Saviynt

    Remote
    6 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Ingénieur·e SRE / Site Reliability Engineer. Be the first to apply!