Ingénieur·e SRE / Site Reliability Engineer
Mthree Recruiting Portal
**English version below**
Doit être local à Montréal
Vous souhaitez travailler dans le domaine de la technologie au sein d'une banque d'investissement?
Nous recherchons une personne pour rejoindre une équipe dynamique en tant qu’ Ingénieur·e Fiabilité de Site (Site Reliability Engineer) pour l’un de nos clients. Le Site Reliability Engineering (SRE) est une discipline orientée production, axée sur l’amélioration de la disponibilité des services systèmes, de l’observabilité, de l’évolutivité, de la performance et de la fiabilité des produits technologiques, en appliquant de solides principes d’ingénierie logicielle et en adoptant les technologies et outils les plus récents.
Nous serions ravis de vous rencontrer si vous :
- Vous intéressez aux systèmes distribués et au travail sur des services hautement évolutifs, fiables et à grande échelle.
- Aimez évoluer dans un environnement dynamique et n’avez pas peur de changer les choses pour les améliorer.
- Appréciez les nouveaux défis technologiques et la résolution de problèmes complexes.
- Croyez qu’une équipe qui collabore efficacement est véritablement plus intelligente que la personne la plus brillante qui la compose.
- Aspirez à évoluer en tant que personne, coéquipier·e et ingénieur·e.
- Faites preuve de détermination, de motivation et d’un profond sens des responsabilités.
À propos de mtrois :
Depuis 2010, mtrois aide ses clients à résoudre leurs défis commerciaux et technologiques. Nous sommes une société de conseil en technologie et en affaires avec une main-d'œuvre mondiale qui réalise des projets commerciaux et informatiques significatifs dans certaines des plus grandes organisations de services financiers du monde.
Services principaux
- Consulting et Conseil
- Services gérés
- Programme de diplômés Alumni
- Programme Alumni Pro
Nous avons une présence mondiale et sommes experts dans la fourniture d'une qualité exceptionnelle à notre base de clients, offrant des services de conseil dans les domaines du risque, de la réglementation et de la conformité ; Produits des fournisseurs ; Support d'application ; Développement d'application ; Cyber et sécurité de l'information ; Science des données et DevOps.
Notre programme Expert offre aux professionnels expérimentés l'accès à des rôles de premier plan dans la technologie, la finance, l'aviation et l'assurance. Rejoignez-nous pour travailler sur des projets technologiques révolutionnaires, des plateformes de trading internationales aux applications critiques pour les principales compagnies aériennes. Nous recrutons des professionnels désireux de faire progresser rapidement leur carrière dans la technologie ou les opérations au sein d'organisations mondiales prestigieuses.
Responsabilités :
- Travailler en étroite collaboration avec les équipes d’ingénierie et de développement pour concevoir, construire et maintenir des systèmes, tout en les conseillant sur le choix des produits, la conception des schémas et l’optimisation des requêtes.
- Diagnostiquer et résoudre des problèmes à travers l’ensemble de la stack : matériel, logiciel, application et réseau.
- Identifier et piloter les opportunités d’amélioration de l’automatisation de nos plateformes ; définir et créer des automatisations pour le déploiement, la gestion et la visibilité de nos services.
- Identifier de manière proactive les risques liés à la fiabilité des systèmes et y remédier.
- Représenter l’organisation RPE lors des revues de conception et des exercices de préparation opérationnelle pour les services nouveaux et existants.
- Travailler aux côtés des membres des équipes mondiales et régionales existantes selon un modèle de suivi en continu (“follow‑the‑sun”).
- Participer à la rotation d’astreintes ainsi qu’aux appels périodiques avec des spécialistes situés dans d’autres fuseaux horaires.
Compétences Requises :
- Formation en informatique équivalente à un diplôme de niveau licence (B.Sc.), ou expérience pratique équivalente.
Doit impérativement avoir une expérience avec Kubernetes et la gestion d’applications conteneurisées.
- Expérience avérée en automatisation, notamment via des langages de script tels que Python, Bash ou Perl. La maîtrise d’au moins un langage de plus haut niveau est souhaitée.
- Expérience du support d’architectures trois tiers, incluant une exposition aux plateformes UNIX/Linux et aux bases de données telles qu’IBM DB2, Sybase, MongoDB, GreenPlum, etc.
- Expérience avec les dépôts de code source et binaires, les outils de build et CI/CD (Git, Artifactory, Jenkins, Docker), ainsi qu’avec des technologies de streaming de données comme Spark ou Kafka.
- Maîtrise des outils d’entreprise tels que Grafana, Dynatrace ou AppDynamics.
- Connaissance et compréhension des architectures logicielles et systèmes modernes : load balancing, file d’attente, caching, modes de défaillance des systèmes distribués, microservices, etc.
- Solide compréhension des concepts liés au système d’exploitation (processus, allocation mémoire, stack réseau), de leurs impacts sur les applications, et capacité à en effectuer le débogage.
- De manière générale, une expérience pratique dans l’exploitation de systèmes en ligne à grande échelle constitue un avantage certain.
Chez mtrois, nos valeurs soutiennent des coéquipiers courageux, des moteurs d'aiguilles et des champions de l'apprentissage tout en s'efforçant de soutenir la santé et le bien-être de tous les employés. Nous sommes très fiers de célébrer la diversité de chaque individu qui contribue à faire de mtrois l'entreprise qu'elle est aujourd'hui et sera à l'avenir. Nous valorisons la diversité tant au sein de mtrois qu'avec nos entreprises partenaires, et nous sommes fiers de fournir un environnement où tous nos collègues peuvent s'épanouir. Cela signifie promouvoir une forte culture d'égalité mais, surtout, d'inclusion.
Les candidats doivent être actuellement autorisés à travailler au Canada à temps plein. L'entreprise ne parrainera pas les candidats pour des visas de travail
**English Translation**
**Must be local to Montreal**
Want to work in technology at an investment bank?
We are looking for someone to be a part of a dynamic team as a Site Reliability Engineer for one of our clients. Systems Reliability Engineering (SRE) is a production-oriented discipline focused on improving system service availability, observability, scalability, performance, and reliability for technology products by applying sound software engineering principles and adopting the latest technology and tooling.
We would like to talk to you if you:
- Are interested in distributed systems and working with high scale scalable and reliable services.
- Like to work in a fast-moving environment and you aren't afraid to change things to make them better.
- Enjoy new technological challenges and solving hard problems.
- Believe that a team working well together is truly smarter than the single smartest person on that team.
- Aspire to grow as a person, as a teammate, and as an engineer.
- Have Grit, drive and a deep feeling of ownership.
About mthree:
Since 2010, mthree has been helping clients solve their business and technological challenges. We are a technology and business consultancy with a global workforce delivering significant business and IT projects in some of the largest financial services organizations worldwide.
- Core Services
- Consulting and Advisory
- Managed Services
- Alumni Graduate Program
- Alumni Pro Program
We have a global presence and are experts in delivering exceptional quality to our client base, providing consulting services across Risk, Regulation & Compliance; Vendor Products; Application Support; Application Development; Cyber & Information Security; Data Science and DevOps areas.
Our Expert program offers experienced professionals access to top roles in tech, finance, aviation and insurance. Join us to work on groundbreaking technology projects, from international trading platforms to critical applications for leading airlines. We recruit professionals who are eager to fast-track their careers in technology or operations within prestigious global organizations.
Responsibilities:
- Working closely with engineering/development teams to design, build, and maintain systems and help them decide on products to use, schema design and query tuning.
- Troubleshoot issues across the entire stack: hardware, software, application and network.
- Identifying and drive opportunities to improve automation for our platforms; scope and create automation for deployment, management and visibility of our services.
- Proactively identifying and addressing systems reliability risks.
- Represent the RPE organization in design reviews and operational readiness exercises for new and existing services.
- Working alongside existing global and regional team members on a follow-the-sun basis.
- Participate in on-call rotation and periodic conference calls with other specialists from other time zones.
Skills Required:
- Background in Computer Science equivalent to a B.Sc. Equivalent practical experience is a reasonable substitute.
- Must have experience with kubernetes and management of containerized applications
- Automation-related experience is particularly valued using scripting languages such as python, bash, Perl. One higher level language is desired.
- Experience on supporting three tier architecture which includes exposure to UNIX, Linux platforms and databases such IBM DB2, Sybase, Mongo, GreenPlum etc.
- Experience with source code and binary repositories, build tools, and CI/CD (Git, Artifactory, Jenkins, Docker) etc and data streaming technologies like Spark, Kafka etc.
- Hands on experience on enterprise tools set such as Grafana, Dynatrace, AppDynamics etc.
- Awareness of, and ability to reason about modern software & systems architectures, including load-balancing, queueing, caching, distributed systems failure modes, micro services etc
- Deep understanding of operating system level concepts such as processes, memory allocation, and the network stack; understanding of how applications are affected by the above, and ability to debug same.
- Generally speaking, practical experience running large scale online systems is always an advantage.
At mthree, our values support courageous teammates, needle movers, and learning champions all while striving to support the health and well-being of all employees. We take great pride in celebrating the diversity of each individual who contributes to making mthree the company it is today and will be in the future. We value diversity both within mthree and with our partner companies, and we're proud to provide an environment where all our colleagues can flourish. That means promoting a strong culture of equality but, most importantly, inclusion.
Applicants must be currently authorized to work in Canada on a full-time basis. The Company will not sponsor applicants for work visas.
- ...in AUM and 22 global investment banks. For more information, please visit . The Role CMG is looking for a Site Reliability Engineer (SRE) with a strong focus on monitoring, observability, and alerting to ensure the reliability, performance, and scalability...SuggestedRemote jobFull timeLocal area
- ...environments. We are a modern, IoT-enabled, cloud-based tool for reliability, safety, and operations of physical equipment and... ...valuing the company at $2.5 billion. We’re looking for a Site Reliability Engineer (SRE) to help advance MaintainX’s reliability, observability,...SuggestedFull time
$101.2k - $136.9k per year
...strategy across digital banking, core banking, data platforms and member facing services. We are looking for a highly skilled Site Reliability Engineer (SRE) who will help build, operate and continuously improve the reliability, performance, security and automation of our cloud...SuggestedPermanent employmentFull timeInternshipWork at officeImmediate startHome officeFlexible hours2 days per week- ...read on to find out more. The role: We’re looking for a Site Reliability Engineer to manage, maintain, improve and provide support on our... ...operations and support Writing and maintaining documentation on SRE processes and policies Recommending and implementing ways...SuggestedFull timeRemote work
$120k - $200k per year
...Coinbase Ventures, Uniswap Labs, Circle Ventures, Delphi Digital, and many more. ABOUT THE ROLE At LayerZero, our Site Reliability Engineering (SRE) team is at the intersection of software and systems engineering, dedicated to crafting and maintaining large-scale,...SuggestedFull time- ...a good fit for you! About the Role We are looking for a Site Reliability Engineer to join our Network and Security Operations Center (NOC), a team... ...and maintain technical documentation and contribute to SRE best practices Partner with platform engineering, deployment...Long term contractPermanent employmentFull timeWork at officeRemote work
$110k - $160k per year
...hear from you! Role Overview We're seeking a Senior Site Reliability Engineer to join our SaaS-Ops team within Shared Services Engineering.... ...that demonstrates growing depth in cloud infrastructure and SRE practices; ~ Managed production Kubernetes environments at...Full timeInternshipWork at officeLocal areaFlexible hours- ...We are seeking a Senior DevOps & Site Reliability Engineer to own the reliability, scalability, performance, and operational excellence of Medeloop... ...pipelines, infrastructure as code, and cloud architecture—with SRE discipline: SLOs, incident management, capacity planning,...Hourly payFull time
- ...and we are not tiptoeing into it. We are rebuilding our engineering culture around a simple belief: AI changes everything.... ...on — and this role is at the center of keeping it reliable, fast, and scalable. As a Staff SRE, you'll own the infrastructure and reliability practices...Remote jobFull timeInternshipWork at officeLocal areaFlexible hoursShift workWeekend work
$120k - $160k per year
...ABOUT YOU We are looking for a Site Reliability Engineer (Monetization) who is pragmatic, product-minded, and equally comfortable writing code... ...production systems to join our Infrastructure department's SRE team . The best candidate will be someone who thrives in a...Long term contractFull time- ...America, Europe, Australia, and Japan. As one of our Lead Site Reliability Engineers, you will combine hands-on technical expertise with strategic... ...scenarios Lead cross-functional reliability initiatives with SRE and service engineering teams, influencing architectural...Long term contractFull time
- ...checking the market? Well… you might just be in the right place! We’re looking for a Staff Site Reliability Engineer to join our Data team in Canada. As a Staff Data SRE, you are the technical backbone of the Data Office's infrastructure platform. Your scope spans...Full timeWork at officeRemote workFlexible hoursShift work
- ...Gauss Labs is seeking a highly skilled Site Reliability Engineer to join our team in Vancouver. As an SRE at Gauss Labs, you will play a critical role in ensuring our industrial AI platform's reliability, performance, and scalability. You will be responsible for building and...Full time
$123k - $160k per year
...your place here. We are seeking a highly experienced Senior Site Reliability Engineer to own the reliability, performance, and operational... ...environments. You’ll be a good fit if you have: ~6+ years in SRE, systems engineering, or software engineering roles, ideally...Remote jobLong term contractFull timeRelocation- ...The Site Reliability Engineering organization at Pinterest is accountable for ensuring overall Pinterest availability as well as enhancing Engineering... ...as Pinterest continues to grow and scale. As a Pinterest SRE, you will design and build systems, platforms, tools, frameworks...Full timeWork at officeRelocationRelocation package
- ...une grande organisation mondiale. Les ingénieurs prospèrent chez Tower tout en... ...the world’s best systematic trading and engineering talent. We empower portfolio managers to... ...the operational burden low and to enhance reliability Monitoring daily reports and run of scheduled...Daily paidApprenticeshipCasual workWork at officeWorldwideRotating shift
$180.4k - $230.4k per year
...Coalition. About the role We are looking for a Staff Site Reliability Engineer to lead AI enablement across our engineering organization. As... ...foundations trustworthy. This role sits within our Platform SRE team, and you'll participate in the team's ad-hoc support...Full timeRemote workHome officeFlexible hoursShift work$145k - $185k per year
...makes simulation at scale possible. We're hiring a Senior Site Reliability Engineer to help build and operate that infrastructure. This role sits... ...Required Qualifications Experience. 5+ years in SRE, DevOps, or infrastructure engineering roles, with a track record...Remote jobFull time$197.5k - $225k per year
...GV and Riverwood Capital. About the Team: As a Senior Site Reliability Engineer, you will be a key technical leader driving the design and optimization... .... Required Qualifications: ~6+ years in SRE , DevOps , or Infrastructure roles , with significant production...Full time- ...in 2024 – but we're just getting started. As a Sr. Site Reliability Engineer, you'll be the guardian of our platform's reliability and performance... ...continuous improvement across our engineering teams. Our SRE Team: We're a bottom-up, collaborative team that thrives...Full timeWork at officeLocal areaRemote workWork from homeHome officeWeekend work
- ...corporate culture by MSP Today, G2, and Great Place to Work™, to name a few. About the role We’re looking for a Senior Site Reliability Engineer (SRE) to help strengthen and scale our multi-cloud platform and developer experience. This is a hands-on senior individual...Full timeInternshipRemote workWork from home
$150k - $240k per year
...As a contributor in the SRE organization, you are passionate... ...achieving the high quality and reliability our customers demand. You will... ...deliverables will reach the entire engineering organization to enable product... .... Exemplify cloud-native site reliability best practices....Long term contractFull timeRemote work- ...visualizing relationships between entities in the system. As a Site Reliability Engineer you will be responsible for the availability, latency,... ...DevOps, Product and Engineering teams to design and implement SRE practice at Behavox to build foundational infrastructure allowing...Full timeFlexible hours
$153k - $187k per year
...clear signal owners can use to run stronger, more resilient businesses. We’re looking for an incredible Senior Site Reliability Engineer to join our SRE team. We aim to make reliability, security, and speed reinforce one another so that the platform becomes the engine...Full timeInternship- ...senior contributor in the APX SRE organization, you are passionate... ...about the high quality, reliability, and security our customers demand... ...deliverables will reach the entire engineering organization to enable product... .... Exemplify cloud-native site reliability best practices...Long term contractFull timeRemote workFlexible hours
$69k - $90k per year
...accolades from global entities like the World Summit Awards and the UN-endorsed Zero Project. About the role As a Junior Site Reliability Engineer (SRE) at Fable, you will help support the reliability, performance, and scalability of the systems that power our products....Full timeInternship$20 per day
...our careers page to see how you can grow with us! As a Site Reliability Engineer at Hiive, you will be responsible for ensuring the... ...performance and system behavior, and ensuring these services are reliable, scalable, and cost-efficient in production. In this role...Full timeSummer holidayRelocation$140k - $180k per year
...the world and our goal is preserving uncensored Internet access and online privacy for all. Right now we are looking for a Site Reliability Engineer to help us tame DNS. About the Position Linux system administration and troubleshooting Network configuration and...Full timeDirect hireWork at officeRemote work- ...a part of our journey! About the role We are committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building and leading processes to ensure the reliability,...Full timeLocal areaRemote workHome officeFlexible hours
$260k - $275k per year
...enterprises • Solve complex reliability challenges at scale • Influence architecture and engineering culture at a company level •... ...will focus on creating reusable, reliable, and scalable solutions that abstract... ..., Platform Engineering, or Site Reliability Engineering role,...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Ingénieur·e SRE / Site Reliability Engineer. Be the first to apply!
- ingénieur bâtiment Remote
- construction engineering technology Remote
- building operating engineer Remote
- construction engineer Remote
- ingénieur de chantier Remote
- construction civil engineering Remote
- site reliability engineer intern Remote
- site reliability engineer remote Remote
- site reliability engineer sre Remote
- site reliability engineer Remote
