Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

Full-time

Tyk Technologies Limited

Description

Who are Tyk, and what do we do? 

The Tyk API Management platform is helping to drive the connected world and power new products and services. We’re changing the way that organisations connect any number of their systems and services.Whether internal, external, public or highly encrypted systems, Tyk helps businesses drive value across the retail, finance, telecoms, healthcare, or media industries (to name just a few!)

If you’ve banked online, used an app to check the news, or perhaps even driven a connected car, API’s, and by extension, Tyk, make that possible. Founded in 2015 with offices in London – UK, London – Ontario, Atlanta and Singapore, we have many thousands of users of our B2B platform across the globe. Brands using Tyk range from Lotte, Bell, T Mobile, to RBS, Capital One and Vinci. We have a varied user base hailing from every continent – even Antarctica.

Our Mission

Tyk is on a mission to connect every system in the world. We’ve started by building an API Management platform.

Total flexibility, default remote, radical responsibility

We offer  unlimited paid holidays and remote working from anywhere in the world , for everyone, Why? Tyk was founded on the principle of offering flexibility and autonomy to our employees, we believe this allows our employees to achieve their best results. It also means we can build the best possible team, location and working hours are no barrier. 

If this sounds like an environment that you believe could work for you then read on to find out more.

The role:

We’re looking for a Site Reliability Engineer to manage, maintain, improve and provide support on our platform. You will be curious by nature, always looking for ways to improve, as we will look to you for new ideas, solutions and metrics on how we can improve the platform. You will also be our first line of incident management to our clients and will help define our response going forward. This is a great opportunity to become an integral part of Tyk as we continue on our journey.

As a remote first company, you will have the opportunity to work with an industry leading distributed team. Having access to expertise from across the globe will give you both the support and opportunity to help shape not only Tyk’s Cloud platform but also the Tyk as a whole as we continue to grow.

Requirements

Here’s what you’ll be responsible for:

  • Maintaining global Tyk Cloud within SL(A/I/O)s you will help to define
  • Identifying reliability issues and working together with your squad to solve them
  • Identifying and introducing new metrics and building relevant dashboards
  • Participating in the on-call rotation
  • Working with your squad to expand multi-region and multi-cloud reach of the platform
  • Documenting operational knowledge
  • Conducting post-incident analysis
  • Automating common tasks
  • Be a key shaper and contributor to our continuous improvement agenda – be it the clarity of our user stories, how we estimate, communicate with other teams or customers – we expect this role to be advocate of continuous improvement
  • Reliability of our new global Tyk Cloud platform
  • Automation of operations and support
  • Writing and maintaining documentation on SRE processes and policies
  • Recommending and implementing ways of driving operational efficiency and driving down our cost to run, without impacting service
  • Assisting in penetration testing for Cloud through liaising with our provider, providing technical details, and environment setup
  • Incident management

Here’s what we’re looking for:

Experience

  • Strong collaboration skills
  • Launching and operating production scale kubernetes clusters
  • Designing and operating infrastructure on AWS and other providers
  • Operating MongoDB (or other document database) clusters
  • Operating Redis (or other key-value storage) clusters
  • Administering Linux servers
  • Maintaining distributed software
  • Operating Prometheus and Grafana
  • Operating logging collection and analysis systems
  • Participating in the on-call rotation(16:00pm – 4:00am UTC)

Skills:

  • Kubernetes & containers (advanced)
  • AWS / EKS (advanced)
  • Linux (advanced)
  • Terraform and IaC in general (proficient)
  • Helm (proficient)
  • Go (familiar)
  • MongoDB (or similar)
  • Redis (or similar)
  • Monitoring – prometheus, grafana, thanos (familiar)
  • Grasp of networking concepts (subnets, routing, peering, load balancing, NAT, etc.)
  • Common networking protocols (DNS, TCP/IP, TLS, UDP)
  • Proactive, energetic, innovative and change oriented

Nice to have:

  • GCP or Azure
  • Bare metal infrastructure engineering
  • API management experience
  • Large scale distributed storage management
  • Familiarity with Rancher
  • CKA/CKAD/CKS
  • Creating and delivering production software in Go language

Benefits

Here’s why you should join us:

  • Everyone  has unlimited paid holiday. 
  • We have total flexibility in hours, as we believe creativity flows better when our people are given freedom to decide when they are most productive. Everyone is unique after all.
  • Employee share scheme
  • Generous maternity and paternity leave
  • Company retreats

We all share the same vision – we value authenticity, respect, responsibility, independence, honesty, diversity and inclusion and most importantly treating others how you wish to be treated. We look for like-minded people who bring their personalities to work everyday, strive to achieve their personal goals and who are willing to challenge the way we do things, why? – to make what we do even better!

Our values tell the story of Tyk – here’s how:

  • It’s ok to screw up!  

We’ve found that it’s often the ‘stupid’ or unexpected ideas that turn out to be the successful ones – so try it, at least we can say we have!

  • The only stupid idea, is the untested one! 

It’s in our DNA – starting a business with founders 12 hours apart, giving our gateway away for free – sure, we did that, and we’d do it again!

  • Trust starts with you – make it count! 

Trust is a two-way street – instill it from day one!

  • Assume best intent! 

We have each other’s back – we’re all on the same team. Think before you speak or act. 

  • Make things, better! 

Always try to leave things better than when you found them – change is constant, inevitable and embraced! Be that change we want to see.

What’s it like to work here?! check it out: 

Tyk is an equal opportunities employer and we are determined to ensure that no applicant or employee receives less favourable treatment on the grounds of gender, age, disability, religion, belief, sexual orientation, marital status, or race, or is disadvantaged by conditions or requirements which cannot be shown to be justifiable.

You can see more about us here 

Vacancy posted 6 hours ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Remote vacancy
  •  ...trillion in AUM and 22 global investment banks. For more information, please visit .     The Role     CMG is looking for a Site Reliability Engineer (SRE) with a strong focus on monitoring, observability, and alerting to ensure the reliability, performance, and... 
    Suggested
    Remote job
    Full time
    Local area

    Capital Markets Gateway

    Remote
    6 hours ago
  • $101.2k - $136.9k per year

     ...strategy across digital banking, core banking, data platforms and member facing services. We are looking for a highly skilled Site Reliability Engineer (SRE) who will help build, operate and continuously improve the reliability, performance, security and automation of our... 
    Suggested
    Permanent employment
    Full time
    Internship
    Work at office
    Immediate start
    Home office
    Flexible hours
    2 days per week

    Vancity

    Remote
    6 hours ago
  •  ...environments. We are a modern, IoT-enabled, cloud-based tool for reliability, safety, and operations of physical equipment and...  ...valuing the company at $2.5 billion. We’re looking for a Site Reliability Engineer (SRE) to help advance MaintainX’s reliability, observability... 
    Suggested
    Full time

    Maintainx

    Remote
    7 hours ago
  • $120k - $200k per year

     ...and many more.   ABOUT THE ROLE At LayerZero, our Site Reliability Engineering (SRE) team is at the intersection of software and systems engineering...  ...internal systems to those external users interact with—are reliable, meet the uptime expectations of our users, and... 
    Suggested
    Full time

    Layer Zero Labs Llc

    Remote
    6 hours ago
  •  ...challenges with continuous learning opportunities, then Tescys could be a good fit for you! About the Role We are looking for a Site Reliability Engineer to join our Network and Security Operations Center (NOC), a team at the heart of platform reliability for mission-critical... 
    Suggested
    Long term contract
    Permanent employment
    Full time
    Work at office
    Remote work

    Tecsys Inc.

    Remote
    6 hours ago
  • $20 per day

     ...our careers page to see how you can grow with us! As a Site Reliability Engineer at Hiive, you will be responsible for ensuring the...  ...performance and system behavior, and ensuring these services are reliable, scalable, and cost-efficient in production. In this role... 
    Full time
    Summer holiday
    Relocation

    Hiive

    Remote
    6 hours ago
  • $110k - $160k per year

     ...to join our team working towards this goal, we would love to hear from you!  Role Overview  We're seeking a Senior Site Reliability Engineer to join our SaaS-Ops team within Shared Services Engineering. The team owns reliability and operational excellence for our highly... 
    Full time
    Internship
    Work at office
    Local area
    Flexible hours

    Magnet Forensics

    Remote
    6 hours ago
  •  ...Gauss Labs is seeking a highly skilled Site Reliability Engineer to join our team in Vancouver. As an SRE at Gauss Labs, you will play a critical role in ensuring our industrial AI platform's reliability, performance, and scalability. You will be responsible for building and... 
    Full time

    Gauss Labs

    Remote
    6 hours ago
  •  ...made, and we are not tiptoeing into it. We are rebuilding our engineering culture around a simple belief: AI changes everything. How teams...  ...team builds on — and this role is at the center of keeping it reliable, fast, and scalable. As a Staff SRE, you'll own the infrastructure... 
    Remote job
    Full time
    Internship
    Work at office
    Local area
    Flexible hours
    Shift work
    Weekend work

    Babylist, Inc

    Remote
    7 hours ago
  • $120k - $160k per year

     ...ABOUT YOU We are looking for a Site Reliability Engineer (Monetization) who is pragmatic, product-minded, and equally comfortable writing...  ...developers ship features and what infrastructure needs to stay reliable - and to bring the reliability lens into design decisions... 
    Long term contract
    Full time

    Xsolla

    Remote
    6 hours ago
  •  ...for a new opportunity? Or just checking the market? Well… you might just be in the right place! We’re looking for a Staff Site Reliability Engineer to join our Data team in Canada. As a Staff Data SRE, you are the technical backbone of the Data Office's... 
    Full time
    Work at office
    Remote work
    Flexible hours
    Shift work

    Lightspeed Commerce

    Remote
    6 hours ago
  •  ...client base with operations throughout North America, Central America, Europe, Australia, and Japan. As one of our Lead Site Reliability Engineers, you will combine hands-on technical expertise with strategic technical leadership across infrastructure and software development... 
    Long term contract
    Full time

    Movable Ink

    Remote
    6 hours ago
  •  ...We are seeking a Senior DevOps & Site Reliability Engineer to own the reliability, scalability, performance, and operational excellence of Medeloop’s platform. This role blends deep DevOps engineering—CI/CD pipelines, infrastructure as code, and cloud architecture—with SRE... 
    Hourly pay
    Full time

    Medeloop

    Remote
    7 hours ago
  • $123k - $160k per year

     ...your place here. We are seeking a highly experienced Senior Site Reliability Engineer to own the reliability, performance, and operational...  ...Kotlin, Bash, or similar languages, with an emphasis on building reliable automation and tooling. ~ Hands-on experience with modern... 
    Remote job
    Long term contract
    Full time
    Relocation

    Branch Metrics

    Remote
    6 hours ago
  •  ...The Site Reliability Engineering organization at Pinterest is accountable for ensuring overall Pinterest availability as well as enhancing Engineering teams’ capability to design, build and operate robust systems at scale. Pinterest’s applications and infrastructure that... 
    Full time
    Work at office
    Relocation
    Relocation package

    Pinterest

    Remote
    6 hours ago
  • $140k - $180k per year

     ...the world and our goal is preserving uncensored Internet access and online privacy for all. Right now we are looking for a Site Reliability Engineer to help us tame DNS. About the Position Linux system administration and troubleshooting Network configuration and... 
    Full time
    Direct hire
    Work at office
    Remote work

    Funded.club

    Remote
    6 hours ago
  •  ...you’re relentless about the high quality, reliability, and security our customers demand. You...  ...technical deliverables will reach the entire engineering organization to enable product teams to...  ...confidence. Exemplify cloud-native site reliability best practices with a strong... 
    Long term contract
    Full time
    Remote work
    Flexible hours

    Axon

    Remote
    6 hours ago
  • $69k - $90k per year

     ...UN-endorsed Zero Project.  About the role  As a Junior Site Reliability Engineer (SRE) at Fable, you will help support the reliability,...  ...build more accessible digital experiences, and maintaining reliable, high-performing systems is critical to delivering that impact... 
    Full time
    Internship

    Fable

    Remote
    7 hours ago
  •  ...corporate culture by MSP Today, G2, and Great Place to Work™, to name a few.  About the role We’re looking for a Senior Site Reliability Engineer (SRE) to help strengthen and scale our multi-cloud platform and developer experience. This is a hands-on senior individual... 
    Full time
    Internship
    Remote work
    Work from home

    Scalepad

    Remote
    7 hours ago
  •  ...a part of our journey! About the role We are committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building and leading processes to ensure the reliability,... 
    Full time
    Local area
    Remote work
    Home office
    Flexible hours

    Clickhouse

    Remote
    6 hours ago
  • $260k - $275k per year

     ...enterprises • Solve complex reliability challenges at scale • Influence architecture and engineering culture at a company level •...  ...will focus on creating reusable, reliable, and scalable solutions that abstract...  ..., Platform Engineering, or Site Reliability Engineering role,... 
    Full time

    Saviynt

    Remote
    6 hours ago
  • $150k - $240k per year

     ...obsessed about achieving the high quality and reliability our customers demand. You will work...  ...technical deliverables will reach the entire engineering organization to enable product teams to...  ...-effective. Exemplify cloud-native site reliability best practices. Write code... 
    Long term contract
    Full time
    Remote work

    Axon

    Remote
    6 hours ago
  •  ...through millions of data items, by searching, filtering, and visualizing relationships between entities in the system. As a Site Reliability Engineer you will be responsible for the availability, latency, performance, efficiency, change management, monitoring, emergency... 
    Full time
    Flexible hours

    Behavox

    Remote
    6 hours ago
  • $153k - $187k per year

     ...resilient businesses. We’re looking for an incredible Senior Site Reliability Engineer to join our SRE team. We aim to make reliability, security,...  ...reinforce one another so that the platform becomes the engine of Relay’s growth. Your love of making high-impact decisions... 
    Full time
    Internship

    Relay

    Remote
    6 hours ago
  • $197.5k - $225k per year

     ...investors including Silver Lake Waterman, Moody’s, Sequoia Capital, GV and Riverwood Capital. About the Team: As a Senior Site Reliability Engineer, you will be a key technical leader driving the design and optimization of our Kubernetes-based infrastructure and CI/CD... 
    Full time

    Securityscorecard

    Remote
    6 hours ago
  •  ...in 2024 – but we're just getting started.      As a Sr. Site Reliability Engineer, you'll be the guardian of our platform's reliability and performance...  .... What You Bring to the Team: Design and implement reliable and scalable AWS architecture to meet the needs of the... 
    Full time
    Work at office
    Local area
    Remote work
    Work from home
    Home office
    Weekend work

    Third-Party Job Posts

    Remote
    6 hours ago
  • $180.4k - $230.4k per year

     ...Coalition. About the role We are looking for a Staff Site Reliability Engineer to lead AI enablement across our engineering organization....  ...and tooling infrastructure to ensure AI-generated output is reliable, secure, and production-worthy. This role owns that layer.... 
    Full time
    Remote work
    Home office
    Flexible hours
    Shift work

    Coalition

    Remote
    7 hours ago
  • $145k - $185k per year

     ...environments, and the infrastructure underneath that platform is what makes simulation at scale possible. We're hiring a Senior Site Reliability Engineer to help build and operate that infrastructure. This role sits at the core of how we run large-scale, distributed simulation... 
    Remote job
    Full time

    Parallel Domain

    Remote
    6 hours ago
  •  ...une équipe dynamique en tant qu’ Ingénieur·e Fiabilité de Site (Site Reliability Engineer) pour l’un de nos clients. Le Site Reliability Engineering...  ...systems and working with high scale scalable and reliable services. Like to work in a fast-moving environment and... 
    Full time
    Apprenticeship
    Local area
    Worldwide

    Mthree Recruiting Portal

    Remote
    6 hours ago
  •  ...overview of this role You'll join the Dedicated team as a Site Reliability Engineer focused on Environment Automation , where your work will...  .... In this role, you'll help keep these environments reliable, scalable, secure, and consistent by treating everything as... 
    Full time
    Remote work
    Home office

    Gitlab

    Remote
    6 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!