Company: TYK TECHNOLOGIES LIMITED

Apply for the Remote Site Reliability Engineer — Global Cloud Platform

Location: London

Job Description:

The Tyk API Management platform is helping to drive the connected world and power new products and services. We’re changing the way that organisations connect any number of their systems and services.Whether internal, external, public or highly encrypted systems, Tyk helps businesses drive value across the retail, finance, telecoms, healthcare, or media industries (to name just a few!)

If you’ve banked online, used an app to check the news, or perhaps even driven a connected car, API’s, and by extension, Tyk, make that possible. Founded in 2015 with offices in London – UK, London – Ontario, Atlanta and Singapore, we have many thousands of users of our B2B platform across the globe. Brands using Tyk range from Lotte, Bell, T Mobile, to RBS, Capital One and Vinci. We have a varied user base hailing from every continent – even Antarctica.

Our Mission

Tyk is on a mission to connect every system in the world. We’ve started by building an API Management platform.

Total flexibility, default remote, radical responsibility

We offer unlimited paid holidays and remote working from anywhere in the world, for everyone, Why? Tyk was founded on the principle of offering flexibility and autonomy to our employees, we believe this allows our employees to achieve their best results. It also means we can build the best possible team, location and working hours are no barrier.

The role

We’re looking for a Site Reliability Engineer to manage, maintain, improve and provide support on our platform. You will be curious by nature, always looking for ways to improve, as we will look to you for new ideas, solutions and metrics on how we can improve the platform. You will also be our first line of incident management to our clients and will help define our response going forward. This is a great opportunity to become an integral part of Tyk as we continue on our journey.

Here’s what you’ll be responsible for

Maintaining global Tyk Cloud within SL(A/I/O)s you will help to define
Identifying reliability issues and working together with your squad to solve them
Identifying and introducing new metrics and building relevant dashboards
Participating in the on-call rotation
Working with your squad to expand multi-region and multi-cloud reach of the platform
Documenting operational knowledge
Conducting post-incident analysis
Be a key shaper and contributor to our continuous improvement agenda – be it the clarity of our user stories, how we estimate, communicate with other teams or customers – we expect this role to be advocate of continuous improvement
Reliability of our new global Tyk Cloud platform
Automation of operations and support
Writing and maintaining documentation on SRE processes and policies
Recommending and implementing ways of driving operational efficiency and driving down our cost to run, without impacting service
Assisting in penetration testing for Cloud through liaising with our provider, providing technical details, and environment setup

Here’s what we’re looking for

Experience

Launching and operating production scale kubernetes clusters
Designing and operating infrastructure on AWS and other providers
Operating MongoDB (or other document database) clusters
Operating Redis (or other key-value storage) clusters
Operating Prometheus and Grafana
Operating logging collection and analysis systems
Participating in the on-call rotation(16:00pm – 4:00am UTC)

Skills

AWS / EKS (advanced)
Terraform and IaC in general (proficient)
Helm (proficient)
MongoDB (or similar)
Redis (or similar)
Monitoring – prometheus, grafana, thanos (familiar)
Grasp of networking concepts (subnets, routing, peering, load balancing, NAT, etc.)
Common networking protocols (DNS, TCP/IP, HTTP, TLS, UDP)
Proactive, energetic, innovative and change oriented

Nice to have

Bare metal infrastructure engineering
Familiarity with Rancher
CKA/CKAD/CKS
Creating and delivering production software in Go language

Here’s why you should join us

Everyone has unlimited paid holiday.
We have total flexibility in hours, as we believe creativity flows better when our people are given freedom to decide when they are most productive. Everyone is unique after all.
Employee share scheme
Generous maternity and paternity leave
Company retreats

Equal Opportunity Statement

Tyk is an equal opportunities employer and we are determined to ensure that no applicant or employee receives less favourable treatment on the grounds of gender, age, disability, religion, belief, sexual orientation, marital status, or race, or is disadvantaged by conditions or requirements which cannot be shown to be justifiable.

#J-18808-Ljbffr…

Posted: May 27th, 2026