Tyk Logo

Tyk

Site Reliability Engineer - AMER

Reposted 2 Months Ago
Remote
Hiring Remotely in Canada
Senior level
Remote
Hiring Remotely in Canada
Senior level
Operate, maintain and improve the global Tyk Cloud platform: run production Kubernetes clusters, manage cloud infrastructure, automate operations, run on-call incident response, create monitoring and dashboards, conduct post-incident analysis, document SRE processes, and drive reliability, efficiency and multi-region/multi-cloud expansion.
The summary above was generated by AI

Who are Tyk, and what do we do? 

The Tyk API Management platform is helping to drive the connected world and power new products and services. We’re changing the way that organisations connect any number of their systems and services.Whether internal, external, public or highly encrypted systems, Tyk helps businesses drive value across the retail, finance, telecoms, healthcare, or media industries (to name just a few!)

If you’ve banked online, used an app to check the news, or perhaps even driven a connected car, API’s, and by extension, Tyk, make that possible. Founded in 2015 with offices in London – UK, London – Ontario, Atlanta and Singapore, we have many thousands of users of our B2B platform across the globe. Brands using Tyk range from Lotte, Bell, T Mobile, to RBS, Capital One and Vinci. We have a varied user base hailing from every continent – even Antarctica.

Our Mission

Tyk is on a mission to connect every system in the world. We’ve started by building an API Management platform.

Total flexibility, default remote, radical responsibility

We offer unlimited paid holidays and remote working from anywhere in the world, for everyone, Why? Tyk was founded on the principle of offering flexibility and autonomy to our employees, we believe this allows our employees to achieve their best results. It also means we can build the best possible team, location and working hours are no barrier. 

If this sounds like an environment that you believe could work for you then read on to find out more.

The role:

Tyk Cloud is our managed API management platform, running on multi-region Kubernetes at scale for customers around the world.

We're looking for an SRE who's as comfortable in the code as in the infrastructure. You'll spend most of your time improving and automating the platform, and when you're on call, you'll handle incidents independently. You'll join a small, centralised SRE team that works closely with our product teams.

You don't need to have done everything below. We care most about how you reason through problems and how quickly you learn. You'll have three to four months to get up to speed, shadowing first, before you go on call.

What you'll do:

Most of your time

  • Deliver the team's planned work each quarter, such as optimising the platform, building self-serve tooling for other teams, and rearchitecting parts of the platform as it grows.
  • Help expand Tyk Cloud across regions and clouds, and bring down what it costs to run.
  • Automate operations in Go, including building and maintaining our custom Kubernetes operators.
  • Run the platform's services and databases, including MongoDB and Redis.
  • Improve our observability: find the metrics that matter, and build the dashboards and alerts to act on them.
  • Keep runbooks and documentation current, and support security work such as SOC 2 audits.

When you're on call

You'll be doing one week in three initially (one in four as we grow), Monday-Friday, on a 12-hour shift with secondary backup support. Rotas: 14:00–02:00 UTC

  • Be first line for platform alerts and incidents: restore service, escalate or help fix product bugs, and lead post-incident reviews.
  • Act as second line for our Customer Success team, on requests that come directly from customers.
  • Handle ad hoc requests from other teams across the organisation regarding Tyk Cloud.

RequirementsWhat you'll need:
  • 3+ years in SRE, platform or infrastructure roles, across more than one company or production platform.
  • Experience owning on-call and leading incidents yourself.
  • Hands-on experience running production Kubernetes at scale, ideally EKS: operating, upgrading and debugging large, multi-tenant clusters.
  • Experience designing and operating infrastructure on AWS, with Terraform or similar.
  • The ability to write, test and ship Go tooling or services.
  • Experience with Prometheus and Grafana, and with logging systems.
  • Solid Linux and networking fundamentals (DNS, TCP/IP, HTTP, TLS, load balancing).
  • Clear communication across time zones and teams.
Our stack

EKS, Terraform/Terragrunt, Helm, GitHub Actions, Argo CD, MongoDB, Redis, Prometheus and Grafana.


BenefitsHere’s why you should join us:
  • Everyone has unlimited paid holiday. 
  • We have total flexibility in hours, as we believe creativity flows better when our people are given freedom to decide when they are most productive. Everyone is unique after all.
  • Employee share scheme
  • Generous maternity and paternity leave
  • Company retreats

We all share the same vision – we value authenticity, respect, responsibility, independence, honesty, diversity and inclusion and most importantly treating others how you wish to be treated. We look for like-minded people who bring their personalities to work everyday, strive to achieve their personal goals and who are willing to challenge the way we do things, why? – to make what we do even better!

Our values tell the story of Tyk – here’s how:

  • It’s ok to screw up! 

We’ve found that it’s often the ‘stupid’ or unexpected ideas that turn out to be the successful ones – so try it, at least we can say we have!

  • The only stupid idea, is the untested one! 

It’s in our DNA – starting a business with founders 12 hours apart, giving our gateway away for free – sure, we did that, and we’d do it again!

  • Trust starts with you – make it count! 

Trust is a two-way street – instill it from day one!

  • Assume best intent! 

We have each other’s back – we’re all on the same team. Think before you speak or act. 

  • Make things, better! 

Always try to leave things better than when you found them – change is constant, inevitable and embraced! Be that change we want to see.

What’s it like to work here?! check it out: https://tyk.io/worklife/

Tyk is an equal opportunities employer and we are determined to ensure that no applicant or employee receives less favourable treatment on the grounds of gender, age, disability, religion, belief, sexual orientation, marital status, or race, or is disadvantaged by conditions or requirements which cannot be shown to be justifiable.

You can see more about us here https://tyk.io

Similar Jobs

One Month Ago
Easy Apply
Remote
Canada
Easy Apply
Senior level
Senior level
Cloud • Security • Software • Cybersecurity • Automation
Maintain and improve reliability, scalability, and automation for user-facing production systems. Build infrastructure tooling, operate Kubernetes-based services, write IaC, participate in on-call and incident response, and advance observability and runbooks to reduce toil and improve platform reliability.
Top Skills: AWSCi/CdGCPGitopsGoInfrastructure As Code (Iac)KubernetesKubernetes Operators/ControllersLoggingMetricsRubySlos/SlisTerraform
15 Minutes Ago
Easy Apply
Remote
Canada
Easy Apply
Expert/Leader
Expert/Leader
Cloud • Security • Software • Cybersecurity • Automation
Own GitLab’s competitive positioning for AI and agentic software development. Analyze competitors and market signals, shape product roadmap recommendations, develop sales and executive enablement assets, lead displacement campaigns, and brief analysts, investors, and executives. Establish thought leadership, define competitive narratives, measure win-rate impact, and elevate competitive marketing practices across the team.
Top Skills: Agentic Software DevelopmentAIAi Productivity ToolsDevsecopsGitlabSoftware Development Lifecycle (Sdlc)
15 Minutes Ago
Easy Apply
Remote
Canada
Easy Apply
Entry level
Entry level
Cloud • Security • Software • Cybersecurity • Automation
Manages GitLab for Nonprofits and GitLab for Education programs from strategy through execution. Responsibilities include building nonprofit and university partnerships, overseeing operations, improving processes, supporting technical implementation, managing communications and inquiries, tracking program performance, collaborating with Sales and Marketing, creating advocacy and storytelling resources, and researching social impact best practices.
Top Skills: GitlabSalesforce

What you need to know about the Vancouver Tech Scene

Raincouver, Vancity, The Big Smoke — Vancouver is known by many names, and in recent years, it has gained a reputation as a growing hub for both tech and sustainability. Renowned for its natural beauty, the city has become a magnet for professionals eager to create environmental solutions, and with an emphasis on clean technology, renewable energy and environmental innovation, it's attracted companies across various industries, all working toward a shared goal: advancing clean technology.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account