Tyk Logo

Tyk

Site Reliability Engineer - AMER

Reposted 2 Months Ago
Remote
Hiring Remotely in Canada
Senior level
Remote
Hiring Remotely in Canada
Senior level
Operate, maintain and improve the global Tyk Cloud platform: run production Kubernetes clusters, manage cloud infrastructure, automate operations, run on-call incident response, create monitoring and dashboards, conduct post-incident analysis, document SRE processes, and drive reliability, efficiency and multi-region/multi-cloud expansion.
The summary above was generated by AI

Who are Tyk, and what do we do? 

The Tyk API Management platform is helping to drive the connected world and power new products and services. We’re changing the way that organisations connect any number of their systems and services.Whether internal, external, public or highly encrypted systems, Tyk helps businesses drive value across the retail, finance, telecoms, healthcare, or media industries (to name just a few!)

If you’ve banked online, used an app to check the news, or perhaps even driven a connected car, API’s, and by extension, Tyk, make that possible. Founded in 2015 with offices in London – UK, London – Ontario, Atlanta and Singapore, we have many thousands of users of our B2B platform across the globe. Brands using Tyk range from Lotte, Bell, T Mobile, to RBS, Capital One and Vinci. We have a varied user base hailing from every continent – even Antarctica.

Our Mission

Tyk is on a mission to connect every system in the world. We’ve started by building an API Management platform.

Total flexibility, default remote, radical responsibility

We offer unlimited paid holidays and remote working from anywhere in the world, for everyone, Why? Tyk was founded on the principle of offering flexibility and autonomy to our employees, we believe this allows our employees to achieve their best results. It also means we can build the best possible team, location and working hours are no barrier. 

If this sounds like an environment that you believe could work for you then read on to find out more.

The role:

Tyk Cloud is our managed API management platform, running on multi-region Kubernetes at scale for customers around the world.

We're looking for an SRE who's as comfortable in the code as in the infrastructure. You'll spend most of your time improving and automating the platform, and when you're on call, you'll handle incidents independently. You'll join a small, centralised SRE team that works closely with our product teams.

You don't need to have done everything below. We care most about how you reason through problems and how quickly you learn. You'll have three to four months to get up to speed, shadowing first, before you go on call.

What you'll do:

Most of your time

  • Deliver the team's planned work each quarter, such as optimising the platform, building self-serve tooling for other teams, and rearchitecting parts of the platform as it grows.
  • Help expand Tyk Cloud across regions and clouds, and bring down what it costs to run.
  • Automate operations in Go, including building and maintaining our custom Kubernetes operators.
  • Run the platform's services and databases, including MongoDB and Redis.
  • Improve our observability: find the metrics that matter, and build the dashboards and alerts to act on them.
  • Keep runbooks and documentation current, and support security work such as SOC 2 audits.

When you're on call

You'll be doing one week in three initially (one in four as we grow), Monday-Friday, on a 12-hour shift with secondary backup support. Rotas: 14:00–02:00 UTC

  • Be first line for platform alerts and incidents: restore service, escalate or help fix product bugs, and lead post-incident reviews.
  • Act as second line for our Customer Success team, on requests that come directly from customers.
  • Handle ad hoc requests from other teams across the organisation regarding Tyk Cloud.

RequirementsWhat you'll need:
  • 3+ years in SRE, platform or infrastructure roles, across more than one company or production platform.
  • Experience owning on-call and leading incidents yourself.
  • Hands-on experience running production Kubernetes at scale, ideally EKS: operating, upgrading and debugging large, multi-tenant clusters.
  • Experience designing and operating infrastructure on AWS, with Terraform or similar.
  • The ability to write, test and ship Go tooling or services.
  • Experience with Prometheus and Grafana, and with logging systems.
  • Solid Linux and networking fundamentals (DNS, TCP/IP, HTTP, TLS, load balancing).
  • Clear communication across time zones and teams.
Our stack

EKS, Terraform/Terragrunt, Helm, GitHub Actions, Argo CD, MongoDB, Redis, Prometheus and Grafana.


BenefitsHere’s why you should join us:
  • Everyone has unlimited paid holiday. 
  • We have total flexibility in hours, as we believe creativity flows better when our people are given freedom to decide when they are most productive. Everyone is unique after all.
  • Employee share scheme
  • Generous maternity and paternity leave
  • Company retreats

We all share the same vision – we value authenticity, respect, responsibility, independence, honesty, diversity and inclusion and most importantly treating others how you wish to be treated. We look for like-minded people who bring their personalities to work everyday, strive to achieve their personal goals and who are willing to challenge the way we do things, why? – to make what we do even better!

Our values tell the story of Tyk – here’s how:

  • It’s ok to screw up! 

We’ve found that it’s often the ‘stupid’ or unexpected ideas that turn out to be the successful ones – so try it, at least we can say we have!

  • The only stupid idea, is the untested one! 

It’s in our DNA – starting a business with founders 12 hours apart, giving our gateway away for free – sure, we did that, and we’d do it again!

  • Trust starts with you – make it count! 

Trust is a two-way street – instill it from day one!

  • Assume best intent! 

We have each other’s back – we’re all on the same team. Think before you speak or act. 

  • Make things, better! 

Always try to leave things better than when you found them – change is constant, inevitable and embraced! Be that change we want to see.

What’s it like to work here?! check it out: https://tyk.io/worklife/

Tyk is an equal opportunities employer and we are determined to ensure that no applicant or employee receives less favourable treatment on the grounds of gender, age, disability, religion, belief, sexual orientation, marital status, or race, or is disadvantaged by conditions or requirements which cannot be shown to be justifiable.

You can see more about us here https://tyk.io

Similar Jobs

One Month Ago
Easy Apply
Remote
Canada
Easy Apply
Senior level
Senior level
Cloud • Security • Software • Cybersecurity • Automation
Maintain and improve reliability, scalability, and automation for user-facing production systems. Build infrastructure tooling, operate Kubernetes-based services, write IaC, participate in on-call and incident response, and advance observability and runbooks to reduce toil and improve platform reliability.
Top Skills: AWSCi/CdGCPGitopsGoInfrastructure As Code (Iac)KubernetesKubernetes Operators/ControllersLoggingMetricsRubySlos/SlisTerraform
2 Minutes Ago
Remote or Hybrid
Ontario, ON, CAN
Senior level
Senior level
Artificial Intelligence • Cloud • Information Technology • Consulting
Leads complex, cross-functional client management projects from planning through implementation. Manages scope, budgets, financial forecasts, stakeholders, risks, dependencies, vendors, and teams exceeding 20 members across multiple countries. Consolidates workstream information for executive reporting, governance, escalation, and mitigation planning. Provides project management expertise, process improvements, mentoring, and guidance while ensuring delivery meets time, quality, and budget expectations.
An Hour Ago
Remote
Canada
Mid level
Mid level
Artificial Intelligence • Cloud • Consumer Web • Productivity • Software • App development • Data Privacy
Manage adoption, customer health, renewal readiness, and growth across a scaled B2B SaaS portfolio. Design lifecycle campaigns, analyze customer and retention signals, coordinate renewal actions, triage inbound requests, conduct targeted customer conversations, and surface qualified expansion opportunities. Partner cross-functionally with Renewals, Sales, Product, Marketing, Support, and Customer Experience teams while delivering portfolio forecasts and executive updates.
Top Skills: Ai-Enabled SaasB2B SaasBusiness Intelligence ToolsCrm SystemsCustomer Success PlatformsCustomer-Data SystemsMarketing Automation

What you need to know about the Toronto Tech Scene

Although home to some of the biggest names in tech, including Google, Microsoft and Amazon, Toronto has established itself as one of the largest startup ecosystems in the world. And with over 2,000 startups — more than 30 percent of the country's total startups — Toronto continues to attract new businesses. Be it helping entrepreneurs manage their finances, simplifying business operations by automating payroll or assisting pharmaceutical companies in launching new drugs, the city's tech scene is just getting started.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account