Serigor Inc Logo

Serigor Inc

Site Reliability Engineer(SRE)

Reposted One Month Ago
In-Office
Toronto, ON, CAN
Mid level
In-Office
Toronto, ON, CAN
Mid level
Site Reliability Engineers apply software and systems engineering to improve reliability, performance, and operability. They deploy, configure, monitor, recover, and scale services, participate in on-call rotations, evaluate products before and after releases, and spend at least half their time engineering away problems while collaborating with teammates.
The summary above was generated by AI
Company Description

Serigor is all about helping you make the right decision about the right technical support for the right fineness in management utilities at any time in a firm standing. Serigor helps organizations stay ahead by building sustainable competitive advantage.

Job Description

The SRE Role

·         SREs are engineers with the right mix of knowledge and skills in software engineering (i.e. programming, data structures, and algorithms) and systems engineering (i.e. applying scientific principles of experimentation and observation to entire systems to improve reliability, performance and operability).

·         We constantly evaluate products and services before and after production releases to prevent, identify and fix problems that impact service availability in deploying, configuring, monitoring, recovering, and scaling.

·         We participate in on-call rotations to monitor and support our products and services, taking recovery actions prior to and after disruptions.

·         We dedicate at least 50% of our time 'engineering away' problems both, directly and through pairing and coaching our team.

·         We work side-by-side with SREs in our team applying software engineering principles to resolve problems impacting service uptime or our operational efficiency.

Qualifications

Required Core Skills for all SREs

·         Programming in at least one language such as: Java, C#, Javascript, Python or Ruby - experience with other languages is also valuable such as Shell scripting, PowerShell, PERL or PHP.

·         Systems configuration and administration: Windows or Linux.

·         Analyzing and discovering how all components of a distributed system work together using a broad range of skills and tools.

Possess or will learn quickly

·         Applying an evidence based approach to solving system problems under pressure and in real time to provide the fastest path to service recovery.

·         System and software configuration management using tools such as puppet, chef or ansible.

·         Cloud technologies and platforms such as AWS or Azure using API or configuration tools.

Additional Information

All your information will be kept confidential according to EEO guidelines.

Similar Jobs

22 Days Ago
Easy Apply
Hybrid
Toronto, ON, CAN
Easy Apply
Entry level
Entry level
Artificial Intelligence • Cloud • Information Technology • Machine Learning • Software • Big Data Analytics • Automation
Build and operate foundational infrastructure for PagerDuty’s real-time platform, including networking, compute, Kubernetes, and ingress systems. Improve reliability, scalability, and security; monitor system health through metrics, logs, and alerts; participate in 24/7 on-call rotations; support infrastructure rollouts; and contribute to agile planning and technical improvements.
Top Skills: Amazon EksAWSAzureCloudFormationDatadogDnsEnvoyGCPGoGrafanaIstioKubernetesLinuxNew RelicNginxPrometheusPythonRubySplunkSumo LogicTerraformTls
27 Days Ago
Easy Apply
Hybrid
Toronto, ON, CAN
Easy Apply
Senior level
Senior level
Marketing Tech • Mobile • Software
Lead reliability engineering for Braze’s NGINX ingress fleets, Kubernetes infrastructure, Ruby on Rails monolith, and Go API services. Design scalable, highly available systems; develop automated scaling routines; define SLIs, SLOs, and error budgets; conduct capacity planning; and participate in PagerDuty on-call rotations. Drive incident response, root-cause analysis, blameless retrospectives, runbook improvements, and resilient architecture decisions with product engineering teams.
Top Skills: AnsibleAWSAzureChefDatadogGCPGoGrafanaHorizontal Pod Autoscaler (Hpa)JavaKafkaKubernetesLinuxMongoDBNginxPagerdutyPostgresPrometheusPythonRedisRubyTcp/IpTerraformUnix
8 Days Ago
In-Office
Toronto, ON, CAN
Expert/Leader
Expert/Leader
Software
Lead reliability strategy across KEV’s production platforms by defining SLOs, SLIs, and error budgets; owning incident management and blameless postmortems; designing observability, monitoring, logging, and tracing systems; automating operational work; and leading capacity and performance engineering. The role also involves evaluating AI-assisted reliability tools, mentoring engineers, partnering with DevOps and product engineering teams, and communicating reliability trade-offs to stakeholders.
Top Skills: .Net.Net FrameworkIisAzure

What you need to know about the Toronto Tech Scene

Although home to some of the biggest names in tech, including Google, Microsoft and Amazon, Toronto has established itself as one of the largest startup ecosystems in the world. And with over 2,000 startups — more than 30 percent of the country's total startups — Toronto continues to attract new businesses. Be it helping entrepreneurs manage their finances, simplifying business operations by automating payroll or assisting pharmaceutical companies in launching new drugs, the city's tech scene is just getting started.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account