SimCorp Logo

SimCorp

Senior Site Reliability Engineer

Posted Yesterday
Be an Early Applicant
Hybrid
Toronto, ON, CAN
Senior level
Hybrid
Toronto, ON, CAN
Senior level
Owns reliability, monitoring, observability, release management, vulnerability management, and cost management for Azure cloud-native products. Responsibilities include infrastructure automation, deployment pipelines, incident response, capacity planning, SLOs, disaster recovery, configuration management, identity integration, and operational improvements. The role collaborates with engineering teams, clients, vendors, and stakeholders while supporting rotational shifts and on-call operations.
The summary above was generated by AI
WHAT MAKES US, US

Join some of the most innovative thinkers in FinTech as we lead the evolution of financial technology. If you are an innovative, curious, collaborative person who embraces challenges and wants to grow, learn and pursue outcomes with our prestigious financial clients, say Hello to SimCorp!
 

At its foundation, SimCorp is guided by our values – caring, customer success-driven, collaborative, curious, and courageous. Our people-centered organization focuses on skills development, relationship building, and client success. We take pride in cultivating an environment where all team members can grow, feel heard, valued, and empowered.

If you like what we’re saying, keep reading!

WHY THIS ROLE IS IMPORTANT TO US

As a Senior Site Reliability Engineer, you will be working on Cloud Native Products & Services, taking ownership of various responsibility domains like monitoring, observability, release management, vulnerability management, cost management, audit & compliance etc. You will work closely with DevOps engineers, clients, and stakeholders to ensure reliability, performance, and automation for both existing and new cloud native products & services. Onboard and long-running clients on them. Your contributions will drive stability, continuous improvement, and operational excellence in our Azure-based environments. This role blends hands-on engineering, incident response, platform configuration, and service quality, - guided by ITIL and SRE best practices.

WHAT YOU WILL BE RESPONSIBLE FOR

  • Support the operational and enhancement of mission-critical environments for both new and existing Cloud Native products & services

  • Collaborate with product development teams to enhance monitoring, observability, reliability, and performance of these services.

  • Collaborate deeply across engineering teams to understand systems at the code level.

  • Manage & improve our infrastructure deployment pipelines and troubleshoot onboarding and operational issues

  • Drive capacity planning efforts to ensure our platform is resilient and scalable as we grow.

  • Build tools and automation to eliminate manual TOIL, improve engineering velocity, developer experience, and improve system reliability.

  • Define and manage SLOs and error budgets in partnership with Engineering teams.

  • Contribute to incidents, problems, and change management processes.

  • Execute disaster recovery, configuration management, and platform readiness tasks. 

  • Flexible working in regular & evening shift on rotational basis and provide weekend or On-Call support as needed. 

  • Collaborate with Agile teams and take part in design discussions with clients, vendors, and stakeholders. 

  • Contribute to knowledge sharing across multiple Product Areas. 

  • Leverage a strong foundation in ITIL practices, including problem, change, and incident management.

WHAT WE VALUE

  • Bachelor’s degree in Computer Science or related field (Master’s is a plus)

  •  3+ years in Site Reliability, DevOps, or Cloud Engineering roles

  • Must have expertise with Microsoft Azure Cloud.

  • Expertise in Infrastructure as Code (IaC) using Bicep, ARM and Terraform.

  • Solid experience in monitoring and logging tools (Azure Monitor, Application Insights, DataDog, Log Analytics).

  • Hand-on experience in IdP Onboarding and integrating, configuring IdP solutions like Azure Entra ID, Okta, KeyCloak or PingFederate.

  • Experience in centralizing authentication, managing user identities, and implementing secure access protocols (SAML, OAuth, OIDC)

  • Experience working with observability frameworks like Open Telemetry and distributed tracing systems

  • Experience working with application reliability platforms like Checkly or equivalent

  • Experience setting up synthetic monitoring using Playwright or equivalent

  • Knowledge of AI/ML-based anomaly detection, log aggregation and analysis tools like Microsoft Azure Anomaly Detector or equivalent

  • Experience working with Microsoft Defender Suite (EDR, XDR) and Sentinel. Proficient in KQL for threat hunting and improving compliance scores using Defender for Cloud. Able to identify and remediate vulnerabilities

  • Understanding of networking, containerization (Kubernetes, Docker)

  • Good understanding of APIs, scripting languages like PowerShell, Bash, Kusto and databases like SQL, Cosmos DB and Postgres SQL

  • Familiarity with SimCorp Dimension & Sales force is a plus

  • Proficiency in IT service management (ITSM) frameworks like ITIL, focusing on incident, change, and problem management to improve operational efficiency

  • Experience managing both onboarding projects and live production operations

  • Collaborative mindset and ability to work in cross-functional teams

  • Interest in continuous learning and growth within your Product Area

Benefits

  • Global hybrid work policy - We ask you to work 2 days a week from the office. If you choose you can work remotely the other days. Of course, you are welcome at the office if that is your preference.

  • Culture – Inclusive and diverse company culture 

  • Work-life balance – We believe that an equilibrium between professional responsibilities makes us all the best version of ourselves, both in private life and as colleagues in the workplace

  • Empowerment – We believe that all voices are valuable and must be heard. You will be involved in shaping our work processes

  • Career & Growth – Simcorp does offer opportunities for professional development: there is never just only one route - we offer an individual approach to professional development to support the direction you want to take.

NEXT STEPS

Please send us your application in English via our career site as soon as possible, we process incoming applications continually. Please note that only applications sent through our system will be processed. At SimCorp, we recognize that bias can unintentionally occur in the recruitment process. To uphold fairness and equal opportunities for all applicants, we kindly ask you to exclude personal data such as photos, age, or any non-professional information from your application. Thank you for aiding us in our endeavor to mitigate biases in our recruitment process.
 

We are eager to continually improve our talent acquisition process and make everyone’s experience positive and valuable. Therefore, during the process we will ask you to provide your feedback, which is highly appreciated.

WHO WE ARE

For over 50 years, we have worked closely with investment and asset managers to become the world’s leading provider of integrated investment management solutions. We are 3,000+ colleagues with a broad range of nationalities, education, professional experiences, ages, and backgrounds.
SimCorp is an independent subsidiary of the Deutsche Börse Group. Following the recent merger with Axioma, we leverage the combined strength of our brands to provide an industry-leading, full, front-to-back offering for our clients.

SimCorp is an equal opportunity employer and welcome applicants from all backgrounds, without regard to race, gender, age, disability, or any other protected status under applicable law. We are committed to building a culture where diverse perspectives and expertise are integrated into our everyday work. We believe in the continual growth and development of our employees, so that we can provide best-in-class solutions to our clients.

For Toronto City only: The annual base salary range for this position is 100 000,00 - 140 600,00 CAD. Additionally, employees are eligible for an annual discretionary bonus, and benefits including health care, leave, and retirement plans.

Your total compensation may vary based on role, location, department and individual performance. 

#Li-Hybrid

SimCorp Toronto, Ontario, CAN Office

100 Wellington St W, TD West Tower, Suite 2204 (PO Box 123), Toronto, Ontario, Canada, M5J

Similar Jobs

15 Days Ago
Easy Apply
Hybrid
Toronto, ON, CAN
Easy Apply
Senior level
Senior level
Marketing Tech • Mobile • Software
Lead reliability engineering for Braze’s NGINX ingress fleets, Kubernetes infrastructure, Ruby on Rails monolith, and Go API services. Design scalable, highly available systems; develop automated scaling routines; define SLIs, SLOs, and error budgets; conduct capacity planning; and participate in PagerDuty on-call rotations. Drive incident response, root-cause analysis, blameless retrospectives, runbook improvements, and resilient architecture decisions with product engineering teams.
Top Skills: AnsibleAWSAzureChefDatadogGCPGoGrafanaHorizontal Pod Autoscaler (Hpa)JavaKafkaKubernetesLinuxMongoDBNginxPagerdutyPostgresPrometheusPythonRedisRubyTcp/IpTerraformUnix
3 Days Ago
Hybrid
Toronto, ON, CAN
Senior level
Senior level
Software
Senior Site Reliability Engineer role at Pigment, an AI-powered SaaS business planning platform. The provided description emphasizes Pigment’s growth, global presence, company culture, equal opportunity commitment, and background-check process, but does not specify detailed responsibilities, technologies, or qualifications.
4 Days Ago
Hybrid
Toronto, ON, CAN
Senior level
Senior level
Software
Operate and maintain cloud-based SimCorp Dimension services for financial-services clients. Manage batch scheduling, incidents, escalations, environments, upgrades, patches, and service requests while meeting ITSM and SLA requirements. Troubleshoot mission-critical applications across infrastructure, middleware, databases, networking, and operating systems. Collaborate with global teams and customers, implement standardized solutions, maintain knowledge documentation, and improve service operations.
Top Skills: ActivebatchAws BatchAzureControl-MItilJams SchedulerJIRAOracle Saas Batch SchedulerRemedyforceServicenowSimcorp Dimension

What you need to know about the Toronto Tech Scene

Although home to some of the biggest names in tech, including Google, Microsoft and Amazon, Toronto has established itself as one of the largest startup ecosystems in the world. And with over 2,000 startups — more than 30 percent of the country's total startups — Toronto continues to attract new businesses. Be it helping entrepreneurs manage their finances, simplifying business operations by automating payroll or assisting pharmaceutical companies in launching new drugs, the city's tech scene is just getting started.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account