CIBC Logo

CIBC

Site Reliability Engineer

Posted 3 Days Ago
Be an Early Applicant
In-Office
Toronto, ON, CAN
Senior level
In-Office
Toronto, ON, CAN
Senior level
Provide senior production support and SRE leadership for Tier 1 Wealth Management trading platforms. Ensure availability, resiliency, observability, and performance across 24x7 operations. Lead incident, problem, change, recovery, and release activities; manage vendor performance; coordinate stakeholders; investigate root causes; implement automation and monitoring; and drive continuous service improvements. Apply knowledge of trading products and workflows to support operational readiness, business transformation, and reliable technology delivery.
The summary above was generated by AI

We’re building a relationship-oriented bank for the modern world. We need talented, passionate professionals who are dedicated to doing what’s right for our clients.

At CIBC, we embrace your strengths and your ambitions, so you are empowered at work. Our team members have what they need to make a meaningful impact and are truly valued for who they are and what they contribute.

To learn more about CIBC, please visit CIBC.com

What You'll Be Doing

You'll join CIBC's Wealth Technology team. As a Site Reliability Engineer, responsible for the operational stability, reliability, and support of Tier 1 Wealth Management Trading platforms. These mission-critical applications are primarily vendor-managed platforms and support business operations across trading, investment management, and client servicing functions. You will act as a senior reliability and production support specialist, ensuring high availability, resiliency, and performance of critical applications. You will work closely with business stakeholders, technology teams, and external vendors to drive operational excellence, incident resolution, continuous improvement, and implementation of Site Reliability Engineering (SRE) practices.This role requires participation in 24x7 production support, major incident management, change coordination and implementation, problem management, and recovery coordination activities.

At CIBC we enable the work environment most optimal for you to thrive in your role. Details on your work arrangement (proportion of on-site and remote work) will be discussed at the time of your interview.

How you’ll succeed     

  • Application Reliability & SRE: Drive the adoption and execution of Site Reliability Engineering (SRE) practices to improve platform availability, resiliency, observability, and operational excellence. Define, monitor, and report reliability metrics, service level objectives (SLOs), operational health indicators, and performance trends. Proactively identify reliability concerns and implement automation, monitoring, alerting, and remediation solutions to minimize service disruptions. Lead root cause investigations and reliability improvement initiatives to reduce recurring incidents and operational inefficiencies. Continuously improve application resiliency through incident reviews, problem management, trend analysis, and operational readiness assessments.
  • Production Support & Service Management: Provide senior-level production support for Tier 1 Wealth Management Trading applications operating in a 24x7 environment. Lead and coordinate resolution of critical production incidents, service disruptions, and business escalations. Apply strong Incident, Problem, and Change Management practices to maintain stable and highly available services. Conduct impact assessments, operational readiness reviews, and post-implementation validation for technology changes and releases. Identify opportunities for process optimization and operational efficiency improvements.
  • Vendor & Stakeholder Management: Act as the primary operational contact for strategic vendor partners supporting Wealth Trading applications. Manage vendor performance, service quality, issue resolution, escalation management, and operational accountability. Collaborate closely with internal technology teams, infrastructure teams, business partners, project teams, and external vendors to ensure effective delivery and support. Facilitate regular service reviews and operational governance discussions with key stakeholders. Ensure vendor-related incidents and service issues are addressed within established service levels and business expectations.
  • Trading Platform & Business Partnership: TDevelop deep understanding of Wealth Management products and trading processes, including Equities, Bonds, GICs, Mutual Funds, and related operational workflows. Translate business requirements and operational challenges into sustainable technology solutions. Partner with business stakeholders to identify opportunities to improve application reliability, user experience, and service delivery. Support strategic initiatives, upgrades, regulatory changes, and business transformation programs impacting trading platforms.
  • Change & Continuous Improvement: Serve as a subject matter expert for production readiness, release management, and operational risk assessments. Support implementation of monitoring, automation, and operational tooling improvements. Develop and maintain support documentation, operational procedures, runbooks, and recovery processes. Drive continuous service improvement initiatives to enhance stability, supportability, and operational maturity. Develop deep understanding of Wealth Management products and trading processes, including Equities, Bonds, GICs, Mutual Funds, and related operational workflows. Translate business requirements and operational challenges into sustainable technology solutions. Partner with business stakeholders to identify opportunities to improve application reliability, user experience, and service delivery. Support strategic initiatives, upgrades, regulatory changes, and business transformation programs impacting trading platforms.

Who you are

  • Experience & Expertise: You can demonstrate 5+ years of experience in Application Support, Application Reliability Engineering, Site Reliability Engineering (SRE), Production Operations, or Technology Operations within a financial institution. You have experience supporting Tier 1 mission-critical applications operating in a 24x7 production environment. You possess strong knowledge of Incident, Problem, Change, and Major Incident Management processes. You have experience working with and managing external technology vendors supporting enterprise applications. You have experience implementing or operating SRE practices, including monitoring, observability, automation, reliability measurement, and service resiliency improvements.
  • Business & Industry Knowledge: You have strong knowledge of Wealth Management and Trading platforms supporting Equities, Bonds, GICs, Mutual Funds, and related investment products. You understand the operational and business impact of technology disruptions within trading and investment management environments. You can effectively communicate technical issues and business impacts to both technical and non-technical stakeholders.
  • Technical & Analytical Skills: You're digitally savvy and continuously seek innovative solutions to improve reliability, efficiency, and service quality. You possess strong troubleshooting, analytical, and problem-solving skills with the ability to quickly assess complex production issues. You are skilled in monitoring, observability, automation, and operational support tools. You proactively identify opportunities to reduce manual effort and improve service reliability through automation and continuous improvement.
  • Values Matter: You bring your authentic self to work and live our values of Trust, Teamwork, and Accountability. You put clients first and understand the importance of maintaining highly available systems that support critical business operations and client experiences.

#LI-TA

What CIBC Offers

At CIBC, your goals are a priority. We start with your strengths and ambitions as an employee and strive to create opportunities to tap into your potential. We aspire to give you a career, rather than just a paycheck.

  • We work to recognize you in meaningful, personalized ways including a competitive salary, incentive pay, banking benefits, a benefits program*, defined benefit pension plan*, an employee share purchase plan, a vacation offering, wellbeing support, and MomentMakers, our social, points-based recognition program.

  • Our spaces and technological toolkit will make it simple to bring together great minds to create innovative solutions that make a difference for our clients.

  • We cultivate a culture where you can express your ambition through initiatives like Purpose Day; a paid day off dedicated for you to use to invest in your growth and development.

*Subject to plan and program terms and conditions

What you need to know

  • CIBC is committed to creating an inclusive environment where all team members and clients feel like they belong. We seek applicants with a wide range of abilities and we provide an accessible candidate experience. If you need accommodation, please contact [email protected]

  • CIBC is committed to clarity in our hiring process. All roles posted are opportunities we’re actively recruiting for, unless stated otherwise.

  • You need to be legally eligible to work at the location(s) specified above and, where applicable, must have a valid work or study permit.

  • We may ask you to complete an attribute-based assessment and other skills test (such as simulation, coding, French proficiency).

  • We use artificial intelligence tools during the recruitment process. Our goal for the application process is to get to know more about you, all that you have to offer, and give you the opportunity to learn more about us.

Job Location

Toronto-81 Bay, 19th Floor

Employment Type

Regular

Weekly Hours

37.5

Skills

Analytical Thinking, Application Production Support, Business Operations, Change Management, Impact Analysis, Implementation Planning, Incident Resolution, IT Operations Support, IT Vendor Management, Operational Efficiency, Problem Management, Reliability Management, Resiliency, Service Levels, Site Reliability Engineering, System Reliability, Technical Knowledge
HQ

CIBC Toronto, Ontario, CAN Office

Square, 81 & 141 Bay, Toronto, Ontario, Canada

CIBC Ontario, CAN Office

Canada

CIBC Ontario, CAN Office

Canada

CIBC Ontario, CAN Office

Canada

CIBC Ajax, Ontario, CAN Office

Ajax, Canada

CIBC Aurora, Ontario, CAN Office

Aurora, Canada

CIBC Brampton, Ontario, CAN Office

Brampton, Canada

CIBC Burlington, Ontario, CAN Office

Burlington, Canada

CIBC Etobicoke, Ontario, CAN Office

Etobicoke, Canada

CIBC Hamilton, Ontario, CAN Office

Hamilton, Canada

CIBC Markham, Ontario, CAN Office

Markham, Canada

CIBC Milton, Ontario, CAN Office

Milton, Canada

CIBC Mississauga, Ontario, CAN Office

Mississauga, Canada

CIBC North York, Ontario, CAN Office

North York, Canada

CIBC Oakville, Ontario, CAN Office

Oakville, Canada

CIBC Oshawa, Ontario, CAN Office

Oshawa, Canada

CIBC Richmond Hill, Ontario, CAN Office

Richmond Hill, Canada

CIBC Scarborough, Ontario, CAN Office

Scarborough, Canada

CIBC Vaughan, Ontario, CAN Office

Vaughan, Canada

CIBC Whitby, Ontario, CAN Office

Whitby, Canada

Similar Jobs

3 Days Ago
In-Office
L5R 0G1, Mississauga, ON, CAN
Entry level
Entry level
eCommerce • Fashion • Retail
Ensure critical enterprise services operate reliably, securely, and efficiently. Responsibilities include building monitoring and observability practices, leading incident response and root cause analysis, establishing SLI/SLO frameworks, improving operational standards, and driving automation, telemetry, analytics, and AIOps initiatives. The role collaborates with application, cloud, infrastructure, support, product, and business teams across hybrid environments, with Azure preferred.
Top Skills: AiopsApp InsightsAzure MonitorDynatraceGrafanaAzureSite Reliability Engineering (Sre)Splunk
11 Days Ago
Easy Apply
Hybrid
Toronto, ON, CAN
Easy Apply
Senior level
Senior level
Marketing Tech • Mobile • Software
Lead reliability engineering for Braze’s NGINX ingress fleets, Kubernetes infrastructure, Ruby on Rails monolith, and Go API services. Design scalable, highly available systems; develop automated scaling routines; define SLIs, SLOs, and error budgets; conduct capacity planning; and participate in PagerDuty on-call rotations. Drive incident response, root-cause analysis, blameless retrospectives, runbook improvements, and resilient architecture decisions with product engineering teams.
Top Skills: AnsibleAWSAzureChefDatadogGCPGoGrafanaHorizontal Pod Autoscaler (Hpa)JavaKafkaKubernetesLinuxMongoDBNginxPagerdutyPostgresPrometheusPythonRedisRubyTcp/IpTerraformUnix
6 Days Ago
In-Office
Toronto, ON, CAN
Expert/Leader
Expert/Leader
Fintech • Insurance • Financial Services
Leads multidisciplinary platform engineering teams supporting quantitative risk management platforms. Owns cloud, database, analytics, and infrastructure roadmaps; drives modernization, automation, reliability, architecture governance, operational excellence, compliance, and stakeholder alignment. Develops engineers through coaching, hiring, succession planning, and accountability while ensuring platforms remain secure, scalable, resilient, available, and cost-effective in a regulated banking environment.
Top Skills: AnsibleAWSLinuxAzurePowershellPythonSQL ServerTerraformWindows Server

What you need to know about the Toronto Tech Scene

Although home to some of the biggest names in tech, including Google, Microsoft and Amazon, Toronto has established itself as one of the largest startup ecosystems in the world. And with over 2,000 startups — more than 30 percent of the country's total startups — Toronto continues to attract new businesses. Be it helping entrepreneurs manage their finances, simplifying business operations by automating payroll or assisting pharmaceutical companies in launching new drugs, the city's tech scene is just getting started.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account