CI Financial Logo

CI Financial

Senior Observability Engineer

Posted 4 Days Ago
Be an Early Applicant
In-Office
Toronto, ON, CAN
Senior level
In-Office
Toronto, ON, CAN
Senior level
Build and manage enterprise observability across cloud and hybrid environments using Dynatrace and AWS CloudWatch. Responsibilities include platform architecture, APM instrumentation, dashboard and alert development, incident triage, SLI/SLO and error-budget management, observability-as-code automation, telemetry governance, cost optimization, CI/CD integration, operational reporting, and mentoring engineering teams. The role requires advanced Dynatrace, AWS, Kubernetes, scripting, distributed-systems, and incident-management expertise.
The summary above was generated by AI

At CI, we see a great place to work as one that is a safe place for everyone to have a voice, where people are empowered to take ownership over meaningful work, where there is an opportunity to grow through stretching themselves, where they can work on innovative products and projects, and where employees are supported and engaged in doing so. 

We are seeking a Mid-Level Observability Engineer to help build, maintain, and enhance our enterprise monitoring and observability capabilities across cloud and hybrid environments. This role is hands-on and execution-focused, supporting Dynatrace and AWS CloudWatch implementations, dashboard development, alert tuning, instrumentation, and operational reporting for critical platforms and applications. The ideal candidate will partner with Cloud Engineering, DevOps, SRE, and application teams to improve service visibility, strengthen monitoring coverage, and embed observability practices into ongoing operational and delivery workflows.

Observability Platform Ownership & Architecture

  • Design, deploy, and optimize enterprise-grade observability solutions using Dynatrace SaaS or Managed, including OneAgent, ActiveGate, full-stack monitoring, RUM, synthetic monitoring, Davis AI, distributed tracing, dashboards, and log monitoring on Grail.
  • Define platform standards for tagging, management zones, network segmentation, alerting profiles, access control, dashboards, and telemetry governance across hybrid environments.
  • Architect observability coverage across AWS and on-prem platforms, including containerized and serverless workloads such as EKS, ECS, Lambda, EC2, RDS, and API Gateway.
  • Lead migration from legacy monitoring tools into Dynatrace and drive closure of enterprise monitoring gaps through structured onboarding and platform modernization. 2. Application Performance Management & Incident Triage
  • Configure and optimize APM instrumentation for distributed applications, APIs, microservices, databases, and business transactions.
  • Serve as the escalation point for complex incidents, using Smartscape, Distributed Traces, Davis AI, Live Debugger, and method-level diagnostics to accelerate root cause identification and reduce MTTR.
  • Define and maintain SLIs, SLOs, and error budgets, aligning platform telemetry to business reliability targets and engineering commitments
  • Lead post-incident reviews using observability evidence and drive corrective improvements in instrumentation, thresholds, dashboards, and alerting logic3. Telemetry Automation & Observability as Code
  • Standardize monitoring configurations using Terraform and/or Dynatrace Monaco, including alerting profiles, dashboards, SLOs, tagging rules, synthetic tests, and management zones
  • Build automation for platform operations, integration workflows, reporting, and remediation using Python, Bash, or PowerShell, along with REST APIs and webhooks.
  • 5–10 years of experience in Observability, Monitoring Engineering, SRE, APM, DevOps, or Infrastructure Engineering, including several years of hands-on Dynatrace administration and architecture.
  • Deep hands-on expertise with Dynatrace across full-stack monitoring, Davis AI, Smartscape, RUM, synthetic monitoring, distributed tracing, Grail log monitoring, DQL, management zones, Workflows/AutomationEngine, and access governance.
  • Strong experience with AWS cloud services, especially CloudWatch, EKS, ECS, Lambda, EC2, RDS, API Gateway, networking, and modern cloud architecture patterns.
  • Advanced knowledge of Kubernetes and cloud-native observability patterns, including instrumentation for microservices and distributed systems.
  • Strong proficiency in observability-as-code using Terraform and/or Monaco, plus scripting in Python, Bash, or PowerShell
  • Solid understanding of distributed application architecture, networking fundamentals, telemetry pipelines, performance engineering, and incident management.
  • Dynatrace certification at Associate or Professional level required; higher-level certification is strongly preferred.

Preferred Qualifications

  • Experience with tools such as Nagios/SolarWinds/Prometheus/Grafana, Splunk, or ELK
  • Experience with OpenTelemetry, Dynatrace Grail, advanced log analytics, and enterprise telemetry standardization
  • Experience integrating observability with ITSM or event-management platforms such as ServiceNow
  • Background in SRE practices such as reliability reviews, error budget management, and incident reduction programs.
  • Integrate observability controls into CI/CD pipelines and establish telemetry quality standards for new application and infrastructure deployments
  • Use Dynatrace Query Language (DQL) and Grail capabilities for advanced log analysis, event correlation, notebooks, and custom operational insights. 4. Governance, Cost Control & Enablement
  • Own monitoring governance practices related to telemetry quality, alert design, data retention, platform usage standards, and operational reporting
  • Manage Dynatrace usage and consumption responsibly by monitoring ingest patterns, tuning retention, and optimizing log, metric, and trace collection for value and efficiency.
  • Build executive and engineering dashboards that communicate service health, reliability KPIs, error budgets, and infrastructure visibility to multiple audiences
  • Mentor engineers and partner teams on observability best practices, onboarding, dashboarding, instrumentation, and platform self-sufficiency.

This opportunity is for an existing vacancy with the company. The anticipated base salary range for this position is $85,000 to $125,000. Exact salary depends on several factors such as experience, skills, education, and budget. Salary range may vary based on geographic location. In addition to base salary, this position is eligible for participation in a bonus program. In addition, The Company offers a variety of benefits to eligible employees, including health insurance coverage, wellness programs, life and disability insurance, retirement savings plans, paid leave programs, education-related programs, paid holidays and vacation time, and many others. Many of these benefits are subsidized or fully paid for by the company.

    CI Financial is an independent company offering global wealth management and asset management advisory services through diverse financial services firms. Since 1965, we have consistently anticipated and responded to the changing needs of investors. We are driven by a commitment to provide individuals and institutions with the highest-quality investments and advice.   Our commitment to the highest levels of performance means that whatever their position, CI employees must be comfortable in a fast-paced environment that will stretch them to tap into their highest potential.  Employees with a healthy dose of ambition, a desire to commit to a curious mindset for continuous learning, and a willingness to go the extra mile thrive at CI. 

    A Supportive Environment for Success

    We offer an in-office environment, competitive benefits, and a supportive workplace to help our employees thrive both personally and professionally.

    WHAT WE OFFER 

    • Modern HQ location within walking distance from Union Station
    • Training Reimbursement
    • Paid Professional Designations
    • Employee Savings Plan (ESP)
    • Corporate Discount Program
    • Enhanced group benefits
    • Parental Leave Top–up program
    • Paid time off for Volunteering 

    We are focused on building a diverse and inclusive workforce. If you are excited about this role and are not confident you meet all the qualification requirements, we encourage you to apply to investigate the opportunity further.

    Please submit your resume in confidence by clicking “Apply”. Only qualified candidates selected for an interview will be contacted. CI Financial Corp. and all of our affiliates (“CI”) are committed to fair and accessible employment practices and provide reasonable accommodations for persons with disabilities. If you require accommodations in order to apply for any job opportunities, require this posting in an additional format, or require accommodation at any stage of the recruitment process please contact us at [email protected], or call 416-364-1145 ext. 4747. 

    HQ

    CI Financial Toronto, Ontario, CAN Office

    Toronto, Canada

    CI Financial Toronto, Ontario, CAN Office

    15 York St, 2nd Floor, , Toronto, Ontario , Canada, M5J 0A3

    Similar Jobs

    5 Days Ago
    In-Office or Remote
    Toronto, ON, CAN
    Senior level
    Senior level
    Software
    Design, build, deploy, and support scalable distributed observability systems for internal developers and customers. Lead the software development lifecycle, develop and review designs, ensure testing and production readiness, investigate reliability and performance issues, participate in on-call support, and collaborate across teams to improve Temporal’s infrastructure and platform capabilities.
    Top Skills: AWSClickhouseGCPGoGrafanaKubernetesLokiPrometheusSQLTemporalThanos
    26 Days Ago
    In-Office or Remote
    Canada
    Senior level
    Senior level
    Cloud • Information Technology • Software • Infrastructure as a Service (IaaS)
    Build ingestion pipelines for logs and metrics, scalable alerting engines, and observability APIs. Interface with product teams and develop microservices using Golang and Rust.
    Top Skills: AnsibleGoGraphQLGrpcRustTerraformTypescript
    26 Days Ago
    In-Office or Remote
    Canada
    Senior level
    Senior level
    Software
    The Senior Infra Engineer will build and maintain ingestion pipelines, scalable alerting engines, and observability APIs, while ensuring resilience and scalability in infrastructure. They will work with tools like Golang, Rust, Terraform, and Ansible, documenting requirements and interfacing with product teams.
    Top Skills: AnsibleGoGraphQLGrpcRustTerraformTypescript

    What you need to know about the Toronto Tech Scene

    Although home to some of the biggest names in tech, including Google, Microsoft and Amazon, Toronto has established itself as one of the largest startup ecosystems in the world. And with over 2,000 startups — more than 30 percent of the country's total startups — Toronto continues to attract new businesses. Be it helping entrepreneurs manage their finances, simplifying business operations by automating payroll or assisting pharmaceutical companies in launching new drugs, the city's tech scene is just getting started.

    Sign up now Access later

    Create Free Account

    Please log in or sign up to report this job.

    Create Free Account