Hiring.Camp

Senior Observability Engineer

Ci

·

Today

Salary
$85k – $125k
Location
CDA ON Head Office - 15 York, Canada
Workplace
Remote, Hybrid
Type
Full-time
Department
Engineering
Seniority
Senior
Experience
5+ years
Source
Workday

Description

At CI, we see a great place to work as one that is a safe place for everyone to have a voice, where people are empowered to take ownership over meaningful work, where there is an opportunity to grow through stretching themselves, where they can work on innovative products and projects, and where employees are supported and engaged in doing so. 

We are seeking a Mid-Level Observability Engineer to help build, maintain, and enhance our enterprise monitoring and observability capabilities across cloud and hybrid environments. This role is hands-on and execution-focused, supporting Dynatrace and AWS CloudWatch implementations, dashboard development, alert tuning, instrumentation, and operational reporting for critical platforms and applications. The ideal candidate will partner with Cloud Engineering, DevOps, SRE, and application teams to improve service visibility, strengthen monitoring coverage, and embed observability practices into ongoing operational and delivery workflows.

Observability Platform Ownership & Architecture

  • Design, deploy, and optimize enterprise-grade observability solutions using Dynatrace SaaS or Managed, including OneAgent, ActiveGate, full-stack monitoring, RUM, synthetic monitoring, Davis AI, distributed tracing, dashboards, and log monitoring on Grail.
  • Define platform standards for tagging, management zones, network segmentation, alerting profiles, access control, dashboards, and telemetry governance across hybrid environments.
  • Architect observability coverage across AWS and on-prem platforms, including containerized and serverless workloads such as EKS, ECS, Lambda, EC2, RDS, and API Gateway.
  • Lead migration from legacy monitoring tools into Dynatrace and drive closure of enterprise monitoring gaps through structured onboarding and platform modernization. 2. Application Performance Management & Incident Triage
  • Configure and optimize APM instrumentation for distributed applications, APIs, microservices, databases, and business transactions.
  • Serve as the escalation point for complex incidents, using Smartscape, Distributed Traces, Davis AI, Live Debugger, and method-level diagnostics to accelerate root cause identification and reduce MTTR.
  • Define and maintain SLIs, SLOs, and error budgets, aligning platform telemetry to business reliability targets and engineering commitments
  • Lead post-incident reviews using observability evidence and drive corrective improvements in instrumentation, thresholds, dashboards, and alerting logic3. Telemetry Automation & Observability as Code
  • Standardize monitoring configurations using Terraform and/or Dynatrace Monaco, including alerting profiles, dashboards, SLOs, tagging rules, synthetic tests, and management zones
  • Build automation for platform operations, integration workflows, reporting, and remediation using Python, Bash, or PowerShell, along with REST APIs and webhooks.

  • 5–10 years of experience in Observability, Monitoring Engineering, SRE, APM, DevOps, or Infrastructure Engineering, including several years of hands-on Dynatrace administration and architecture.
  • Deep hands-on expertise with Dynatrace across full-stack monitoring, Davis AI, Smartscape, RUM, synthetic monitoring, distributed tracing, Grail log monitoring, DQL, management zones, Workflows/AutomationEngine, and access governance.
  • Strong experience with AWS cloud services, especially CloudWatch, EKS, ECS, Lambda, EC2, RDS, API Gateway, networking, and modern cloud architecture patterns.
  • Advanced knowledge of Kubernetes and cloud-native observability patterns, including instrumentation for microservices and distributed systems.
  • Strong proficiency in observability-as-code using Terraform and/or Monaco, plus scripting in Python, Bash, or PowerShell
  • Solid understanding of distributed application architecture, networking fundamentals, telemetry pipelines, performance engineering, and incident management.
  • Dynatrace certification at Associate or Professional level required; higher-level certification is strongly preferred.

Preferred Qualifications

  • Experience with tools such as Nagios/SolarWinds/Prometheus/Grafana, Splunk, or ELK
  • Experience with OpenTelemetry, Dynatrace Grail, advanced log analytics, and enterprise telemetry standardization
  • Experience integrating observability with ITSM or event-management platforms such as ServiceNow
  • Background in SRE practices such as reliability reviews, error budget management, and incident reduction programs.
  • Integrate observability controls into CI/CD pipelines and establish telemetry quality standards for new application and infrastructure deployments
  • Use Dynatrace Query Language (DQL) and Grail capabilities for advanced log analysis, event correlation, notebooks, and custom operational insights. 4. Governance, Cost Control & Enablement
  • Own monitoring governance practices related to telemetry quality, alert design, data retention, platform usage standards, and operational reporting
  • Manage Dynatrace usage and consumption responsibly by monitoring ingest patterns, tuning retention, and optimizing log, metric, and trace collection for value and efficiency.
  • Build executive and engineering dashboards that communicate service health, reliability KPIs, error budgets, and infrastructure visibility to multiple audiences
  • Mentor engineers and partner teams on observability best practices, onboarding, dashboarding, instrumentation, and platform self-sufficiency.

This opportunity is for an existing vacancy with the company. The anticipated base salary range for this position is $85,000 to $125,000. Exact salary depends on several factors such as experience, skills, education, and budget. Salary range may vary based on geographic location. In addition to base salary, this position is eligible for participation in a bonus program. In addition, The Company offers a variety of benefits to eligible employees, including health insurance coverage, wellness programs, life and disability insurance, retirement savings plans, paid leave programs, education-related programs, paid holidays and vacation time, and many others. Many of these benefits are subsidized or fully paid for by the company.

    CI Financial is an independent company offering global wealth management and asset management advisory services through diverse financial services firms. Since 1965, we have consistently anticipated and responded to the changing needs of investors. We are driven by a commitment to provide individuals and institutions with the highest-quality investments and advice.   Our commitment to the highest levels of performance means that whatever their position, CI employees must be comfortable in a fast-paced environment that will stretch them to tap into their highest potential.  Employees with a healthy dose of ambition, a desire to commit to a curious mindset for continuous learning, and a willingness to go the extra mile thrive at CI. 

    A Supportive Environment for Success

    We offer an in-office environment, competitive benefits, and a supportive workplace to help our employees thrive both personally and professionally.

    WHAT WE OFFER 

    • Modern HQ location within walking distance from Union Station
    • Training Reimbursement
    • Paid Professional Designations
    • Employee Savings Plan (ESP)
    • Corporate Discount Program
    • Enhanced group benefits
    • Parental Leave Top–up program
    • Paid time off for Volunteering 

    We are focused on building a diverse and inclusive workforce. If you are excited about this role and are not confident you meet all the qualification requirements, we encourage you to apply to investigate the opportunity further.

    Please submit your resume in confidence by clicking “Apply”. Only qualified candidates selected for an interview will be contacted. CI Financial Corp. and all of our affiliates (“CI”) are committed to fair and accessible employment practices and provide reasonable accommodations for persons with disabilities. If you require accommodations in order to apply for any job opportunities, require this posting in an additional format, or require accommodation at any stage of the recruitment process please contact us at [email protected], or call 416-364-1145 ext. 4747. 

    Skills

    PythonAWSKubernetesTerraformCI/CDServiceNowSplunkRESTDevOpsSREMicroservices

    Similar Jobs

    30

    Senior Observability Engineer

    Saxobank · Headquarters, Denmark · Remote, Hybrid

    6 days ago

    Senior Observability Engineer

    Experian · Hyderabad, India · Hybrid

    1 month ago

    Senior Observability Engineer

    Rbc · 250 NICOLLET MALL:MINNEAPOLIS, United States of America

    1 month ago

    Senior Observability Engineer

    modernatx · POL - Mazowieckie - Warsaw - MESH Rondo Ignacego Daszynskiego 1, Poland +1 · Remote, Hybrid

    1 month ago

    Senior Observability Engineer

    Rent the Runway · Galway, Ireland +1

    2 months ago

    Senior Observability Engineer

    M&G · Pune, India

    2 months ago

    Sr. Observability Engineer

    Micron Technology · Remote

    8 months ago

    Sr. Observability Engineer

    Micron · Hyderabad - Phoenix Aquila, India

    8 months ago

    Senior Enterprise Observability Operations Engineer

    Careers Home · ABC Manila Office, Philippines · Onsite

    Today

    Senior AWS DevOps Engineer - AWS, Kubernetes, HCP, CI/CD, Observability with AI-focus (REMOTE)

    Koniag Government Services · Remote

    Yesterday

    Senior Observability Systems Engineer

    Q2Ebanking · Bangalore, India

    Yesterday

    Senior Enterprise Observability Operations Engineer

    Careers Home · ABC Manila Office, Philippines · Onsite

    Yesterday

    Senior Observability Automation Engineer

    Q2Ebanking · Austin, Texas, United States of America

    3 days ago

    Observability Lead-Sr. Infrastructure Engineer

    Truist · Charlotte NC - 2320 Cascade Pointe Boulevard, United States of America

    1 week ago

    Sr Software Engineer Enterprise Tools (Observability) Architecture

    Staples Canada · Framingham, MA, United States, US · Onsite

    1 week ago

    Senior Software Engineer – DevOps/Observability Platform

    Nexthink · Madrid, MD, Spain · Hybrid

    1 week ago

    Senior Software Engineer - Observability & IRM

    The Trade Desk · Boulder; Denver; Seattle +3

    2 weeks ago

    Senior Elastic Security & Observability engineer

    Professional Kyndryl · KNL51596 Hoofddorp (KNL51596), Netherlands · Remote

    2 weeks ago

    Senior Elastic Security & Observability Engineer

    Professional Kyndryl · KNL51596 Hoofddorp (KNL51596), Netherlands · Remote

    2 weeks ago

    Sr. Observability Engineer – Integration Layer & OpenTelemetry (OTEL)

    Zensar Technologies · Bangalore, Karnataka, India

    2 weeks ago

    Senior Lead Engineer, Platform Operations & Observability

    The Future of Health Starts with You · Saint-Laurent, QC, CAN - 4705 Dobrin Street (MC41), Canada · Hybrid

    2 weeks ago

    Senior Software Engineer, Observability platform

    Grab · HCMC, Vietnam · Remote, Onsite

    3 weeks ago

    Senior Site Reliability Engineer (SRE) – OpenShift & Observability

    Swift · OPC NL, Netherlands

    3 weeks ago

    Sr. Observability Engineer – Grafana & Prometheus

    Zensar Technologies · Bangalore, Karnataka, India

    3 weeks ago

    Sr. Security Engineer, Security Search and Observability, Field Innovation, Security Search and Observability (SSO)

    Amazon · Remote

    3 weeks ago

    Senior Datadog Security & Observability Engineer

    Keepersecurity · Remote, US · Remote

    4 weeks ago

    Senior SRE – Unified Observability Engineer

    Ncr · Hyderabad (Western Aqua) Mixed Use Office, India · Remote, Hybrid

    1 month ago

    Senior Dynatrace Observability Engineer

    Mufgub · BCIT Bengaluru Office (MGS), India

    1 month ago

    #127606 - Senior Software Engineer / SRE (Observability Focus)

    Lifted · Singapore, Singapore · Remote

    1 month ago

    Senior Site Reliability & Observability Engineer (SRE)

    TransUnion · Eagle - D.F. Del. Miguel Hidalgo, Mexico · Remote, Onsite

    1 month ago