Hiring.Camp

Site Reliability Expert

Valtech

·

Today

Salary
$100k – $150k
Location
Canada - Remote
Workplace
Remote
Department
Engineering
Source
Greenhouse

Description

Why Valtech? We’re the experience innovation company - a trusted partner to the world’s most recognized brands. To our people we offer growth opportunities, a values-driven culture, international careers and the chance to shape the future of experience. 

The opportunity

At Valtech, you’ll find an environment designed for continuous learning, meaningful impact, and professional growth. Whether you're pioneering new digital solutions, challenging conventional thinking or building the next generation of customer experiences, your work will help transform industries. 

We are proud of: 

 

The role  

Please be aware thet French speaking skills are needed for this role.

We are seeking a highly experienced Site Reliability Expert to lead and drive observability, reliability, and operational excellence initiatives across complex cloud-native environments. This role goes beyond platform administration and requires a strong Site Reliability Engineering (SRE) background, combining observability expertise with production operations, automation, and reliability best practices. 

The ideal candidate will be an experienced technical leader capable of defining observability standards and strategies, supporting product teams, implementing reliability practices, and enabling scalable monitoring solutions across distributed microservices architectures. 

You will thrive in this role if you are: 

  • A curious problem solver who challenges the status quo 
  • A collaborator who values teamwork and knowledge-sharing 
  • Excited by the intersection of technology, creativity and data 
  • Experienced in Agile methodologies and consulting (a plus) 

Role responsibilities

  • Define and implement observability strategies, standards, and governance across applications and platforms. 
  • Design and maintain monitoring, alerting, dashboarding, and reporting solutions using Dynatrace or equivalent observability platforms. 
  • Establish and drive SRE best practices, including Service Level Indicators (SLIs), Service Level Objectives (SLOs), error budgets, and symptom-based alerting. 
  • Partner with engineering and product teams to improve system reliability, performance, and operational maturity. 
  • Develop standards for tagging, ownership, dashboard design, access management, and alerting governance. 
  • Support teams that are not specialized in observability by providing guidance, coaching, and knowledge transfer. 
  • Lead technical workstreams, prioritize initiatives, and ensure successful delivery within defined timelines and budgets. 
  • Analyze distributed systems and troubleshoot complex production issues using monitoring and tracing data. 
  • Promote documentation, operational rigor, and continuous improvement across engineering teams. 
  • Collaborate effectively within a distributed, multilingual environment. 

 

Must have qualifications

To be considered for this role, you must meet the following essential qualifications: 

Site Reliability Engineering & Observability 

  • Significant experience in Site Reliability Engineering (SRE) within large-scale production environments. 
  • Deep understanding of:  
  • Service Level Indicators (SLIs) 
  • Service Level Objectives (SLOs) 
  • Error Budgets 
  • Symptom-based Alerting 
  • Proven expertise with enterprise observability platforms such as:  
  • Dynatrace 
  • Datadog 
  • New Relic 
  • AppDynamics 
  • Strong experience with:  
  • Application Performance Monitoring (APM) 
  • Real User Monitoring (RUM) 
  • Monitoring agents and instrumentation 
  • Alerting strategies 
  • Role-Based Access Control (RBAC) 
  • SLO management 
  • Tagging and governance models 

Distributed Systems & Cloud Platforms 

  • Strong knowledge of OpenTelemetry (OTEL) and distributed tracing. 
  • Experience working within composable, microservices-based architectures. 
  • Hands-on production experience with:  
  • AWS 
  • Kubernetes 

Automation & DevOps 

  • Experience with infrastructure and operational automation. 
  • Practical knowledge of:  
  • Terraform 
  • Bash scripting 
  • Python scripting 
  • Experience with CI/CD tools such as GitLab CI or equivalent pipeline/workflow platforms. 

Delivery & Collaboration 

  • Demonstrated ability to lead technical initiatives and workstreams. 
  • Experience working within complex operational and agile environments. 
  • Strong stakeholder management and collaboration skills. 
  • Excellent communication skills in both French and English. 
  • Strong documentation practices, organizational skills, and attention to detail. 
  • High degree of autonomy and ownership. 

 

Nice to have qualifications  

  • Experience monitoring and supporting Java Spring Boot applications. 
  • Experience within e-commerce platforms and high-transaction environments. 
  • Experience establishing enterprise-wide observability frameworks and governance models. 
  • Consulting or advisory experience supporting multiple engineering teams. 

If you do not meet all the listed qualifications or have gaps in your experience, we still encourage you to apply. At Valtech, we recognize that talent comes in many forms, and we value diverse perspectives and a willingness to learn. 

 

 

The benefits  

This is a Full time position based in Canada. The offered salary range is $100,000 - $150,000 CAD annually, depending on experience and location. 

Valtech offers a comprehensive benefits package effective after three months of continuous service:

  • A comprehensive insurance plan, where you can choose the module that best suits your needs—Gold, Silver, or Bronze. The employer may contribute up to 80% of your coverage depending on the selected module. This plan includes short- and long-term disability coverage.
  • Dialogue via Sun Life provides virtual healthcare services, allowing you to consult with a healthcare professional for emergencies, prescription renewals, and more. You also have access to the Employee and Family Assistance Program, as well as a complete mental health support program.
  • A $500 Personal Spending Account, which can be used for healthcare reimbursements, gym memberships, public transit passes, office supplies, or contributions to your RRSP through Valtech.
  • A retirement plan where Valtech will match 100% of your RRSP contributions through a Deferred Profit Sharing Plan (DPSP), up to a maximum of 4%. You can start contributing to your RRSP immediately, and to the DPSP after 3 months. The vesting of the DPSP will be after a 24 months of service. 
  • Access to a flexible vacation under Valtech's policy to support your work-life balance, with 5 days available during your probation period and a prorated amount calculated for the remainder of the year.
  • Personal Technology Reimbursement – $30/month for every employee-offered on day 1. 
  • We close during the winter holidays and offer flexible scheduling throughout the year, so you can enjoy those sunny Friday afternoons—provided your weekly hours are completed.

 

Your application process

Once you apply, our Talent Acquisition team will review your application. If your skills and experience align with the role, we’ll reach out for next steps. Your CV should cover key information on relevant experiences and expertise. We do not require information such as age, gender, marital status, or a headshot in your application. We review all candidates based on skills, experience, and potential.

⚠️ Beware of recruitment fraud: Only engage with official Valtech email addresses.

We are committed to inclusion and accessibility. If you need reasonable accommodations during the interview process, please either indicate it in your application or let your Talent Partner know. 

  

About Valtech

Valtech is the experience innovation company that exists to unlock a better way to experience the world. By blending crafts, categories, and cultures, we help brands unlock new value in an increasingly digital world. 

At the intersection of data, AI, creativity, and technology, we drive transformation for leading organizations, including L’Oréal, Mars, Audi, P&G, Volkswagen Dolby, and more. 

At Valtech, we don’t just talk about transformation. We make it happen. Our people are the heart of our success, and we foster a workplace where everyone has the support to thrive, grow and innovate. 

Are you ready to create what’s next? Join us.

Skills

PythonJavaSpring BootAWSKubernetesTerraformCI/CDGitLabAgileDevOpsSREMicroservices

Similar Jobs

30

Site Reliability Expert

Valtech · Montreal

Yesterday

Expert DevOps

fr Orange · Sala Al Jadida, MA

1 month ago

SRE Expert

Ing · MILAN, Italy · Remote, Hybrid

2 months ago

Expert DEVOPS

Babelgroup · CASABLANCA, Morocco · Onsite

1+ year ago

Sr. Platform Engineer - Jfrog Expert - Remote US

Situsamc · Remote, United States of America · Remote

Yesterday

Expert DevOps Software Engineer

Applicant Portal · Lowell, AR - JB Hunt Corporate E, United States of America

Yesterday

Google Cloud Platform Engineer Expert (Remote, US)

Allstate · USA - IL (Remote), United States of America · Remote

2 days ago

Cloud DevOps Expert Engineer

Ing · PB_Cen_Katowice (ul. Sokolska 34), Poland +1

6 days ago

DevOps Engineer Expert

Leidos · 1471 Liberty Ctr Chantilly VA, United States of America

1 week ago

GPU DC East-West Network SRE Expert (SME)

Bitdeer Technologies Group · San Jose, CA, US · Remote

2 weeks ago

South North Network SRE Expert (SME)

Bitdeer Technologies Group · San Jose, CA, US · Remote

2 weeks ago

Subject Matter Expert (SME) DevOps Engineer

Redhorse · Chantilly, VA · Onsite

2 weeks ago

Ingénieur DevOps Expert Conteneurisation F/H (DSI/FAB)

RATP · BATIMENT NOISY LE GRAND, France · Hybrid

3 weeks ago

Ingénieur DevOps Expert Conteneurisation F/H (DSI/FAB)

Ratp · BATIMENT NOISY LE GRAND, France · Hybrid

3 weeks ago

Integration Platform Engineer Expert

Trabajos en Demo Datos · Madrid

1 month ago

Data Integration Platform Engineer Expert

Trabajos en Demo Datos · Madrid

1 month ago

Expert DevOps Engineer

Ing · PB_Cen_Katowice (ul. Chorzowska 50), Poland +1

1 month ago

Expert Devops Kubernetes H/F

Devoteam · Toulouse, Occitanie, France · Hybrid

1 month ago

Senior Software Engineer - DevOps Expert, Officer

Statestreet · Hyderabad, India

2 months ago

SRE Expert Lead - WB Tech @ING Hubs Romania

Ing · Bucharest - Dacia One, Romania

3 months ago

Cloud & DevOps Expert

ZutaCore · Sderot, Israel

3 months ago

SRE Expert Lead -PSS @ING Hubs Romania

Ing · Bucharest - Dacia One, Romania

3 months ago

Senior Site Reliability Engineer (SRE) – Dynatrace & Azure Observability Expert

RaceTrac · 1168 RaceTrac Store Support Center, United States of America · Remote, Hybrid, Onsite

3 months ago

Expert DevOps GenAI H/F

Inetum · LA CHAPELLE SUR ERDRE, France · Hybrid

3 months ago

Expert Software Engineer, Data Platform

Alegeus · Bangalore - India

3 months ago

DevOps Expert Engineer

HiNext · (HE)Office_KRK Pawia, Poland · Hybrid

3 months ago

Lead Expert, D&T Infra Platform Engineer

dsm-firmenich · Onsite

6 months ago

[Cadre] Expert DevOps Outillage de Production F/H (DSI/FAB)

Ratp · BATIMENT NOISY LE GRAND, France

8 months ago

Expert DevOps SRE Engineer | Devoteam Maroc Nearshore

Devoteam · Rabat, Morocco

10 months ago

Senior DevOps Expert (m/w/d)

Qvestgroup · Deutschland · Onsite

10 months ago