Hiring.Camp

Site Reliability Engineering Lead

Pepsi Co

·

Yesterday

Location
MIGUEL HIDALGO, DF, MX
Type
Full-time
Department
IT
Seniority
Lead
Experience
3+ years
Closing date
Today
Source
iCIMS

Description

Overview

We Are PepsiCo  

  

Join PepsiCo and Dare for Better! We are the perfect place for curious people, thinkers and change agents. From leadership to front lines, we're excited about the future and working together to make the world a better place.  

  

Being part of PepsiCo means being part of one of the largest food and beverage companies in the world, with our iconic brands consumed more than a billion times a day in more than 200 countries.   

Our product portfolio, which includes 22 of the world's most iconic brands, such as Sabritas, Gamesa, Quaker, Pepsi, Gatorade and Sonrics, has been a part of Mexican homes for more than 116 years.  

  

A career at PepsiCo means working in a culture where all people are welcome. Here, you can dare to be you. No matter who you are, where you're from, or who you love, you can always influence the people around you and make a positive impact in the world.  

  

Know more: PepsiCoJobs 

Join PepsiCo, dare for better. 

 

Responsibilities

The Opportunity 

 

We are looking for a self-driven, software engineering mindset SRE engineer to:

• Drive new shift left activities critical to apply Site Reliability Engineering (SRE) and quality assurance principles within the application design / Project roadmap that enablees resilient outcomes.

• Apply pre-emptive approach into production minimizing business impact, via SRE-driven orchestration of connecting all components of the ecosystem diagnosing anomalies prior to user & remediating through automation.

 

This is a critical enabler achieving a high resiliency during operations and also continuously improving through design during the software development lifecycle. The Lead SRE design & support engineer is integral part of the global team with its main purpose to provide a delightful customer experience for the user of the global consumer, commercial, supply chain and enablement functions in the PepsiCo digital products application portfolio of 260+ applications, enabling a full SRE Practice incident prevention / proactive resolution model. The scope of this role is focussed on the cloud architecture application full stack devlopment, B2B pepsiconnect and Direct to Customer and other S&T roadmap applications. Ensures that PepsiCo DPA applications service performance, reliability and availability expected by our customers and internal groups. It requires a blend of technical expertise on SRE tools, modern applications cloud architecture i.e. full stack, IT operations experience, and analytics & influence skills.

  

Your Impact  

 

• Ensure ecosystem availability and performance in production environments, Pro-actively preventing P1, P2, potential P3s.

• Engage & influence product and engineering teams during the design and development phases to embed reliability and operability into new services defining & enforce events, logging, monitoring, and observability standards across applications.

• Accountable to institute non-functional requirements (NFRs) are embedded early including SLA/SLO/SLI and error budgets into the product’s offerings as part of the engineering solution.

• Leads the team diagnosing any anomalies prior to any user and driving the necessary remediations across the teams involved in end-to-end ecosystem availability, performance and consumption of the cloud architected application ecosystem leveraging SRE Orchestration solutions

• Collaborates with Engineering & support teams, including participation in escalations, and blameless postmortems.

• Work closely with customer-facing support teams to empower them with SRE insights and tooling.

• Observe, diagnose & improve the end-2-end ecosystem performance of the Modern architected application portfolio i.e. technical “understanding of interactions" of a full stack application alongside with peer SRE team member.

• Continuously optimize the L2/support operations work via SRE workflow automation

• Shape the SRE orchestration platform design with inputs from Production Operations, Business usage & Product and engineering teams. • Actively engage and drive AI Ops adoption across teams.

Qualifications

Who Are We Looking For?  

  

• 8+ years of work experience evolving to a SRE engineer with 3-5 years of experience in continuously improving and transforming IT operations ways of working.

• Bachelor’s degree in Computer Science, Information Technology or a related field.

• Proven experience as an SRE in designing the events diagnostics, performance measures and alert solutions to meet the SLA/SLO/SLIs.

• The ideal Engineer will be highly quantitative, have great judgment, able to connect dots across ecosytems, and efficiently work cross-functionally across teams to ensure SRE orchestrating solutions are meeting customer/end-user expectations.

• The candidate will take a pragmatic approach resolving incidents, including the ability to systemically triangulate root causes and work effectively with external and internal teams to meet objectives.

• A strong expertise of SRE (Software Reliability Engineering) and IT Service Management (ITSM) processes with a track record for improving service offerings – pro-actively resolving incidents, providing a seamless customer/end-user experience and proactively identifying and mitigating areas of risk.

• Hands on experience in Python, SQL /No-SQl( MySQL, Mongo DB, Cassandra, Postgress), AppDynamics, ELK Stack Grafana, Splunk, Dynatrace, Kafka and any SRE Ops toolsets.

• A firm understanding of cloud archticture for distributed environments.

• Front-end technologies: HTML, CSS, JavaScript, and frameworks like React, Angular, or Vue.js.

• Back-end technologies: Server-side languages (Java, Spring Boot, and related technologies that build the server-side logic, APIs, and database interaction with MySQL, MongoDB, Cassandra, Couchbase)

• Infrastructure: Azure/AWS cloud platforms and/or Client / server environments.

• Prior experience involving in shaping transformation developing SRE solutions would be a plus.

 

If this is an opportunity that interests you, we encourage you to apply even if you do not meet 100% of the requirements. 

  

What can you expect from us: 

  • Opportunities to learn and develop every day through a wide range of programs.  

  • Internal digital platforms that promote self-learning.  

  • Development programs according to Leadership skills.  

  • Specialized training according to the role.  

  • Learning experiences with internal and external providers.  

  • We love to celebrate success, which is why we have recognition programs for seniority, behavior, leadership, moments of life, among others.  

  • Financial wellness programs that will help you reach your goals in all stages of life.  

  • A flexibility program that will allow you to balance your personal and work life, adapting your working day to your lifestyle. 

  • And because your family is also important to us, they can also enjoy benefits such as our Wellness Line, thousands of Agreements and Discounts, Scholarship programs for your children, Aid Plans for different moments of life, among others.   

We are an equal opportunity employer and value diversity at our company. We do not discriminate based on race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status. We respect and value diversity as a work force and innovation for the organization. 

Skills

PythonJavaScriptJavaReactAngularVue.jsSpring BootAWSAzureSQLMySQLMongoDBCassandraSplunkSRE

Similar Jobs

30

Site Reliability Engineering

Zoox · Foster City, CA · Hybrid

2 months ago

Site Reliability Engineering

LSEG · IND-BLR-Divyasree Technopolis, India

11 months ago

Senior Engineer 2 - Site Reliability Engineering (Hybrid, Seattle)

Us We · Seattle WA, United States of America · Hybrid, Onsite

Yesterday

Senior Manager, Site Reliability Engineering

Finastra · Mississauga - Avebury, Canada

2 days ago

Manager, Site Reliability Engineering

LSEG · Bucharest - Iuliu Maniu Boulevard, Romania

2 days ago

Director of Site Reliability Engineering

JPMorgan Chase · Jersey City, NJ, United States, US

2 days ago

Director of Site Reliability Engineering

JP Morgan Chase · Jersey City, NJ, United States, US

2 days ago

Site Reliability Engineering (SRE) Leader

Patsnap · Remote, UK · Remote

3 days ago

Senior Engineer, Site Reliability Engineering

General Motors · Dublin LMR - Long Mile Works - Data Centre, Ireland · Hybrid

3 days ago

Team Lead - Site Reliability Engineering (all genders)

FactFinder · Pforzheim, Baden-Württemberg +3 · Remote, Hybrid

3 days ago

Site Reliability Engineering Manager

NationsBenefits, LLC · Plantation, FL, US · Remote

4 days ago

Manager Site Reliability Engineering

Sabre · Poland - Poland -Tischnera · Hybrid

4 days ago

Manager, Site Reliability Engineering

Oracle · BENGALURU, KARNATAKA, India

6 days ago

Lead, Site Reliability Engineering (Application Support)

Omers · Head Office Toronto, Canada

1 week ago

Software Engineer Lead - Site Reliability Engineering Center

PNC Bank · The Tower at PNC Plaza (PAA86), United States of America +5 · Onsite

1 week ago

Manager, Site Reliability Engineering

Oracle · Reston, VA, United States, US

1 week ago

Software Engineer, Site Reliability Engineering (Application Software)

Spacex · Hawthorne, CA +1

1 week ago

Senior Manager, Site Reliability Engineering

Oracle · United States, US

1 week ago

Senior Manager, Site Reliability Engineering

Oracle · Reston, VA, United States, US

1 week ago

Manager, Site Reliability Engineering

Octanner · USA - Utah-Salt Lake City-Headquarters, United States of America

1 week ago

Head of Site Reliability Engineering (SRE)

Computershare · Bristol, United Kingdom, GB · Hybrid

1 week ago

Engineer, Site Reliability Engineering

LSEG · IND-BLR-Divyasree Technopolis, India

1 week ago

Site Reliability Engineering (SRE)

DualEntry · Remote (EU, LATAM) · Remote

1 week ago

Site Reliability & Engineering Coach

Leland · Remote - USA · Remote

2 weeks ago

Site Reliability Engineering Senior Manager

Alliancedata · P1 - Easton Campus Building A, United States of America +1

2 weeks ago

Contract Lead, Site Reliability Engineering — AI Accelerator Infrastructure

D Matrix · Santa Clara · Hybrid

2 weeks ago

Sr. Site Reliability Engineering (Agentic Builders Experience team)

Adobe · San Jose, United States of America

2 weeks ago

Cloud Performance Engineering - Site Reliability Engineer

Smile Digital Health · Toronto, Ontario · Remote

2 weeks ago

Site Reliability Engineering Manager

Cae · Krakow, Poland · Hybrid

3 weeks ago

Director, Site Reliability Engineering

Thomson Reuters · United States of America, Frisco, Texas +1 · Hybrid

3 weeks ago