Hiring.Camp

Site Reliability Engineer

Pg

·

Today

Location
MANILA NET PARK OFFICE, Philippines
Type
Full-time
Department
Engineering
Experience
5+ years
Education
Bachelor
Source
Workday

Description

Job Location

Taguig City

Job Description

Information Technology (IT) at Procter & Gamble is where business, innovation and technology integrate to build a competitive advantage for P&G. Our mission is clear -- you deliver IT to help P&G win with consumers. 

 

Do you love implementing continuous improvement in IT solutions to drive efficiency and agility in meeting constantly evolving business needs? Then this job might be for you! 

 

As a Site Reliability Engineer, you will be instrumental in ensuring the high availability and reliability of our digital IT products in P&G Your primary focus will be on enhancing system performance through faster detection, response, and resolution of issues, while also implementing strategies to prevent recurrence and reduce operational toil. You will use robust Observability and Monitoring tools, automate incident response systems, and optimize IT architecture to create a resilient and reliable infrastructure. 

This is a Managerial position. Being a manager at P&G involves leading teams and / or end-to-end processes, managing P&G resources, and driving business results. Managers are responsible for overseeing various aspects of the business, including strategy, operations, and team performance. They play a crucial role in ensuring that P&G's brands continue to grow and succeed in the market. Managers at P&G are expected to have strong leadership skills, a growth mindset, and the ability to make data-driven decisions roles lead and initiatives, significantly impacting business results through independent judgment and minimal guidance.

Responsibilities: 

  • Implement and lead comprehensive monitoring solutions and tools to provide real-time insights into system performance, enabling proactive incident detection and ensuring accurate, actionable alerts for prompt responses. 

  • Continuously refine monitoring strategies and develop automation scripts to address recurring issues, enhancing system visibility, resource optimization, and overall efficiency. 

  • Establish and maintain Service Level Indicators (SLIs) and Service Level Objectives (SLOs) to improve service quality and reliability, 

  • Collect and share data and insights from observability tools to drive continuous improvement initiatives. 

  • Work closely with Software Engineers, Product Teams, and Infrastructure Teams to develop and implement initiatives that enhance IT reliability. 

  • Engage with customers to understand their needs and difficulties regarding Observability and Monitoring tools, providing exceptional support in all interactions, including communications, updates, and feedback. 

  • Stay updated on industry trends and effective strategies in Site Reliability Engineering while continuously enhancing technical skills in system architecture, automation, cloud technologies, and operational processes. 

Job Qualifications

Candidates must demonstrate strong leadership in the application of technical expertise to drive business results. 

We are looking for candidates who possess the following core qualities: 

  • A Bachelor's degree in related field such as Engineering, Information Technology and Computer Science discipline, and up to 5 years experience at most. 

  • Experience or familiarity with monitoring and observability tools (e.g., Prometheus, preferably Grafana) 

  • Knowledge and familiarity in system administration, including Linux/Unix environments, cloud platforms (Azure or GCP preferred, but AWS is acceptable) 

  • Experience with configuration management tools and infrastructure-as-code frameworks (e.g., Terraform) 

  • Proficiency in at least one programming language (e.g., Python, C#) and a background in scripting for automation tasks 

  • Understanding of networking protocols, network infrastructures, load balancing, and DNS management 

  • Familiarity with containerization and Orchestration Technologies (e.g., Docker, Kubernetes) 

  • Familiarity with databases and proficiency in writing SQL queries 

  • Understanding of best practices in security and experience with implementing secure systems 

  • Knowledge of incident response methodologies, root cause analysis, and implementing preventive measures (ITIL and/or SRE) 

  • Familiarity with ticketing systems and task management (preferably ServiceNow) 

  • Problem-solving skills with ability to analyze complex issues and devise effective solutions 

  • Learning agility as there will be new topics to learn and new spaces to understand 

  • Communication and collaboration skills to work effectively with multi-functional teams, partners, and customers 

  • Teamwork and interpersonal skills, with an ability to build relationships and work effectively in a collaborative environment 

  • Operational excellence / execution skills as the work requires discipline 

Job Schedule

Full time

Job Number

R000156530

Job Segmentation

Entry Level

Skills

PythonAWSAzureGCPDockerKubernetesTerraformLinuxSQLServiceNowSREITIL

Similar Jobs

30

Site Reliability Engineer

Outpost · Remote · Remote

3 days ago

Site Reliability Engineer

Kong · Milan, Italy · Hybrid

3 days ago

Site Reliability Engineer

PNC Bank · Phoenix - East Camelback Exec Suite (AZ014), United States of America +5

3 days ago

Site Reliability Engineer

Cisco · GBR-LONDON, United Kingdom · Hybrid

3 days ago

Site Reliability Engineer

Synechron · MDC – Montreal, Canada

3 days ago

Site Reliability Engineer

Vynca · Remote - United States · Remote

4 days ago

Site Reliability Engineer

Ghr · Charlotte, United States of America +1 · Onsite

4 days ago

Site Reliability Engineer

Empower · KS Overland Park, United States of America · Remote

4 days ago

Engineer, Site Reliability

Vanguard · USA - Majestic, United States of America +1

5 days ago

Engineer, Site Reliability

Vanguard · USA - Majestic, United States of America +1

5 days ago

Site Reliability Engineer

Willisre · London, United Kingdom

5 days ago

Site Reliability Engineer

Sabre · India - Bangalore-Navigator Bldg

5 days ago

Site Reliability Engineer

Pluxee · BEL_ Bruxelles (1000), Belgium · Hybrid

5 days ago

Site Reliability Engineer

Andurilindustries · Waltham, Massachusetts, United States

6 days ago

Site Reliability Engineer

Incidentiq · Incident IQ North (Alpharetta) +1

6 days ago

Site Reliability Engineer

Ad Astra · Overland Park, KS

1 week ago

Site Reliability Engineer

Myfitnesspal · Remote - US · Remote

1 week ago

Site Reliability Engineer

Trading Technologies · Ahmedabad/GiftCity · Remote, Hybrid, Onsite

1 week ago

Site Reliability Engineer

Calix · Bangalore, India

1 week ago

Site Reliability Engineer

Clearscoretechnologylimited · London, England, United Kingdom

1 week ago

Site Reliability Engineer

Recordedfuture · Gothenburg, Sweden

1 week ago

Site Reliability Engineer

NMRK-Property Management-PM Northeast · Belfast, County Antrim, United Kingdom, GB

1 week ago

Site Reliability Engineer

Cisco · USA-SAN FRANCISCO, United States of America +18 · Remote

1 week ago

Site Reliability Engineer

LSEG · USA-St. Louis-795 Office Pkwy, United States of America

1 week ago

Site Reliability Engineer

Mizuho Mizuho is in growth · MetroPark, United States of America +1 · Remote, Hybrid

1 week ago

Site Reliability Engineer

LSEG · St. Loui, Missouri, United States of America +1

1 week ago

Site Reliability Engineer

Electrolux · US-CLT-001, United States of America

1 week ago

Site Reliability Engineer

Picogrid · El Segundo, CA

1 week ago

Site Reliability Engineer

Unum · Atlanta, Georgia, USA, United States of America

1 week ago

Site Reliability Engineer

Kiongroup · Kraków, Poland · Hybrid

1 week ago
Site Reliability Engineer at Pg | Hiring.Camp