Hiring.Camp

Senior Advisory Software Engineer

Pitneybowes

·

Yesterday

Location
IN Pune, India
Workplace
Hybrid, Onsite
Type
Full-time
Department
Engineering
Seniority
Senior
Education
Bachelor
Source
Workday

Description

We’re hiring at Pitney Bowes, where top talent builds meaningful careers and lasting impact. We Move fast, Deliver excellence, and Win together…that’s The Pitney Bowes way. Here, how we work matters just as much as what we achieve.

We’re looking for people who:

  • Act with urgency, accountability, and purpose

  • Deliver high quality work with consistency and pride

  • Collaborate effectively and elevate those around them

  • Focus on outcomes that drive impact and growth

Job Description:

Join Pitney Bowes as Senior Advisory Software Engineer

Years of experience: 8–10 years

Job Location: Pune

Work mode: Hybrid(3days/week in office)

Impact

We are looking for a hands-on, senior technical architect who can act as a trusted reliability advisor across our hybrid estate – AWS Cloud and on-premises data centers. This is a role for someone who has designed, operated, and hardened large-scale production systems. You will not simply maintain systems; you will critically evaluate them, expose hidden risk, and drive permanent, architectural fixes.

We actively look for prospects who demonstrate:

  • Deep architectural judgment across compute, containers, networking, storage, and data – spanning AWS and on-premises data centers

  • The ability to assess redundancy, high availability, and disaster recovery (DR) posture, and to close the gap between designed resilience and actual resilience

  • A rigorous, evidence-driven approach to reliability: chaos engineering, failure-mode analysis, and validated recovery

  • The instinct to ask the right questions, challenge assumptions, and prompt effective investigations rather than accept surface-level explanations

  • Strong ownership of blameless post-mortems and the discipline to convert incidents into permanent, systemic fixes

Collaboration: Partnering closely with SRE, engineering, cloud, security, and infrastructure teams, and advising leadership on reliability strategy and risk.

Problem-solving: Tackling complex, ambiguous challenges in system design, resilience, scalability, and performance across a hybrid production estate.

Efficiency: Driving automation, standardization, and permanent remediation to reduce operational toil and recurring incidents.

If this sounds like you, then you may be a great fit for Pitney Bowes.

Desired Profile: Technical Architect – Senior SRE Advisor

Education: B.E./B.Tech/MCA or equivalent in Computer Science, Engineering, or a related field. 8–13 years of progressive experience in infrastructure engineering, cloud architecture, and/or site reliability engineering, including significant time operating large-scale production systems. Excellent interpersonal, written, and verbal communication skills in English. Self-motivated and able to prioritize and drive outcomes independently across multiple stakeholders. Relevant certifications (e.g., AWS Solutions Architect – Professional, CKA/CKAD, ITIL) are a strong plus.

The Job

Operate as a senior technical architect and reliability advisor within the corporate IT operations and security function responsible for the availability, performance, and resilience of production services running across on-premises data centers and AWS Cloud. You will independently evaluate the architecture of critical systems, expose single points of failure, and validate that redundancy and DR mechanisms work as designed – not just on paper.

You will design and lead chaos engineering exercises and DR tests, mature the team's SRE practices, review and elevate runbooks, and set the standard for high-quality post-mortems and permanent remediation. Working alongside SRE, cloud, and infrastructure teams, you will ask the incisive questions that steer investigations to true root cause and ensure fixes are architectural rather than temporary. You will advise leadership on reliability risk and roadmap, and help embed observability, automation, and resilience-by-design across the estate (New Relic, Grafana, LogicMonitor, Site24x7, CloudWatch; ITSM via JIRA/JSM).

Responsibilities

  • Evaluate the architecture of critical production systems across AWS and on-premises data centers, identifying single points of failure, capacity risks, and resilience gaps.

  • Assess and validate redundancy, high-availability, and disaster-recovery (DR) mechanisms – including failover, backup/restore, RTO/RPO adherence, and multi-AZ/region posture – and drive remediation of gaps.

  • Design and lead chaos engineering and game-day exercises to proactively surface failure modes, and turn findings into hardening actions.

  • Provide deep architectural guidance on AWS services, EKS/Kubernetes and container platforms, networking, and database reliability (HA, replication, backup/recovery, performance).

  • Review, standardize, and elevate runbooks and operational procedures so that response is fast, consistent, and safe.

  • Ask the right questions and prompt effective investigations – steering incident response and problem management toward true root cause rather than symptoms.

  • Lead and contribute to blameless post-incident reviews (PIR) and root cause analysis (RCA), and ensure incidents result in permanent, systemic fixes – not recurring workarounds.

  • Champion observability best practices across New Relic, Grafana, LogicMonitor, Site24x7, and CloudWatch to improve signal quality, alerting, and mean-time-to-detect/resolve.

  • Advise leadership on reliability strategy, risk posture, and the resilience roadmap; mentor SRE and engineering team members and raise the overall technical bar.

  • Drive automation of operational and remediation tasks (IaC, scripting, JSM automation) to eliminate manual effort and reduce human error.

Required Skills

  • 8–13 years of hands-on experience in infrastructure, cloud architecture, and/or SRE, including a proven track record as a Technical/Cloud Architect designing and operating large-scale, mission-critical AWS production systems.

  • Deep expertise across the AWS platform and the Well-Architected Framework — compute (EC2, Lambda, ECS, EKS, Fargate), networking (VPC, Transit Gateway, Route 53, ELB, CloudFront, Direct Connect), storage (S3, EBS), and IAM/security (KMS, Guard Duty).

  • Strong command of AWS Regions and Availability Zones, with the ability to architect multi-AZ and multi-Region topologies for fault isolation and high availability.

  • Proven ability to design and validate Disaster Recovery strategies — Backup & Restore, Pilot Light, Warm Standby, and Multi-Site Active/Active — mapped to defined RTO/RPO targets, including failover, cross-Region replication, and DR drills.

  • Solid experience assessing redundancy, high availability, and single points of failure, and engineering self-healing and graceful degradation into production systems.

  • Strong experience with containers and orchestration — Docker and Kubernetes, ideally Amazon EKS — including production operations and troubleshooting.

  • Solid database reliability experience — HA/replication, backup and recovery, and performance considerations across relational and/or NoSQL engines (RDS, Aurora, DynamoDB).

  • Practical experience with chaos engineering / resilience testing (e.g., AWS Fault Injection Service) and failure-mode analysis.

  • Hands-on with observability and monitoring tooling (e.g., New Relic, Grafana, Prometheus, LogicMonitor, CloudWatch) and with ITSM platforms (JIRA/JSM).

  • Automation and infrastructure-as-code skills (e.g., Terraform, Ansible, CloudFormation) and scripting (Python, Bash, PowerShell), with exposure to CI/CD pipelines for infrastructure delivery.

  • Excellent analytical and investigative skills — the ability to ask incisive questions, drive to true root cause, and communicate findings clearly to both engineers and leadership.

Good to have

  • Hands-on experience operating and maintaining on-premises data center infrastructure alongside cloud.

  • AWS Solutions Architect – Professional, Certified Kubernetes Administrator (CKA/CKAD), or equivalent certifications.

  • Experience defining and running formal DR programs and enterprise resilience frameworks.

  • Exposure to security tooling and practices (e.g., CrowdStrike, Zscaler, Idira) in a hybrid environment.

  • Experience with AI-driven operations (AIOps) – alert correlation, anomaly detection, and AI-assisted RCA workflows.

  • Experience mentoring senior engineers and influencing architecture across multiple teams.

About Pitney Bowes

Pitney Bowes (NYSE: PBI) is a global shipping and mailing company that provides technology, logistics, and financial services to more than 90 percent of the Fortune 500. Small business, retail, enterprise, and government clients around the world rely on Pitney Bowes to remove the complexity of sending mail and parcels. For additional information visit Pitney Bowes at www.pitneybowes.com.

Only Talent Matters at Pitney Bowes

Pitney Bowes is an equal opportunity workplace. To remove unconscious biases from our hiring process, we encourage 'Blind Applications' from candidates applying for jobs at Pitney Bowes. This means that details such as gender, caste, religion, nationality, and age are omitted from applications. And candidates can choose to reveal only their first or last name on the application.

We will:


• Provide the will: opportunity to grow and develop your career
• Offer an inclusive environment that encourages diverse perspectives and ideas
• Deliver challenging and unique opportunities to contribute to the success of a transforming organization
• Offer comprehensive benefits globally (PB Benefits and Wellbeing Programs)

Pitney Bowes is an equal opportunity employer that values diversity and inclusiveness in the workplace.

All interested individuals must apply online.

Skills

PythonAWSDockerKubernetesTerraformAnsibleCI/CDDynamoDBJiraSREITIL
Senior Advisory Software Engineer at Pitneybowes | Hiring.Camp