- Location
- IN Pune, India
- Workplace
- Remote, Hybrid, Onsite
- Type
- Full-time
- Department
- Engineering
- Seniority
- Senior
- Education
- Bachelor
- Source
- Workday
Description
We’re hiring at Pitney Bowes, where top talent builds meaningful careers and lasting impact. We Move fast, Deliver excellence, and Win together…that’s The Pitney Bowes way. Here, how we work matters just as much as what we achieve.
We’re looking for people who:
Act with urgency, accountability, and purpose
Deliver high quality work with consistency and pride
Collaborate effectively and elevate those around them
Focus on outcomes that drive impact and growth
Job Description:
Join Pitney Bowes as Senior Advisory Software Engineer
Years of experience: 8–10 years
Job Location: Pune
Work mode: Hybrid(3days/week in office)
Impact
We are looking for a hands-on, senior infrastructure architect who owns the reliability of our on-premises data center estate – the compute, virtualization, hyperconverged, storage, and network-security backbone that our production services depend on. This is a role for someone who has designed, operated, and hardened enterprise data centers at scale, and who can raise the resilience bar of an established 24/7 Enterprise Operations Center. You will not simply keep the lights on; you will critically evaluate the estate, expose hidden risk, and drive permanent, architectural fixes.
We actively look for prospects who demonstrate:
Deep expertise across on-premises data center infrastructure – hyperconverged (Nutanix), virtualization (VMware vSphere/ESXi), and software-defined networking and security (VMware NSX)
The ability to assess redundancy, high availability, and disaster recovery (DR) posture across data centers, and to close the gap between designed resilience and actual resilience
A rigorous, evidence-driven approach to reliability: DR testing, failover validation, failure-mode analysis, and validated recovery
Strong grounding in compute, storage, and network architecture, including capacity planning and performance optimization
The instinct to ask the right questions, challenge assumptions, and prompt effective investigations rather than accept surface-level explanations
Strong ownership of blameless post-mortems and the discipline to convert incidents into permanent, systemic fixes
Collaboration: Partnering closely with SRE, engineering, network, security, and cloud teams, and advising leadership on infrastructure resilience strategy and risk.
Problem-solving: Tackling complex, ambiguous challenges in data center design, virtualization, resilience, scalability, and performance across a hybrid estate.
Efficiency: Driving automation, standardization, and permanent remediation to reduce operational toil and recurring incidents.
If this sounds like you, then you may be a great fit for Pitney Bowes.
Desired Profile: Data Center & Virtualization Architect
Education: B.E./B.Tech/MCA or equivalent in Computer Science, Engineering, or a related field. 8–13 years of progressive experience in data center infrastructure, virtualization, and hyperconverged platforms, including significant time operating large-scale, mission-critical on-premises environments. Excellent interpersonal, written, and verbal communication skills in English. Self-motivated and able to prioritize and drive outcomes independently across multiple stakeholders. Relevant certifications (e.g., Nutanix NCP/NCM, VMware VCP/VCAP – DCV or NV, ITIL) are a strong plus.
The Job
Operate as a senior data center and virtualization architect within the Enterprise Operations Center (EOC), a 24/7 corporate IT operations and security function responsible for the availability, performance, and resilience of production services running across on-premises data centers (KDC, PDC, MDC) and AWS Cloud. You will independently evaluate the architecture of the on-premises estate – Nutanix hyperconverged clusters, VMware vSphere/ESXi hosts, and NSX network-security fabric – expose single points of failure, and validate that redundancy and DR mechanisms actually work as designed, not just on paper.
You will design and lead DR tests and failover exercises, mature the team's reliability practices, review and elevate runbooks, and set the standard for high-quality post-mortems and permanent remediation. Working alongside SRE, network, and cloud teams, you will ask the incisive questions that steer investigations to true root cause and ensure fixes are architectural rather than temporary. You will advise leadership on infrastructure risk and roadmap, and help embed observability, automation, and resilience-by-design across the estate (New Relic, Grafana, LogicMonitor, Site24x7, CloudWatch; ITSM via JIRA/JSM).
Responsibilities
Own and evaluate the architecture of the on-premises data center estate – Nutanix hyperconverged infrastructure, VMware vSphere/ESXi virtualization, and NSX-based networking and micro-segmentation – identifying single points of failure, capacity risks, and resilience gaps.
Assess and validate redundancy, high-availability, and disaster-recovery (DR) mechanisms across data centers – including host/cluster failover, replication, backup/restore, and RTO/RPO adherence – and drive remediation of gaps.
Design, plan, and lead DR tests, failover drills, and data center migration/cutover activities, and turn findings into hardening actions.
Provide deep architectural guidance on Nutanix (AOS, Prism, AHV), VMware (vCenter, ESXi, vSAN, SRM), and NSX (distributed firewall, micro-segmentation, edge/routing), plus underlying compute, storage, and network fabric.
Mature the EOC's reliability practices – define and operationalize monitoring, capacity, and performance baselines, reduce toil, and embed resilience-by-design into change and release processes.
Review, standardize, and elevate runbooks and operational procedures so that response is fast, consistent, and safe.
Ask the right questions and prompt effective investigations – steering incident response and problem management toward true root cause rather than symptoms.
Lead and contribute to blameless post-incident reviews (PIR) and root cause analysis (RCA), and ensure incidents result in permanent, systemic fixes – not recurring workarounds.
Champion observability best practices across New Relic, Grafana, LogicMonitor, Site24x7, and CloudWatch to improve signal quality, alerting, and mean-time-to-detect/resolve for infrastructure.
Advise leadership on infrastructure resilience strategy, risk posture, lifecycle/refresh, and the data center roadmap; mentor engineering team members and raise the overall technical bar.
Drive automation of operational and remediation tasks (IaC, scripting, JSM automation) to eliminate manual effort and reduce human error.
Primary / Must-have Skills
8–13 years of hands-on experience in on-premises data center and infrastructure engineering, with a strong track record of operating large-scale, mission-critical environments.
Deep expertise with Nutanix hyperconverged infrastructure – AOS, Prism, AHV, cluster operations, scaling, and troubleshooting.
Strong VMware virtualization expertise – vSphere, ESXi host management, vCenter, clustering (HA/DRS), and storage (vSAN / datastores).
Hands-on with VMware NSX – distributed firewall, micro-segmentation, security policies, edge/routing, and software-defined networking.
Solid data center fundamentals – compute, storage (SAN/NAS), networking, and physical/logical resilience across sites.
Proven ability to assess redundancy, high availability, and DR, and to design and validate recovery mechanisms (RTO/RPO, failover, replication, backup/restore).
Experience planning and executing data center migrations, cutovers, and refresh/lifecycle activities.
Excellent analytical and investigative skills – the ability to ask incisive questions, drive to true root cause, and communicate findings clearly to both engineers and leadership.
Secondary / Good to have
Exposure to AWS Cloud – core compute, networking, storage, and hybrid connectivity between on-premises and cloud (secondary to the on-prem focus of this role).
Experience with containers and orchestration (Docker, Kubernetes) and hybrid platform integration.
Nutanix (NCP/NCM) and VMware (VCP/VCAP – DCV or NV) certifications, or equivalent.
Experience defining and running formal DR programs and enterprise resilience frameworks.
Exposure to security tooling and practices (e.g., CrowdStrike, Zscaler, CyberArk) in a hybrid environment.
Experience with AI-driven operations (AIOps) – alert correlation, anomaly detection, and AI-assisted PIR/RCA workflows.
Experience mentoring engineers and influencing architecture across multiple teams.
Automation and scripting skills (e.g., PowerShell/PowerCLI, Python, Ansible) for infrastructure operations and remediation.
Hands-on with observability and monitoring tooling (e.g., LogicMonitor, New Relic, Grafana, Site24x7) and with ITSM platforms (JIRA/JSM).
Shifts / On-call
Primarily a senior advisory/architecture role, with participation in a rotational on-call escalation model for major incidents.
Weekend and public-holiday on-call escalation support as required.
About Pitney Bowes
Pitney Bowes (NYSE: PBI) is a global shipping and mailing company that provides technology, logistics, and financial services to more than 90 percent of the Fortune 500. Small business, retail, enterprise, and government clients around the world rely on Pitney Bowes to remove the complexity of sending mail and parcels. For additional information visit Pitney Bowes at www.pitneybowes.com.
We will:
• Provide the will: opportunity to grow and develop your career
• Offer an inclusive environment that encourages diverse perspectives and ideas
• Deliver challenging and unique opportunities to contribute to the success of a transforming organization
• Offer comprehensive benefits globally (PB Benefits and Wellbeing Programs)
Pitney Bowes is an equal opportunity employer that values diversity and inclusiveness in the workplace.
All interested individuals must apply online.