- Location
- SG-01-SINGAPORE-083 ~ 83 Clemenceau Ave ~ UE SQAURE
- Workplace
- Remote, Onsite
- Type
- Full-time
- Department
- Engineering
- Seniority
- Senior
- Experience
- 8+ years
- Closing date
- Today
- Source
- Workday
Description
Date Posted:
2026-08-14Country:
SingaporeLocation:
SG-01-SINGAPORE-083 ~ 83 Clemenceau Ave ~ UE SQAUREPosition Role Type:
OnsiteAt RTX, the world's largest aerospace and defense company, 185,000 great minds are united by purpose and inspired to make a difference solving the world’s most complex problems. With our three market leading businesses, world-class operations and investments in research and development, we offer capabilities and opportunity no one else can. Together, we push the boundaries of known science and find new ways to connect and protect our world.
Collins Aerospace is a leader in technologically advanced, intelligent solutions that help redefine the aerospace and defense industry. With a comprehensive portfolio and deep technical expertise, we help customers meet the demands of the global market. Join us and help shape the future of aerospace and defense.
Senior System Deployment Engineer
Role Overview
We are seeking an experienced Senior System Deployment to join our team. The ideal candidate will be responsible for the deployment, configuration, testing, recovery, and ongoing support of mission-critical airline and airport applications across Windows and Linux environments.
The role covers the complete lifecycle of application infrastructure, including server staging, OS installation, VM provisioning, application deployment, database configuration, high-availability and failover testing, disaster recovery, and production transition.
A key responsibility is providing critical second-level/third-level technical support for major incidents, including the ability to recover and rebuild corrupted virtual machines, operating systems, databases, and application environments from scratch using available backups, recovery media, infrastructure templates, and documented recovery procedures.
The engineer will work across Windows Server, Red Hat Linux, VMware, Hyper-V, AWS Cloud, SQL Server, PostgreSQL, Kafka, MongoDB, OpenShift, and related infrastructure technologies, supporting highly available and operationally critical airport systems.
Key Responsibilities
1. Server Staging, Build & Deployment
· Stage, provision, and prepare Windows and Linux servers for application deployment.
· Perform OS installation, hardening, patching, configuration, and baseline validation.
· Deploy and configure airline and airport applications across physical and virtual environments.
· Configure application services, ports, certificates, connectivity, service accounts, and system dependencies.
· Ensure servers meet required performance, security, availability, and operational standards.
· Prepare production, DR, test, FAT, UAT, and staging environments.
2. Critical Incident Support & Infrastructure Recovery
· Provide second-level/third-level technical support for mission-critical airport and airline systems.
· Participate in major incident, problem management, and ITSM processes.
· Troubleshoot critical failures involving servers, VMs, operating systems, databases, applications, storage, and infrastructure services.
· Lead or support full infrastructure recovery when systems become corrupted or unrecoverable.
· Restore and rebuild VMs from scratch, including:
o Recreating virtual machines.
o Reinstalling operating systems.
o Applying required OS configuration and security hardening.
o Restoring application components and services.
o Reconfiguring network, storage, DNS, certificates, and connectivity.
o Restoring application configuration and validating dependencies.
· Recover corrupted or failed servers using available VM backups, snapshots, images, templates, backup repositories, or DR infrastructure.
· Perform complete application environment restoration when the original server/VM cannot be recovered.
· Work with infrastructure, backup, network, database, security, and application teams during major recovery activities.
· Conduct post-recovery validation and ensure systems are returned to operational readiness.
3. Database Recovery & Restoration
· Support SQL Server and PostgreSQL environments supporting mission-critical applications.
· Troubleshoot database corruption, service failures, connectivity issues, performance issues, and abnormal database behavior.
· Perform or coordinate database restoration from backup following database corruption or infrastructure failure.
· Restore databases from full, differential, transaction-log, WAL, or other applicable backup mechanisms.
· Rebuild database servers when required, including:
o OS and VM restoration.
o Database software installation.
o Database configuration.
o Storage and permissions configuration.
o Database restoration.
o User, role, service account, and connectivity configuration.
o Application reconnection and validation.
· Validate database integrity following restoration.
· Support database failover, recovery, replication, and DR testing.
· Coordinate with DBA teams for complex database recovery and corruption scenarios.
· Ensure appropriate RPO/RTO requirements are achieved during recovery.
4. Virtualization & Infrastructure Management
· Manage and troubleshoot VMware and Hyper-V virtual environments.
· Provision, configure, resize, clone, restore, and recover virtual machines.
· Troubleshoot VM, host, datastore, virtual network, resource allocation, and performance issues.
· Perform VM recovery from backup or rebuild VMs from scratch when required.
· Support VMware HA, clustering, snapshots, templates, and DR-related activities.
· Investigate virtualization-related failures affecting critical applications.
· Support data center infrastructure troubleshooting involving compute, storage, virtualization, and server connectivity.
5. High Availability, Failover & Disaster Recovery
· Design, execute, and document server, application, database, and VM failover testing.
· Conduct planned and unplanned failover exercises.
· Validate application recovery following infrastructure or database failures.
· Perform DR restoration and recovery testing.
· Validate RPO and RTO against operational requirements.
· Test recovery procedures for complete server/VM loss and database corruption scenarios.
· Identify recovery gaps and recommend improvements to HA/DR architecture.
· Maintain detailed recovery procedures and runbooks.
6. FAT, UAT & Production Validation
· Plan and execute Factory Acceptance Testing (FAT) and User Acceptance Testing (UAT).
· Perform technical validation before applications are transitioned into production.
· Validate server, application, database, network, security, and infrastructure dependencies.
· Conduct failover, recovery, performance, and resilience testing.
· Support application cutover and go-live activities.
· Ensure production handover is completed with appropriate documentation and support procedures.
7. Database & Middleware Services
· Configure and support:
o Microsoft SQL Server
o PostgreSQL
o Kafka
o MongoDB
o OpenShift
o Related application middleware and services
· Troubleshoot application-to-database connectivity and service dependencies.
· Support Kafka brokers, topics, connectivity, and service availability.
· Support MongoDB configuration, availability, and recovery activities.
· Work with development and application teams to resolve middleware and database-related incidents.
8. AWS Cloud & Hybrid Infrastructure
· Support applications deployed across on-premises data centers and AWS Cloud.
· Troubleshoot cloud infrastructure, connectivity, compute, storage, and application dependencies.
· Support VM/server recovery and application restoration in hybrid environments.
· Assist with cloud DR and backup/recovery activities.
· Understand connectivity between on-premises infrastructure and AWS environments.
9. Monitoring, Troubleshooting & Operational Support
· Monitor system health, availability, capacity, and performance.
· Analyze server, application, database, and infrastructure logs.
· Troubleshoot CPU, memory, disk, network, service, database, and application failures.
· Work with monitoring and ITSM tools to identify and resolve recurring incidents.
· Participate in root cause analysis (RCA) and problem management.
· Identify opportunities for automation and proactive monitoring.
10. Scripting & Automation
· Develop scripts and automation to reduce manual deployment and recovery activities.
· Use PowerShell, Bash, Python, or similar scripting technologies where appropriate.
· Automate server health checks, service validation, deployment activities, backup validation, and recovery procedures.
· Develop automated health-check and auto-healing capabilities for critical services where feasible.
11. Documentation & Knowledge Management
· Create and maintain:
o Server build documents
o VM configuration documents
o Application deployment procedures
o Database recovery procedures
o VM recovery procedures
o Disaster recovery runbooks
o Failover test procedures
o FAT/UAT test cases
o Troubleshooting guides
o Operational handover documents
· Maintain accurate recovery procedures to enable complete rebuild and restoration from a failed/corrupted environment.
· Ensure documentation is regularly tested and updated following system changes or incidents.
12. Collaboration
· Collaborate with application development, DBA, network, security, cloud, infrastructure, service desk, and operations teams.
· Coordinate with vendors and technology partners during critical incidents.
· Participate in technical reviews, change management, CAB activities, and production implementation.
· Provide technical guidance to junior engineers and operations teams.
Qualifications
· Bachelor’s degree in Computer Science, Information Technology, Engineering, or related discipline.
· 8+ years of experience in system deployment, infrastructure support, system recovery, or mission-critical IT environments.
· Strong experience with Windows Server and Red Hat Linux.
· Strong experience in server staging, OS build, configuration, deployment, troubleshooting, and recovery.
· Hands-on experience with VMware and/or Hyper-V.
· Strong experience with VM restoration, rebuilding, and disaster recovery.
· Experience recovering databases and application environments following corruption or infrastructure failure.
· Experience with SQL Server and PostgreSQL.
· Working knowledge of Kafka, MongoDB, and OpenShift.
· Experience with high-availability systems, failover, backup, restore, and DR.
· Experience conducting FAT, UAT, failover, recovery, and production validation testing.
· Working knowledge of AWS Cloud and hybrid infrastructure.
· Experience with ITSM, incident management, change management, and problem management.
· Basic networking knowledge including TCP/IP, DNS, DHCP, firewall rules, routing, VLANs, and connectivity troubleshooting.
· Experience with scripting/automation using PowerShell, Bash, Python, or equivalent.
· Relevant certifications such as Microsoft, VMware, Red Hat, AWS, or other senior infrastructure certifications are desirable.
· Strong analytical, troubleshooting, communication, and documentation skills.
Preferred Skills
· Prior experience supporting mission-critical airline, airport, transportation, other 24x7 environments.
· Experience supporting systems with strict availability, RTO, and RPO requirements.
· Hands-on experience with complete infrastructure recovery from bare/clean VM or server build through application restoration and production validation.
· Experience with VMware HA, clustering, snapshots, templates, and backup/restore solutions.
Basic network working experience.
Experience in scripting and automation to streamline deployment processes.
Please ensure the role type defined below is appropriate for your needs before applying to this role. This position is classified as:
Onsite: Employees who are working in Onsite roles will work primarily onsite. This includes all production and maintenance employees, as they are essential to the development of our products.]
Hybrid: Employees who are working in Hybrid roles will work regularly both onsite and offsite. Ratio of time working onsite will be determined in partnership with your leader.
Remote: Employees who are working in Remote roles will work primarily offsite (from home). If you live within a reasonable commute of an RTX site with other colleagues you interact with, your manager will discuss whether there is a degree of onsite presence associated with this role.
Candidates will learn more about role type and current site status throughout the recruiting process. For onsite and hybrid roles, commuting to and from the assigned site is the employee’s personal responsibility.
Understands basic management approaches such as work scheduling,prioritizing, coaching and process execution.
Typically requires specialized knowledge of technical or operational practices.
Typically requires: A University Degree or equivalent experience and minimum 5 years prior relevant experience, or An Advanced Degree in a related field and minimum 3 years experience
Engineering/Other Technical Positions: Typically requires a degree in Science, Technology, Engineering or Mathematics (STEM) and a minimum of 5 years of prior relevant experience unless prohibited by local laws/regulations.
RTX adheres to the principles of equal employment. All qualified applications will be given careful consideration without regard to ethnicity, color, religion, gender, sexual orientation or identity, national origin, age, disability, protected veteran status or any other characteristic protected by law.
Privacy Policy and Terms:
Click on this link to read the Policy and Terms