Job matched to your search
SENIOR SYSTEM DEPLOYMENT ENGINEER
Collins Aerospace · Singapore
Free · Join 5,000+ job seekers using Qarera
How well do you match this role?
Tap the skills you already have — then see your real match score, what’s missing, and your resume fixed for this job.
Job description
Date Posted:
2026-08-14Country:
Singapore Location:
SG-01-SINGAPORE-083 ~ 83 Clemenceau Ave ~ UE SQAURE Position Role Type:
Onsite At RTX, the world's largest aerospace and defense company, 185,000 great minds are united by purpose and inspired to make a difference solving the world’s most complex problems. With our three market leading businesses, world-class operations and investments in research and development, we offer capabilities and opportunity no one else can. Together, we push the boundaries of known science and find new ways to connect and protect our world.
Collins Aerospace is a leader in technologically advanced, intelligent solutions that help redefine the aerospace and defense industry. With a comprehensive portfolio and deep technical expertise, we help customers meet the demands of the global market. Join us and help shape the future of aerospace and defense.
Senior System Deployment Engineer
Role Overview
We are seeking an experienced Senior System Deployment to join our team. The ideal candidate will be responsible for the deployment, configuration, testing, recovery, and ongoing support of mission-critical airline and airport applications across Windows and Linux environments.
The role covers the complete lifecycle of application infrastructure, including server staging, OS installation, VM provisioning, application deployment, database configuration, high-availability and failover testing, disaster recovery, and production transition.
A key responsibility is providing critical second-level/third-level technical support for major incidents, including the ability to recover and rebuild corrupted virtual machines, operating systems, databases, and application environments from scratch using available backups, recovery media, infrastructure templates, and documented recovery procedures.
The engineer will work across Windows Server, Red Hat Linux, VMware, Hyper-V, AWS Cloud, SQL Server, PostgreSQL, Kafka, MongoDB, OpenShift, and related infrastructure technologies, supporting highly available and operationally critical airport systems.
Key Responsibilities
1. Server Staging, Build & Deployment
- Stage, provision, and prepare Windows and Linux servers for application deployment.
- Perform OS installation, hardening, patching, configuration, and baseline validation.
- Deploy and configure airline and airport applications across physical and virtual environments.
- Configure application services, ports, certificates, connectivity, service accounts, and system dependencies.
- Ensure servers meet required performance, security, availability, and operational standards.
- Prepare production, DR, test, FAT, UAT, and staging environments.
2. Critical Incident Support & Infrastructure Recovery
- Provide second-level/third-level technical support for mission-critical airport and airline systems.
- Participate in major incident, problem management, and ITSM processes.
- Troubleshoot critical failures involving servers, VMs, operating systems, databases, applications, storage, and infrastructure services.
- Lead or support full infrastructure recovery when systems become corrupted or unrecoverable.
- Restore and rebuild VMs from scratch, including:
o Recreating virtual machines.
o Reinstalling operating systems.
o Applying required OS configuration and security hardening.
o Restoring application components and services.
o Reconfiguring network, storage, DNS, certificates, and connectivity.
o Restoring application configuration and validating dependencies.
- Recover corrupted or failed servers using available VM backups, snapshots, images, templates, backup repositories, or DR infrastructure.
- Perform complete application environment restoration when the original server/VM cannot be recovered.
- Work with infrastructure, backup, network, database, security, and application teams during major recovery activities.
- Conduct post-recovery validation and ensure systems are returned to operational readiness.
3. Database Recovery & Restoration
- Support SQL Server and PostgreSQL environments supporting mission-critical applications.
- Troubleshoot database corruption, service failures, connectivity issues, performance issues, and abnormal database behavior.
- Perform or coordinate database restoration from backup following database corruption or infrastructure failure.
- Restore databases from full, differential, transaction-log, WAL, or other applicable backup mechanisms.
- Rebuild database servers when required, including:
o OS and VM restoration.
o Database software installation.
o Database configuration.
o Storage and permissions configuration.
o Database restoration.
o User, role, service account, and connectivity configuration.
o Application reconnection and validation.
- Validate database integrity following restoration.
- Support database failover, recovery, replication, and DR testing.
- Coordinate with DBA teams for complex database recovery and corruption scenarios.
- Ensure appropriate RPO/RTO requirements are achieved during recovery.
4. Virtualization & Infrastructure Management
- Manage and troubleshoot VMware and Hyper-V virtual environments.
- Provision, configure, resize, clone, restore, and recover virtual machines.
- Troubleshoot VM, host, datastore, virtual network, resource allocation, and performance issues.
- Perform VM recovery from backup or rebuild VMs from scratch when required.
- Support VMware HA, clustering, snapshots, templates, and DR-related activities.
- Investigate virtualization-related failures affecting critical applications.
- Support data center infrastructure troubleshooting involving compute, storage, virtualization, and server connectivity.
5. High Availability, Failover & Disaster Recovery
- Design, execute, and document server, application, database, and VM failover testing.
- Conduct planned and unplanned failover exercises.
- Validate application recovery following infrastructure or database failures.
- Perform DR restoration and recovery testing.
- Validate RPO and RTO against operational requirements.
- Test recovery procedures for complete server/VM loss and database corruption scenarios.
- Identify recovery gaps and recommend improvements to HA/DR architecture.
- Maintain detailed recovery procedures and runbooks.
6. FAT, UAT & Production Validation
- Plan and execute Factory Acceptance Testing (FAT) and User Acceptance Testing (UAT).
- Perform technical validation before applications are transitioned into production.
- Validate server, application, database, network, security, and infrastructure dependencies.
- Conduct failover, recovery, performance, and resilience testing.
- Support application cutover and go-live activities.
- Ensure production handover is completed with appropriate documentation and support procedures.
7. Database & Middleware Services
- Configure and support:
o Microsoft SQL Server
o PostgreSQL
o Kafka
o MongoDB
o OpenShift
o Related application middleware and services
- Troubleshoot application-to-database connectivity and service dependencies.
- Support Kafka brokers, topics, connectivity, and service availability.
- Support MongoDB configuration, availability, and recovery activities.
- Work with development and application teams to resolve middleware and database-related incidents.
8. AWS Cloud & Hybrid Infrastructure
- Support applications deployed across on-premises data centers and AWS Cloud.
- Troubleshoot cloud infrastructure, connectivity, compute, storage, and application dependencies.
- Support VM/server recovery and application restoration in hybrid environments.
- Assist with cloud DR and backup/recovery activities.
- Understand connectivity between on-premises infrastructure and AWS environments.
9. Monitoring, Troubleshooting & Operational Support
- Monitor system health, availability, capacity, and performance.
- Analyze server, application, database, and infrastructure logs.
- Troubleshoot CPU, memory, disk, network, service, database, and application failures.
- Work with monitoring and ITSM tools to identify and resolve recurring incidents.
- Participate in root cause analysis (RCA) and problem management.
- Identify opportunities for automation and proactive monitoring.
10. Scripting & Automation
- Develop scripts and automation to reduce manual deployment and recovery activities.
- Use PowerShell, Bash, Python, or similar scripting technologies where appropriate.
- Automate server health checks, service validation, deployment activities, backup validation, and recovery procedures.
- Develop automated health-check and auto-healing capabilities for critical services where feasible.
11. Documentation & Knowledge Management
- Create and maintain:
o Server build documents
o VM configuration documents
o Application deployment procedures
o Database recovery procedures
o VM recovery procedures
o Disaster recovery runbooks
o Failover test procedures
o FAT/UAT test cases
o Troubleshooting guides
o Operational handover documents
- Maintain accurate recovery procedures to enable complete rebuild and restoration from a failed/corrupted environment.
- Ensure documentation is regularly tested and updated following system changes or incidents.
12. Collaboration
- Collaborate with application development, DBA, network, security, cloud, infrastructure, service desk, and operations teams.
- Coordinate with vendors and technology partners during critical incidents.
- Participate in technical reviews, change management, CAB activities, and production implementation.
- Provide technical guidance to junior engineers and operations teams.
Qualifications
- Bachelor’s degree in Computer Science, Information Technology, Engineering, or related discipline.
- 8+ years of experience in system deployment, infrastructure support, system recovery, or mission-critical IT environments.
- Strong experience with Windows Server and Red Hat Linux.
- Strong experience in server staging, OS build, configuration, deployment, troubleshooting, and recovery.
- Hands-on experience with VMware and/or Hyper-V.
- Strong experience with VM restoration, rebuilding, and disaster recovery.
- Experience recovering databases and application environments following corruption or infrastructure failure.
- Experience with SQL Server and PostgreSQL.
- Working knowledge of Kafka, MongoDB, and OpenShift.
- Experience with high-availability systems, failover, backup, restore, and DR.
- Experience conducting FAT, UAT, failover, recovery, and production validation testing.
- Working knowledge of AWS Cloud and hybrid infrastructure.
- Experience with ITSM, incident management, change management, and problem management.
- Basic networking knowledge including TCP/IP, DNS, DHCP, firewall rules, routing, VLANs, and connectivity troubleshooting.
- Experience with scripting/automation using PowerShell, Bash, Python, or equivalent.
- Relevant certifications such as Microsoft, VMware, Red Hat, AWS, or other senior infrastructure certifications are desirable.
- Strong analytical, troubleshooting, communication, and documentation skills.
Preferred Skills
- Prior experience supporting mission-critical airline, airport, transportation, other 24x7 environments.
- Experience supporting systems with strict availability, RTO, and RPO requirements.
- Hands-on experience with complete infrastructure recovery from bare/clean VM or server build through application restoration and production validation.
- Experience with VMware HA, clustering, snapshots, templates, and backup/restore solutions.
Basic network working experience.
Experience in scripting and automation to streamline deployment processes.
Please ensure the role type defined below is appropriate for your needs before applying to this role. This position is classifi
More jobs in Singapore
Browse related jobs
Don’t just read the job — see if you’ll get it.
Get your match score, a resume tailored to this exact role, and jobs like it — free.
Check my fit for this job