Platform Engineer/Senior Platform Engineer
Save this job and keep your search organized
Create a free account to save jobs, create alerts and return to this listing from your dashboard.
By continuing, you agree to our Terms & Privacy Policy.
The Next Generation Programme Office (NGPO) Strategy branch is looking for motivated platform engineers with a collaborative, proactive attitude and a passion for continuous learning. You will be part of a team focused on designing, implementing, and maintaining either cloud infrastructure and platform services OR on-premises infrastructure and platform services for the Next Generation ANS (Air Navigation System), depending on the specific deployment requirements. If you are an adaptable, proactive platform engineer with a passion for best practices and continuous learning, we would love to hear from you. Join us and contribute to building robust, scalable infrastructure solutions for the future
What you will be working onKey responsibilities include:
- Design & Maintain Mission-Critical Infrastructure: Architect, implement, and maintain platform solutions for safety-critical applications, ensuring compliance with industry standards for reliability, availability, and performance in either cloud or on-premises environments.
- Platform Management: Work with either cloud-native technologies and services (AWS, Azure) including compute, storage, networking, and managed services, OR on-premises infrastructure including virtualisation platforms, bare metal servers, and traditional networking solutions, adapting quickly to new technologies as needed.
- Infrastructure as Code & Automation: Implement infrastructure automation using Infrastructure as Code tools and configuration management systems to ensure consistent, repeatable deployments in your designated environment.
- Container Orchestration & Platform Services: Design and manage containerised environments using container orchestration platforms and related ecosystem tools to support application deployment and scaling.
- Collaborate with Cross-Functional Teams: Work closely with software developers, architects, and other stakeholders to design, implement and deploy scalable platform solutions. May also perform Site Reliability Engineering (SRE) functions as part of the role responsibilities.
- Monitoring & Observability: Implement comprehensive monitoring, logging, and alerting solutions to ensure platform health, performance, and early issue detection across all environments.
- Security & Compliance: Apply security best practices, implement security controls, and ensure compliance with regulatory requirements for aviation systems in your designated platform environment.
- Data Platform Operations: Maintain and optimise data pipeline infrastructure, streaming platforms, and analytics workloads as secondary responsibility.
- Deployment & Integration: Perform platform deployment, integration testing, and validation in production environments, ensuring seamless service delivery.
- Troubleshooting & Performance Optimisation: Proactively identify and resolve infrastructure, performance, and reliability issues across development and production environments.
- Documentation & Knowledge Sharing: Maintain thorough documentation for infrastructure, processes, and operational procedures, ensuring knowledge transfer and compliance requirements are met.
Trained in Computer Science, Information Technology, Engineering or equivalent. Platform engineering and infrastructure management experience in either cloud or on-premises environments. Strong knowledge of either cloud-native architectures and microservices patterns OR traditional infrastructure patterns and distributed systems design.
Core technical experience with: Infrastructure as Code tools (e.g., Terraform, CloudFormation) Container orchestration platforms (e.g., Kubernetes, Docker Swarm, OpenShift) Package management and deployment tools (e.g., Helm, Kustomize) Monitoring and observability stack (e.g., Prometheus/Grafana, ELK stack, Datadog)
For Cloud Platform Focus - experience with: Major cloud platforms (AWS, Azure) Cloud-native services and managed solutions
For On-Premises Platform Focus - experience with: Virtualisation platforms (e.g., VMware vSphere, Hyper-V, KVM) Configuration management tools (e.g., Ansible, Puppet, Chef) Bare metal server management and data centre operations Knowledge of CI/CD pipelines, DevOps practices, and automation tools. Understanding of cybersecurity concepts, including network security, identity management, and compliance frameworks. Knowledge of networking concepts including VPCs/VLANs, load balancers, DNS, and service mesh technologies. Proficiency in scripting and automation languages (e.g., Python, Bash, PowerShell, Go).
Engineer desired skills and experience Having Site Reliability Engineering (SRE) background and experience, including familiarity with service level objects, incident management, and production system reliability practices. Experience with disaster recovery, backup st