About the Position
We are looking for an experienced Platform Engineer to support the existing AWS infrastructure and contribute to its ongoing modernisation and consolidation.
The role combines hands-on cloud engineering with developer enablement. You should understand Platform Engineering as a product discipline and be able to discuss developer experience, self-service, platform adoption, reliability, and technical trade-offs.
The working schedule is aligned with UK business hours, from 9:00 AM to 5:30 PM UK time. Participation in an on-call rota may be required.
This position includes a sign-in bonus.
About the Team
You will join a Platform Engineering team responsible for supporting and improving several AWS environments. The team treats the platform as a product and internal engineering teams as its customers. You will help engineers use the platform effectively while improving its reliability, usability, and scalability.
Responsibilities
- Support and improve AWS-based production infrastructure.
- Develop and maintain platform capabilities using Kubernetes, Helm, and Terraform.
- Automate repetitive processes and improve developer self-service.
- Support engineering teams in adopting and using platform services.
- Investigate platform issues and take ownership through resolution and follow-up.
- Participate in incident response, communication, post-incident reviews, and completion of corrective actions.
- Improve observability using metrics, logs, traces, dashboards, alerts, and service-level indicators.
- Explain technical decisions, risks, and trade-offs clearly.
- Maintain strong standards for reliability, security, testing, and documentation.
- Participate in an on-call rota when required.
- Use AI-assisted tools to improve engineering productivity and quality.
Requirements
- Strong hands-on experience with AWS production environments.
- Practical experience with Kubernetes and Terraform.
- Experience with Helm and GitHub Actions or similar tools.
- Strong understanding of Platform Engineering principles.
- Experience treating internal platforms as products and engineers as customers.
- Ability to discuss platform adoption, self-service, developer experience, and different implementation approaches.
- Good understanding of observability beyond logs, including metrics, tracing, alerting, dashboards, and SLOs.
- Experience with incident response, post-incident reviews, and operational improvements.
- Ability to work independently and take ownership of outcomes.
- Clear spoken and written English, with the ability to explain technical concepts in sufficient depth.
- Demonstrated growth mindset and ability to learn from feedback.
- Experience working in environments with regular feedback, accountability, and performance expectations.
- Experience with at least one programming or scripting language.
- Experience with observability tools and practices, including OpenTelemetry, Honeycomb, or Splunk.
Nice to Have
- Experience with ECS, S3, SNS, SQS, or RDS.
- Experience with Datadog, or OpenTelemetry.
- Experience building internal developer platforms.
- Experience mentoring or supporting the growth of other engineers.
- Experience in financial services or another regulated industry.
- Experience with cloud migration or infrastructure consolidation.
Technologies
AWS, EKS, ECS, Kubernetes, Docker, Helm, Terraform, GitHub Actions, OpenTelemetry, Honeycomb, Datadog, Claude, ChatGPT