Senior Cloud Infrastructure Engineer

Skills & Experience

  • Job roles: Cloud Engineer

  • Experience level: Senior

  • Tech stack/tooling used: GCP, Terraform, Ansible, Python, CI/CD, Prometheus, Grafana, Datadog, Bash, Cyber Security

  • Core skills considered: GCP, Terraform, Ansible, Python, CI/CD

  • Other skills considered: Prometheus, Grafana, Datadog, Bash, Cyber Security

Logistics

  • Base salary: Undisclosed

  • Employment type: Permanent

  • Remote working: Hybrid

  • Visa sponsorship: Not available

Job Description

Fundment is a fast-growing wealth infrastructure company, building on our success in transforming the £3 trillion UK wealth management market with our cutting-edge digital investment system. We are passionate about revolutionising the investment experience for financial advisers and their clients by combining innovative proprietary technology with exceptional customer service.

About the Role

We are seeking a Senior Cloud Infrastructure Engineer to join our team to focus on the performance, resilience and security of our Cloud Infrastructure. Reporting directly to our Head of IT Infrastructure, you will be responsible for developing and implementing tools and processes that ensure the high availability, performance and security of our market-leading investment system running on Google Cloud Platform (GCP). You will work closely with colleagues across our infrastructure, engineering and product teams to continually improve our rapid deployments while maintaining operational excellence.

Key Responsibilities:

  • System Resilience and Optimisation: Design, build and maintain reliable, secure and scalable cloud infrastructure on Google Cloud Platform.

  • Monitoring and Alerting: Develop and enhance comprehensive monitoring, alerting, and logging systems to allow us to proactively detect issues.

  • Automated Change Management and Infrastructure as Code (IaC): Automate repetitive tasks, deployments, and operational workflows to streamline processes and reduce human error by implementing Infrastructure as Code (IaC) practices using Terraform and similar tools.

  • Incident Management and Troubleshooting: Respond to incidents, triage and resolve system issues, and conduct root cause analysis to prevent recurrence. Act as an escalation point for critical incidents and lead post-incident reviews. Security and Compliance: Continually drive improvements in our cybersecurity posture and work closely with our security advisers to implement security best practices and ensure compliance with industry regulations and standards.

  • Collaboration and Mentorship: Collaborate with software engineers, product managers and infrastructure colleagues to promote and provide guidance on cloud infrastructure, platform reliability, operational best practices and cybersecurity.

Required Skills/Experience:

  • 5+ years of experience in infrastructure engineering, site reliability engineering, DevOps, or a related role.

  • Extensive experience with Google Cloud Platform (GCP) services including VPC Service Control and Cloud Run.

  • Advanced knowledge of Infrastructure as Code tools (e.g., Terraform, Ansible).

  • Deep understanding of monitoring, logging, and observability practices, with experience in tools such as Prometheus, Grafana, Datadog, Google Cloud Monitoring.

  • Strong scripting skills in Python, Bash, or similar languages.

  • Strong understanding of cybersecurity best practices, across both cloud infrastructure and end user devices.

  • Familiarity with SRE principles, including Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets.

  • Excellent problem-solving skills with the ability to troubleshoot complex issues under pressure.

  • Effective communication skills (both written and verbal) with the ability to collaborate with both technical and non-technical teams

  • Expertise in CI/CD pipelines, especially with tools like GitLab CI/CD or similar.

Preferred Skills/Experience:

  • Experience of cybersecurity best practices and regulatory requirements in the financial sector, including SOC 2 and ISO 27001.

  • Experience of leading projects to improve security, efficiency or operational resilience.

  • Proficiency with deploying Kubernetes at scale in production environments.

  • Strong understanding of High Availability Enterprise Cloud SQL deployments using MySQL or PostgreSQL databases

  • Deep understanding of network operations and the ability to configure load balancers, firewalls and CDNs

  • Experience with Microsoft Azure, Entra ID and Microsoft 365

Company Benefits

  • Be part of a modern, inclusive, high-trust engineering culture

  • Take ownership and ship code that directly improves client outcomes

  • Work with a smart, friendly team that values balance, growth, and support

  • Pension 6% employer contribution, minimum 2% employee contribution.

  • BUPA Private Health Insurance – fully paid for by the company, for you and your immediate family.

  • Medicash Cashplan – fully paid for by the company, for you and your immediate family.

  • Travel insurance – fully paid for by the company, for you and your immediate family.

  • Life Assurance – 4 x base salary.

  • Employee Assistance Programme

  • 28 days annual leave plus bank holidays.

  • Paid compassionate leave – up to 5 days per year.

  • Enhanced paternity/maternity/adoption leave

  • Jury service – 10 days at full pay.

  • Hybrid working arrangements – 3 days per week in the Fitzrovia office.

  • Coaching&Counselling sessions

  • Training Budget

  • Annual pay review

  • Annual training budget