All Locations
London

About The Role

FDM is a global business and technology consultancy seeking a Site Reliability Engineer to work for our client within the Finance sector. This is initially a 6 month contract with very good prospects to extend and will be a hybrid role that will be based in London.

Our client is seeking an experienced Site Reliability Engineer (SRE) with a strong focus on Observability and Monitoring Platforms. The successful candidate will play a key role in enhancing the organisation's monitoring, alerting, and operational visibility capabilities across critical engineering systems.
 
This role requires hands-on expertise in the deployment, administration, and optimisation of OpenSearch, alongside experience with Grafana, Geneos, and automation/scripting technologies. Particular emphasis will be placed on the candidate's ability to design, deploy, and support enterprise-grade OpenSearch environments.
 
Responsibilities:  
 
  • Lead the design, deployment, configuration, and ongoing management of OpenSearch clusters and associated observability tooling.
  • Develop and maintain scalable monitoring, logging, and alerting solutions for business-critical applications and infrastructure.
  • Build and enhance observability dashboards using Grafana.
  • Support and optimise existing Geneos monitoring implementations.
  • Create and maintain automation scripts to streamline operational processes and improve reliability.
  • Collaborate with engineering, infrastructure, and support teams to improve system resilience and operational performance.
  • Define and implement SRE best practices, including monitoring standards, alert management, incident response, and operational readiness.
  • Perform troubleshooting and root cause analysis of platform and application issues.
  • Support capacity planning, performance tuning, and platform optimisation initiatives.
  • Contribute to documentation, operational procedures, and knowledge sharing within the engineering team.

About You

OpenSearch (Critical Requirement)

  • Extensive hands-on experience deploying and managing OpenSearch in production environments.
  • Deep understanding of OpenSearch architecture, cluster design, indexing strategies, shard management, and performance tuning.
  • Experience implementing log aggregation, search, analytics, and observability use cases using OpenSearch.
  • Knowledge of OpenSearch security, access controls, backups, upgrades, and operational best practices.

Monitoring & Observability

  • Strong experience with Grafana, including dashboard development, alerting, and data source integration.
  • Experience with enterprise monitoring platforms, specifically Geneos.
  • Understanding of modern observability principles, including metrics, logs, traces, alerting, and service health monitoring.

Scripting & Automation

  • Strong scripting skills in one or more of:
    • Python
    • Shell/Bash
    • PowerShell
  • Experience automating operational tasks and monitoring workflows.

SRE / Platform Engineering

  • Proven experience in an SRE, Platform Engineering, DevOps, or Infrastructure Engineering role.
  • Strong troubleshooting and problem-solving capabilities.
  • Experience supporting highly available and business-critical systems.
  • Understanding of incident management, resilience engineering, and operational excellence practices.

Desirable Skills

  • Experience with cloud platforms (Azure, AWS, or GCP).
  • Knowledge of containerisation technologies (Docker, Kubernetes).
  • Experience with CI/CD pipelines and Infrastructure as Code.
  • Experience working within financial services or regulated environments.
  • Familiarity with Elasticsearch ecosystems and migration strategies to OpenSearch.

Candidate Profile

The ideal candidate will be a hands-on engineer who combines deep technical expertise with a pragmatic operational mindset. They will be comfortable working independently, driving observability improvements, and collaborating across engineering teams to deliver reliable and scalable monitoring solutions.
 
Key attributes:
 
  • Strong ownership mentality.
  • Excellent analytical and troubleshooting skills.
  • Effective stakeholder communication.
  • Ability to operate in fast-paced production environments.
  • Focus on reliability, automation, and continuous improvements

About Us

FDM is an award-winning global leader in tech and business talent solutions, backed by more than 35 years of industry experience. We have centres across Europe, North America, and Asia-Pacific, and a global workforce of over 2500 employees. FDM has shown exponential growth throughout the years, firmly establishing itself as an award-winning employer, currently listed on the FTSE4Good Index and as a 2026 Financial Times UK ‘Best Employer’. 

Diversity and Inclusion

FDM Group is an equal opportunity employer, and all qualified applicants will receive consideration for employment without regard to race, colour, religion, sex, sexual orientation, national origin, age, disability, veteran status or any other status protected by federal, provincial or local laws.

Why join us

  • Career coaching, mentoring and access to upskilling throughout your entire FDM career
  • Assignments with global companies and opportunities to work abroad
  • Opportunity to re-skill and up-skill into new areas, develop non-linear career paths and build a skillset within your field
  • Annual leave and work-place pension

Other jobs like this

All Locations
London
All Locations
UK
All Locations
London