Elasticsearch jobs in London
- CGILondon
- This role is focused on improving the reliability, scalability, observability, and operational performance of critical data-driven platforms and services across…
- As part of the Forensics & Integrity Services, our Forensic Data Analytics team provides advanced analytics services to support high profile and sensitive…
- View all EY jobs - London jobs - Forensic Analyst jobs in London
- Salary Search: Assurance - Financial Services - Forensic Data Analyst - Senior - London salaries in London
- See popular questions & answers about EY
- Capgemini SogetiLondon
- This role focuses on building OSS intelligence and AIOps capabilities to support autonomous network operations.
- OSS assurance and inventory systems.
- View all Capgemini Sogeti jobs - London jobs - Network Engineer jobs in London
- Salary Search: OSS & Autonomous Network Engineer salaries in London
- See popular questions & answers about Capgemini Sogeti
- As the Data & AI Ops Lead, you will be responsible for the strategic design, implementation, and operational management of the ServiceNow CMDB, data, and AI Ops…
- View all Capgemini jobs - London jobs - Operations Lead jobs in London
- Salary Search: Data & AI Network Ops Lead salaries in London
- See popular questions & answers about Capgemini
- London Stock Exchange GroupLondon
- This role is ideal for hands‑on architects who want to be individual contributors and technical leaders.
- You’ll work at the intersection of data modeling,…
View similar jobs with this employerNoirLondon EC3R 8EE- Company pension
- NET MAUI, GitHub Copilot, Docker, Kubernetes, Microservices architecture, Angular 21, Vue.js, TypeScript, Programmer, Full Stack Engineer, Architect, .
View similar jobs with this employerNoirLondon EC3R 8EE- Company pension
- On-site gym
- NET 10.0, C# 14, Azure, JavaScript, Agile - London.
- NET Core, C# 14, Blazor, JavaScript, React, Microservices, Azure, ASP.
- NET, .NET Core, C# and Azure SQL.
- CGILondon
- Company pension
- Although we would like candidates to have all the skills we need, we would consider high quality individuals who meet most of the criteria.
View similar jobs with this employerNoirLondon EC3R 8EE- Gym membership
- Company pension
- Flexible schedule
- .NET Developer - Social Messaging Platform - London.
- NET MAUI, GitHub Copilot, Docker, Kubernetes, Microservices architecture, Angular 21, Vue.js, TypeScript,…
View similar jobs with this employerNoirLondon EC3R 8EE- Gym membership
- Company pension
- Private medical insurance
- Flexible schedule
- NET MAUI, GitHub Copilot, Docker, Kubernetes, Microservices architecture, Angular 21, Vue.js, TypeScript, Programmer, Full Stack Engineer, Architect, .
- CanonicalLondon
- In addition to defining the infrastructure as code, you will improve Canonical products and the open-source technologies they're based on by providing critical…
- View all Canonical jobs - London jobs - Site Reliability Engineer jobs in London
- Salary Search: Site Reliability / Gitops Engineer salaries in London
- See popular questions & answers about Canonical
- CanonicalLondon
- In addition to defining the infrastructure as code, you will improve Canonical products and the open-source technologies they're based on by providing critical…
- View all Canonical jobs - London jobs - Site Reliability Engineer jobs in London
- Salary Search: Site Reliability / Gitops Engineer salaries in London
- See popular questions & answers about Canonical
- Bank of AmericaBromley
- Engineers are encouraged to challenge existing approaches and contribute to the evolution of our technology strategy and platform.
- View all Bank of America jobs - Bromley jobs
- Salary Search: Senior Engineer salaries in Bromley
- See popular questions & answers about Bank of America
- Wells FargoLondon EC4R 9AT
- Wells Fargo is seeking a Principal Engineer in Corporate and Investment Banking to serve as a senior technical authority for AI solutions.
- View all Wells Fargo jobs - London jobs - Principal Software Engineer jobs in London
- Salary Search: Principal Engineer, AI Engineering salaries in London
- See popular questions & answers about Wells Fargo
- Dazzle: Commercial & Office Cleaning | LondonLondon SW7
- With the latest round of fundraising, they're now looking for a rockstar, jack of all trades backend developer to help to build the next generation of the apps…
- Sopra SteriaHemel Hempstead HP2 7AH
- We're looking for an experienced Elastic Subject Matter Expert (SME) to lead the design, implementation and evolution of a critical Elastic Stack environment…
Job Post Details
Site Reliability Engineer - job post
Job details
Job type
- Permanent
- Full-time
Location
Full job description
We are seeking an experienced and proactive Site Reliability Engineer (SRE) to join a team supporting multiple data product and platform groups. This role is focused on improving the reliability, scalability, observability, and operational performance of critical data-driven platforms and services across complex production environments.
The successful candidate will work closely with engineering, platform, and support teams to strengthen monitoring and alerting capabilities, improve logging and traceability, troubleshoot production incidents, support deployments, and automate operational processes wherever possible. The environment includes Kubernetes, Helm, the ELK stack, and a strong focus on modern Site Reliability Engineering practices across cloud and platform services.
This is a hands-on technical role suited to someone who thrives in fast-paced operational environments and is passionate about reliability engineering, automation, and continuous improvement. The role requires strong collaboration with both client stakeholders and engineering teams to ensure platform stability, operational excellence, and high service availability
Your future duties and responsibilities
- Support, maintain, and improve highly available production platforms and services across cloud and containerised environments.
- Manage and support Kubernetes clusters and Helm-based deployments across multiple environments.
- Implement and enhance monitoring, alerting, logging, and observability solutions to improve platform reliability and operational visibility.
- Investigate incidents, analyse logs, identify root causes, and drive timely resolution of production issues.
- Participate in incident response, post-incident reviews, and continuous operational improvement initiatives.
- Automate operational tasks and repetitive support activities to reduce manual effort and improve platform efficiency.
- Work closely with engineering and data platform teams to improve system resilience, scalability, deployment reliability, and operational maturity.
- Develop and maintain operational documentation, support procedures, runbooks, and troubleshooting guides.
- Contribute to reliability engineering practices including proactive monitoring, service health management, and operational readiness.
- Support deployment activities, release processes, and production change management activities.
Required qualifications to be successful in this role
- Strong commercial experience in Site Reliability Engineering, Platform Engineering, DevOps, or Production Support environments.
- Strong hands-on experience with Kubernetes and Helm in enterprise or production environments.
- Proven experience supporting mission-critical production platforms and operational support functions.
- Strong hands-on experience with the ELK stack (Elasticsearch, Logstash, Kibana) for logging, monitoring, troubleshooting, and operational analysis.
- Demonstrated capability in log analysis, incident investigation, troubleshooting, and root cause analysis.
- Strong understanding and practical experience with core SRE practices including:
Incident management and response
Root cause analysis and post-incident reviews
Automation and operational improvement
Production support and reliability engineering
- Experience working with data platforms, analytics platforms, or data product teams would be highly advantageous.
- Experience with scripting and automation tools such as Bash, Python, or similar technologies is desirable.
- Exposure to CI/CD pipelines, Infrastructure as Code, and cloud-native environments would be beneficial.
- Strong communication, stakeholder engagement, and collaboration skills.
- Ability to work effectively in fast-paced support environments and manage competing priorities under pressure.
Security Clearance
- Resource must be willing and able to work onsite at the client location five days per week.
- Candidate must already hold current HLC clearance (mandatory requirement).
- Previous experience working within secure, government, defence, or highly regulated environments will be highly regarded.
- Due to client security requirements, only candidates meeting the required clearance criteria will be considered.
#LI-CGISDI
Together, as owners, let’s turn meaningful insights into action.
Life at CGI is rooted in ownership, teamwork, respect and belonging. Here, you’ll reach your full potential because…
You are invited to be an owner from day 1 as we work together to bring our Dream to life. That’s why we call ourselves CGI Partners rather than employees. We benefit from our collective success and actively shape our company’s strategy and direction.
Your work creates value. You’ll develop innovative solutions and build relationships with teammates and clients while accessing global capabilities to scale your ideas, embrace new opportunities, and benefit from expansive industry and technology expertise.
You’ll shape your career by joining a company built to grow and last. You’ll be supported by leaders who care about your health and well-being and provide you with opportunities to deepen your skills and broaden your horizons.
Come join our team—one of the largest IT and business consulting services firms in the world.