Talent.com
JPMorganChase
Senior Lead Site Reliability DevOps EngineerJPMorganChase • Glasgow, Scotland, UK
Senior Lead Site Reliability DevOps Engineer

Senior Lead Site Reliability DevOps Engineer

JPMorganChase • Glasgow, Scotland, UK
27 days ago
Job type
  • Full-time
Job description
Description

Be an integral part of an agile team thats constantly pushing the envelope to enhance build and deliver top-notch reliability and observability for our most critical platforms.

As a Senior Lead Site Reliability / DevOps Engineer at JPMorgan Chase within the Commercial & Investment Bank you are an integral part of an agile team that works to enhance build and deliver trusted market-leading technology products in a secure stable and scalable way. Drive significant business impact through your capabilities and contributions and apply deep technical expertise and problem-solving methodologies to tackle a diverse array of reliability observability and performance challenges that span multiple technologies and applications.

Job responsibilities

  • Regularly provides technical guidance and direction on site reliability practices to support the business and its technical teams contractors and vendors
  • Develops secure and high-quality production code for reliability tooling and telemetry pipelines and reviews and debugs code written by others
  • Drives decisions that influence reliability design observability architecture application functionality and technical operations and processes
  • Serves as a function-wide subject matter expert in one or more areas of site reliability observability or telemetry engineering
  • Leads resiliency design reviews and breaks up complex reliability problems into digestible work for other engineers acting as a technical lead for large-sized products
  • Acts as the main point of contact during major incidents demonstrating the skills to identify and solve issues quickly to avoid financial losses and champions blameless postmortem culture
  • Collaborates with team members and stakeholders to define comprehensive service level indicators service level objectives and error budgets
  • Designs implements and maintains operational reliability for large-scale OpenTelemetry pipelines on hybrid on-prem/cloud environments supporting telemetry ingestion processing and export to backends such as InfluxDB Prometheus Elasticsearch and OpenSearch
  • Drives the assessment refactoring and incremental migration of custom legacy telemetry collection code to standardized OpenTelemetry instrumentation reducing technical debt while maintaining system stability
  • Actively contributes to the engineering community as an advocate of firmwide frameworks tools and practices and influences peers and project decision-makers to consider the use and application of leading-edge observability and reliability technologies
  • Adds to the team culture of diversity opportunity inclusion and respect

Required qualifications capabilities and skills

  • Formal training or certification on software engineering concepts and advanced applied experience delivering system design application development testing and operational stability
  • Advanced knowledge of reliability scalability performance security enterprise system architecture toil reduction and other site reliability best practices with considerable in-depth knowledge in one or more technical disciplines (e.g. cloud observability distributed systems etc.)
  • Advanced proficiency in one or more programming languages (e.g. Java Python Go etc.)
  • Advanced proficiency and experience in observability such as white and black box monitoring SLO alerting and telemetry collection using tools such as Grafana Dynatrace Prometheus Datadog Splunk Elasticsearch etc.
  • Proficiency in continuous integration and continuous delivery tools (e.g. Jenkins GitLab Terraform etc.)
  • Experience with container and container orchestration (e.g. ECS Kubernetes Docker etc.)
  • Hands-on experience with the design deployment and operation of OpenTelemetry collectors in production environments focusing on technical aspects such as configuring optimizing and troubleshooting OTLP endpoints and receivers
  • Ability to tackle reliability design and functionality problems independently with little to no oversight
  • Practical cloud native experience
  • Ability to expand and collaborate across different levels and stakeholder groups

Preferred qualifications capabilities and skills

  • Knowledge of distributed tracing metrics and logging best practices
  • Certification in AWS Kubernetes or relevant technologies
  • Proven track record in system health monitoring capacity management and blameless postmortems for high-availability services
  • Deep understanding of distributed system design principles networking (TCP/IP DNS load balancing) and Linux internals
  • Contributions to open-source observability or telemetry projects
  • Experience working with agent control planes and management protocols; hands-on knowledge of OpAMP is highly desirable



Required Experience:

Senior IC


Employment Type : Full-Time
Experience: years
Vacancy: 1
Create a job alert for this search

Senior Lead Site Reliability DevOps Engineer • Glasgow, Scotland, UK

Similar jobs

Senior Lead Site Reliability / DevOps Engineer

J.P. MORGANGlasgow, Scotland, GB
Full-time

Be an integral part of an agile team that's constantly pushing the envelope to enhance, build, and deliver top‑notch reliability and observability for our most critical platforms.As a Senior Lead S... Show more

 • Promoted

Dev Ops Engineer

FPSG ConnectGlasgow, Scotland, GB
Full-time

Are you a highly technical, hands-on developer with a deep passion for SDLC tooling and processes? We're seeking a skilled DevOps Developer who excels in automation, development, and continuous int... Show more

 • Promoted

Lead Engineer - Protection Systems

MeritusGlasgow, Scotland, GB
Full-time

Lead Engineer - Protection Systems.Meritus, Glasgow, Scotland, United Kingdom.Lead Engineer role focused on the design, development, and delivery of protection systems for the UK's energy infrastru... Show more

 • Promoted

Senior Systems Engineer (Team Lead): Azure, AD & Security

ARCH EUROPE INSURANCE SERVICES LTDGlasgow, Scotland, GB
Full-time

A leading insurance services company in the UK is seeking a Senior Systems Engineer (Team Lead) to oversee infrastructure management and lead security initiatives.The ideal candidate will have a mi... Show more

 • Promoted

Senior Lead Site Reliability Engineer

J.P. MorganGlasgow, Schottland, GB
Full-time

Be an integral part of an agile team that's constantly pushing the envelope to enhance, build, and deliver top-notch reliability and observability for our most critical platforms.As a Senior Lead S... Show more

 • Promoted

Senior Engineer (Cybersecurity)

SYSTRA UK & IrelandGlasgow, Scotland, GB
Full-time

Around the world, SYSTRA’s specialists plan, design, integrate, test, commission, project manage and deliver mass transit and mobility solutions that are relied on by more than 50 million people ev... Show more

 • Promoted

Senior Platform Engineer - Cloud & DevX Leader

Madison LogicGlasgow, Scotland, GB
Full-time

A leading marketing technology firm in Glasgow is looking for a talented developer proficient in programming languages like NodeJS and Python, with extensive experience in Kubernetes and observabil... Show more

 • Promoted

DevOps Engineer

CititecGlasgow, Scotland, GB
Full-time

Location: Glasgow – 2–3 days per week onsite.My client is searching for an Automation & Observability Engineer to help drive automation initiatives and develop observability and alerting solutions ... Show more

 • Promoted

Hybrid UCaaS Deployment Engineer

GammaGlasgow, Scotland, GB
Full-time

Gamma is seeking an UCaaS Implementation Engineer to design and implement standard UCaaS solutions across our Cloud Platforms, including Microsoft Teams Operator Connect and Direct Routing, Horizon... Show more

 • Promoted

Global SRE Engineer: Storage & Data Reliability

OracleScotland, GB
Full-time

Oracle in the United Kingdom is seeking an experienced SRE to support OCI storage and data services, focusing on reliability, operations, automation, and scalable production systems.The role involv... Show more

 • Promoted

Senior Platform Engineer: Threat Hunt & DevSecOps

Morgan StanleyGlasgow, Scotland, GB
Full-time

Morgan Stanley is looking for a Senior Platform Engineer to join their Cyber Data Risk & Resilience division in Glasgow.This role involves serving as a DevSecOps point of contact, monitoring alerts... Show more

 • Promoted

Senior Systems Engineer – Automation & Low-Latency Infra

ProvnScotland, GB
Full-time

A leading technology solutions provider in the United Kingdom is seeking a Senior System Engineer.The ideal candidate will have extensive experience in Linux administration, automation (especially ... Show more

 • Promoted

Site Engineer - Glasgow

CareysGlasgow, Scotland, GB
Full-time

We have an exciting opportunity for you to join a highly skilled team delivering several new projects across the Glasgow region.As a Site Engineer, you will support the technical delivery of these ... Show more

 • Promoted

Site Reliability Engineer

Paritas RecruitmentGlasgow, Scotland, GB
Full-time

AWS Site Reliability Engineer (Data Platform) – Contract.Contract Length: February 2026 – January 2027.We are recruiting an AWS Site Reliability Engineer (SRE) to support a cloud-native data platfo... Show more

 • Promoted

Senior Software Engineer - Space Reliability

UK Space JobsGlasgow, Scotland, GB
Full-time

Name Your Satellite Program (NYSP).Spire is making a fundamental shift in how it operates its constellation.We are moving from a model where trained operators watch dashboards and escalate to exper... Show more

 • Promoted

Senior Software Engineer - Space Reliability

SpireGlasgow, Scotland, GB
Full-time

Spire is making a fundamental shift in how it operates its constellation.We are moving from a model where trained operators watch dashboards and upscale to experts, to one where the system is fully... Show more

 • Promoted

Site Engineering Lead

Finsbury FoodEast Kilbride, Scotland, GB
Full-time

Johnstones Food Service Ltd, 3 Redwood Place, East Kilbride, United Kingdom, United Kingdom.Johnstone’s Food Service, East Kilbride.We have an opportunity for an experienced Site Engineering Lead t... Show more

 • Promoted

Senior Site Engineer og Site Engineer, Frigate Site office UK

Norwegian Defence Materiel AgencyGlasgow, Scotland, GB

Forsvarsmateriell er Norges største offentlige investeringsaktør med en omsetning i 2026 på over 30 mrd.Forsvarsmateriell er en etat i forsvarssektoren som anskaffer og leverer materiell og tjenest... Show more

 • Promoted

Lead Site Reliability Engineer

慨正橡扯Glasgow, Scotland, United Kingdom
Full-time

Assume a critical role in defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability.As a Lead Site Reliabi... Show more

 • Promoted

Azure DevOps Engineer

OscarGlasgow, Scotland, GB
Full-time

This range is provided by Oscar.Your actual pay will be based on your skills and experience — talk with your recruiter to learn more.Direct message the job poster from Oscar.You'll operate at the i... Show more