LOG IN
SIGN UP
Canary Wharfian - Online Investment Banking & Finance Community.
Sign In
or continue with e-mail and password
Forgot password?
Don't have an account?
Join Canary Wharfian
or continue with e-mail and password
By signing up, you agree to our Terms & Conditions and Privacy Policy.

Sr Lead Infrastructure Engineer- Devops/AWS

ExperiencedNo visa sponsorship
J.P. Morgan logo

at J.P. Morgan

Bulge Bracket Investment Banks

Posted 4 days ago

No clicks

**Lead Infrastructure Engineer - Senior Vice President (AWS/DevOps) in Glasgow:** Oversee CI/CD, observability, and incident management for AI/ML platform owned by top-tier banking tech team. Manage cross-functional team to enhance reliability, security, and scalability of AI products operating worldwide. Core responsibilities include evolving CI/CD, setting SLOs, automating infrastructure, implementing security controls, and mentoring junior engineers. Required: 7+ years in DevOps/SRE, Terraform, Python, strong hands-on Kubernetes and cloud experience. Must have production security and incident response experience. Master's degree preferred.

Compensation
Not specified

Currency: Not specified

City
Not specified
Country
United Kingdom

Full Job Description

Location: GLASGOW, LANARKSHIRE, United Kingdom

We're looking for a hands-on DevOps / SRE engineer ready to take their career to new heights. Join the ranks of top talent at one of the world's most influential companies.

As a Lead DevOps / SRE Engineer - Vice President at JPMorgan Chase within the International Private Bank (IPB) Technology Artificial Intelligence and Machine Learning (AIML) Team, you will own the automation, reliability, and production operations of our agentic AI and machine learning products. As the team takes end-to-end ownership of the platforms it runs, you will build the CI/CD, observability, and incident-management practices that keep those services stable, secure, and performant across international markets.

This is a Vice President-level role and an integral part of the IPB Tech AIML team, reporting to the Head of AI, IPB Tech.

Job responsibilities

  • Owns and evolves the team's CI/CD pipelines, release automation, and deployment tooling
  • Establishes reliability practices (SLOs, error budgets, runbooks) and leads production incident response and post-incident review
  • Builds and operates observability across the team's AI/ML services (metrics, logging, tracing, alerting)
  • Automates infrastructure provisioning and configuration through infrastructure-as-code
  • Implements operational security, secrets management, and access controls to firm-wide standards
  • Partners with platform, data, and AI engineers to harden services for production and reduce dependency on external functions
  • Mentors junior engineers on DevOps and reliability practices and sets standards through review
  • Champions the firm's culture of diversity, Opportunity, inclusion, and respect

Required qualifications, capabilities, and skills

  • Formal training or certification on software engineering or systems concepts and applied experience
  • Advanced proficiency with infrastructure-as-code (e.g., Terraform) and scripting in Python and/or shell
  • Deep hands-on experience with CI/CD tooling and building release automation at scale
  • Strong experience with Kubernetes, containerisation, and cloud-native operations
  • Proven experience running production services: observability, on-call, incident response, and reliability engineering
  • Understanding of production security and change-management controls
  • Strong communication skills and the ability to set operational standards across a team
  • Formal SRE experience in a regulated or high-availability environment
  • Master's degree in Computer Science, Engineering, or a related technical field (or equivalent applied experience)

Preferred qualifications, capabilities, and skills

  • Experience operating ML / LLM workloads in production (MLOps, inference reliability, cost/performance management)
  • Experience within financial services technology
  • Familiarity with JPM-internal platform, cloud, and observability tooling for internal candidates
Lead reliable AI products by owning automation, operations, and incident practices that empower teams to build, deploy, and run services.

Sr Lead Infrastructure Engineer- Devops/AWS

Compensation

Not specified

City: Not specified

Country: United Kingdom

J.P. Morgan logo
Bulge Bracket Investment Banks

4 days ago

No clicks

at J.P. Morgan

ExperiencedNo visa sponsorship

**Lead Infrastructure Engineer - Senior Vice President (AWS/DevOps) in Glasgow:** Oversee CI/CD, observability, and incident management for AI/ML platform owned by top-tier banking tech team. Manage cross-functional team to enhance reliability, security, and scalability of AI products operating worldwide. Core responsibilities include evolving CI/CD, setting SLOs, automating infrastructure, implementing security controls, and mentoring junior engineers. Required: 7+ years in DevOps/SRE, Terraform, Python, strong hands-on Kubernetes and cloud experience. Must have production security and incident response experience. Master's degree preferred.

Full Job Description

Location: GLASGOW, LANARKSHIRE, United Kingdom

We're looking for a hands-on DevOps / SRE engineer ready to take their career to new heights. Join the ranks of top talent at one of the world's most influential companies.

As a Lead DevOps / SRE Engineer - Vice President at JPMorgan Chase within the International Private Bank (IPB) Technology Artificial Intelligence and Machine Learning (AIML) Team, you will own the automation, reliability, and production operations of our agentic AI and machine learning products. As the team takes end-to-end ownership of the platforms it runs, you will build the CI/CD, observability, and incident-management practices that keep those services stable, secure, and performant across international markets.

This is a Vice President-level role and an integral part of the IPB Tech AIML team, reporting to the Head of AI, IPB Tech.

Job responsibilities

  • Owns and evolves the team's CI/CD pipelines, release automation, and deployment tooling
  • Establishes reliability practices (SLOs, error budgets, runbooks) and leads production incident response and post-incident review
  • Builds and operates observability across the team's AI/ML services (metrics, logging, tracing, alerting)
  • Automates infrastructure provisioning and configuration through infrastructure-as-code
  • Implements operational security, secrets management, and access controls to firm-wide standards
  • Partners with platform, data, and AI engineers to harden services for production and reduce dependency on external functions
  • Mentors junior engineers on DevOps and reliability practices and sets standards through review
  • Champions the firm's culture of diversity, Opportunity, inclusion, and respect

Required qualifications, capabilities, and skills

  • Formal training or certification on software engineering or systems concepts and applied experience
  • Advanced proficiency with infrastructure-as-code (e.g., Terraform) and scripting in Python and/or shell
  • Deep hands-on experience with CI/CD tooling and building release automation at scale
  • Strong experience with Kubernetes, containerisation, and cloud-native operations
  • Proven experience running production services: observability, on-call, incident response, and reliability engineering
  • Understanding of production security and change-management controls
  • Strong communication skills and the ability to set operational standards across a team
  • Formal SRE experience in a regulated or high-availability environment
  • Master's degree in Computer Science, Engineering, or a related technical field (or equivalent applied experience)

Preferred qualifications, capabilities, and skills

  • Experience operating ML / LLM workloads in production (MLOps, inference reliability, cost/performance management)
  • Experience within financial services technology
  • Familiarity with JPM-internal platform, cloud, and observability tooling for internal candidates
Lead reliable AI products by owning automation, operations, and incident practices that empower teams to build, deploy, and run services.