LOG IN
SIGN UP
Canary Wharfian - Online Investment Banking & Finance Community.
Sign In
Forgot password?
Don't have an account?
or
Join Canary Wharfian
By signing up, you agree to our Terms & Conditions and Privacy Policy.
or

Lead Systems Operations Engineer - Windows / Unix

ExperiencedNo visa sponsorship

Posted 15 days ago

No clicks

**Lead Systems Operations Engineer - Windows / Unix** Drive reliability, resiliency, and operational excellence across critical platforms. Lead complex projects, consult on systems, and make technical decisions. Proven senior experience required (5+ years) in Systems Engineering, Site Reliability Engineering, and supporting large-scale Java/.NET applications. Expertise in Windows/Unix, CI/CD tools, observability, and automation essential. Both on-prem and cloud environments (AWS, Azure, GCP) experience preferred. Occasional evening/weekend On-Call and min. 3 days/week on-site expected. Subject matter expert for environment reliability, promotes best practices, and supports operational governance.

Compensation
$119,000 – $187,000 USD

Currency: $ (USD)

City
Not specified
Country
United States

Full Job Description

About this role:

Wells Fargo is seeking a Lead Systems Operations Engineer within the Branch Systems and Transformation Technology team. This role is aligned to modern Site Reliability Engineering (SRE) practices and is responsible for driving reliability, resiliency, observability, and operational excellence across critical platform and application services. The role is intended for senior engineers with deep expertise in one core platform domain, applying that expertise to proactively improve platform stability, scalability, and availability.


In this role, you will:

  • Lead complex, broad impact initiatives including provision of high level systems consultation for the technology teams
  • Work as key participant in large scale planning of computer systems and network infrastructure for Systems Operations functional area
  • Review and analyze complex technical challenges, as well as escalated support issues related to core business solutions that require in depth evaluation of multiple factors, such as alternatives, enhancements, periodic systems reviews, or improvements to existing systems
  • Make decisions on technical changes and enhancements
  • Consult with engineering team on change design requiring solid understanding of technical process controls or standards that influence and drive new initiatives
  • Collaborate and consult with technical peers, colleagues, and mid to more experienced level managers to resolve systems support issues and achieve goals


Required Qualifications:

  • 5+ years of Systems Engineering, Technology Architecture experience, or equivalent demonstrated through one or a combination of the following: work experience, training, military experience, education
  • 5+ years of Site Reliability Engineering experience, or equivalent demonstrated through one or a combination of the following:  work experience, training, education


Desired Qualifications:

  • Demonstrated experience leading Production Application Support at scale, including ITIL-aligned incident, problem, and change management
  • Experience supporting AI/ML-enabled production systems, including reliability, scalability, and operational risk management for LLM- or agent-based workflows (e.g., model deployments, inference services, retrieval pipelines).
  • Understanding of AI operational concerns such as model observability, prompt/configuration change management, data drift, failure modes, and integration of AI signals into incident response and SRE practices.
  • Thorough understanding of application environment implementations, including on prem, client server, on prem cloud, hybrid cloud and public cloud
  • Experience using and configuring Continuous Integration Continuous Deploy (CICD) tools including Jenkins, Aritifactory, Udeploy as well as Terraform
  • Proven experience reducing operational toil through automation (Ansible or equivalent), including runbook automation and selfhealing patterns
  • Experience implementing application Observability through tools such as AppDynamics, Splunk, Elastic, BigPanda (AIOPS), Grafana, Microsoft Application Insights
  • Experience designing and operating resilient systems, including traffic management (F5, AVI), fault tolerance, and data replication strategies
  • Experience supporting largescale Java and .NET applications; strong RDBMS expertise (Oracle, MSSQL); NoSQL experience (MongoDB) a plus


Job Expectations:

  • Periodic availability for evening and weekend On-Call
  • Expectation of at least 3 days per week in office
  • Serves as a subject matter expert for environment reliability and operational readiness.
  • Provides technical guidance, promotes reliability engineering best practices, supports operational governance activities, and collaborates with engineering and testing teams to improve environment stability and delivery effectiveness.

Pay Range
 

Reflected is the base pay range offered for this position. Pay may vary depending on factors including but not limited to demonstrated examples of prior performance, skills, experience, or work location. Employees may also be eligible for incentive opportunities.

$119,000.00 - $187,000.00

Benefits

Wells Fargo provides eligible employees with a comprehensive set of benefits, many of which are listed below. Visit Benefits - Wells Fargo Jobs for an overview of the following benefit plans and programs offered to employees.

  • Health benefits
  • 401(k) Plan
  • Paid time off
  • Disability benefits
  • Life insurance, critical illness insurance, and accident insurance
  • Parental leave
  • Critical caregiving leave
  • Discounts and savings
  • Commuter benefits
  • Tuition reimbursement
  • Scholarships for dependent children
  • Adoption reimbursement

Posting End Date:

25 Aug 2026

*Job posting may come down early due to volume of applicants.

We Value Equal Opportunity

Wells Fargo is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, status as a protected veteran, or any other legally protected characteristic.

Employees support our focus on building strong customer relationships balanced with a strong risk mitigating and compliance-driven culture which firmly establishes those disciplines as critical to the success of our customers and company. They are accountable for execution of all applicable risk programs (Credit, Market, Financial Crimes, Operational, Regulatory Compliance), which includes effectively following and adhering to applicable Wells Fargo policies and procedures, appropriately fulfilling risk and compliance obligations, timely and effective escalation and remediation of issues, and making sound risk decisions. There is emphasis on proactive monitoring, governance, risk identification and escalation, as well as making sound risk decisions commensurate with the business units risk appetite and all risk and compliance program requirements.

Applicants with Disabilities

To request a medical accommodation during the application or interview process, visit Disability Inclusion at Wells Fargo.

Drug and Alcohol Policy

 

Wells Fargo maintains a drug free workplace.  Please see our Drug and Alcohol Policy to learn more.

Wells Fargo Recruitment and Hiring Requirements:

a. Third-Party recordings are prohibited unless authorized by Wells Fargo.

b. Wells Fargo requires you to directly represent your own experiences during the recruiting and hiring process.


Lead Systems Operations Engineer - Windows / Unix

Compensation

$119,000 – $187,000 USD

City: Not specified

Country: United States

Wells Fargo Corporate & Investment Banking  logo
Investment Banking

15 days ago

No clicks

at Wells Fargo Corporate & Investment Banking

ExperiencedNo visa sponsorship

**Lead Systems Operations Engineer - Windows / Unix** Drive reliability, resiliency, and operational excellence across critical platforms. Lead complex projects, consult on systems, and make technical decisions. Proven senior experience required (5+ years) in Systems Engineering, Site Reliability Engineering, and supporting large-scale Java/.NET applications. Expertise in Windows/Unix, CI/CD tools, observability, and automation essential. Both on-prem and cloud environments (AWS, Azure, GCP) experience preferred. Occasional evening/weekend On-Call and min. 3 days/week on-site expected. Subject matter expert for environment reliability, promotes best practices, and supports operational governance.

Full Job Description

About this role:

Wells Fargo is seeking a Lead Systems Operations Engineer within the Branch Systems and Transformation Technology team. This role is aligned to modern Site Reliability Engineering (SRE) practices and is responsible for driving reliability, resiliency, observability, and operational excellence across critical platform and application services. The role is intended for senior engineers with deep expertise in one core platform domain, applying that expertise to proactively improve platform stability, scalability, and availability.


In this role, you will:

  • Lead complex, broad impact initiatives including provision of high level systems consultation for the technology teams
  • Work as key participant in large scale planning of computer systems and network infrastructure for Systems Operations functional area
  • Review and analyze complex technical challenges, as well as escalated support issues related to core business solutions that require in depth evaluation of multiple factors, such as alternatives, enhancements, periodic systems reviews, or improvements to existing systems
  • Make decisions on technical changes and enhancements
  • Consult with engineering team on change design requiring solid understanding of technical process controls or standards that influence and drive new initiatives
  • Collaborate and consult with technical peers, colleagues, and mid to more experienced level managers to resolve systems support issues and achieve goals


Required Qualifications:

  • 5+ years of Systems Engineering, Technology Architecture experience, or equivalent demonstrated through one or a combination of the following: work experience, training, military experience, education
  • 5+ years of Site Reliability Engineering experience, or equivalent demonstrated through one or a combination of the following:  work experience, training, education


Desired Qualifications:

  • Demonstrated experience leading Production Application Support at scale, including ITIL-aligned incident, problem, and change management
  • Experience supporting AI/ML-enabled production systems, including reliability, scalability, and operational risk management for LLM- or agent-based workflows (e.g., model deployments, inference services, retrieval pipelines).
  • Understanding of AI operational concerns such as model observability, prompt/configuration change management, data drift, failure modes, and integration of AI signals into incident response and SRE practices.
  • Thorough understanding of application environment implementations, including on prem, client server, on prem cloud, hybrid cloud and public cloud
  • Experience using and configuring Continuous Integration Continuous Deploy (CICD) tools including Jenkins, Aritifactory, Udeploy as well as Terraform
  • Proven experience reducing operational toil through automation (Ansible or equivalent), including runbook automation and selfhealing patterns
  • Experience implementing application Observability through tools such as AppDynamics, Splunk, Elastic, BigPanda (AIOPS), Grafana, Microsoft Application Insights
  • Experience designing and operating resilient systems, including traffic management (F5, AVI), fault tolerance, and data replication strategies
  • Experience supporting largescale Java and .NET applications; strong RDBMS expertise (Oracle, MSSQL); NoSQL experience (MongoDB) a plus


Job Expectations:

  • Periodic availability for evening and weekend On-Call
  • Expectation of at least 3 days per week in office
  • Serves as a subject matter expert for environment reliability and operational readiness.
  • Provides technical guidance, promotes reliability engineering best practices, supports operational governance activities, and collaborates with engineering and testing teams to improve environment stability and delivery effectiveness.

Pay Range
 

Reflected is the base pay range offered for this position. Pay may vary depending on factors including but not limited to demonstrated examples of prior performance, skills, experience, or work location. Employees may also be eligible for incentive opportunities.

$119,000.00 - $187,000.00

Benefits

Wells Fargo provides eligible employees with a comprehensive set of benefits, many of which are listed below. Visit Benefits - Wells Fargo Jobs for an overview of the following benefit plans and programs offered to employees.

  • Health benefits
  • 401(k) Plan
  • Paid time off
  • Disability benefits
  • Life insurance, critical illness insurance, and accident insurance
  • Parental leave
  • Critical caregiving leave
  • Discounts and savings
  • Commuter benefits
  • Tuition reimbursement
  • Scholarships for dependent children
  • Adoption reimbursement

Posting End Date:

25 Aug 2026

*Job posting may come down early due to volume of applicants.

We Value Equal Opportunity

Wells Fargo is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, status as a protected veteran, or any other legally protected characteristic.

Employees support our focus on building strong customer relationships balanced with a strong risk mitigating and compliance-driven culture which firmly establishes those disciplines as critical to the success of our customers and company. They are accountable for execution of all applicable risk programs (Credit, Market, Financial Crimes, Operational, Regulatory Compliance), which includes effectively following and adhering to applicable Wells Fargo policies and procedures, appropriately fulfilling risk and compliance obligations, timely and effective escalation and remediation of issues, and making sound risk decisions. There is emphasis on proactive monitoring, governance, risk identification and escalation, as well as making sound risk decisions commensurate with the business units risk appetite and all risk and compliance program requirements.

Applicants with Disabilities

To request a medical accommodation during the application or interview process, visit Disability Inclusion at Wells Fargo.

Drug and Alcohol Policy

 

Wells Fargo maintains a drug free workplace.  Please see our Drug and Alcohol Policy to learn more.

Wells Fargo Recruitment and Hiring Requirements:

a. Third-Party recordings are prohibited unless authorized by Wells Fargo.

b. Wells Fargo requires you to directly represent your own experiences during the recruiting and hiring process.