
Posted 10 days ago
No clicks
**Principal Site Reliability Engineer**: Leads high-scale, resilient systems using SRE principles and automation. Designs SRE-focused products, enhances cloud capabilities, and ensures reliability via chaos engineering. Key responsibilities include managing datasets and cloud infrastructure, scaling products, and mentoring SRE teams. Requires BS/MS in tech-related fields and 3-5 years of SRE experience. Expertise in Kubernetes, Terraform, Azure, Datadog, Jenkins, and Python is essential. Part of Fidelity's on-site working model transition.
- Compensation
- Not specified
- City
- Not specified
- Country
- Not specified
Currency: Not specified
Full Job Description
Job Description:
Note: Fidelity will not provide immigration sponsorship for this position.
Position Description:
Delivers services at high scale and high availability with resilience by using automation and Infrastructure Code. Builds reliability into the ecosystem by applying best practices in resiliency engineering and observability by developing resiliency tools and capabilities for observability and chaos testing of pipelines. Combines systems and software engineering techniques with site reliability engineering practices to create reliable user experiences to support workplace investing, healthcare and defined benefits organizations. Assists teams scale through production insights, operational automation, developer guidance, real-time metrics, and automation. Provides training, support, and alignment to ensures Site Reliability Engineers have the skills, tools, and opportunities to accomplish engineering reliable systems. Provides Product and Platform teams with engineering expertise that enables them to clarify and achieve their system reliability goals. Partners with our key stakeholders in defining and adopting policies, processes and practices that lead to reliable Information Technology systems and measures compliance with those policies.
Primary Responsibilities:
- Provides cloud support and enhances cloud capabilities according to Site Reliability Engineering principles -- observability, automation, and resiliency.
- Develops and enhances internal chaos framework to streamline chaos executions and reporting.
- Facilitates the adoption of chaos engineering by application teams, helping to perform chaos testing, and to analyze business-critical applications to understand the weaknesses and increase application resiliency.
- Develops and designs products created around the Site Reliability Engineering domain to improve stability and improve platform availability.
- Collaborates with business and technology teams to scale the products and automation across business units.
- Minimizes the impact of operational problems by developing strategies and tools to remediate the issues.
- Provides technical leadership on chaos testing for cloud and on-premises based applications.
- Develops scripts and applications to automate repeatable business processes.
- Advises senior management on technical strategy and tools.
- Mentors team members to build core competencies required in Site Reliability Engineering space.
Education and Experience:
Bachelors degree in Computer Science, Engineering, Information Technology, Information Systems, or a closely related field (or foreign education equivalent) and five (5) years of experience as a Principal Site Reliability Engineer (or closely related occupation) designing and automating container and Cloud-based platform products and infrastructure solutions within a production environment.
Or, alternatively, Masters degree in Computer Science, Engineering, Information Technology, Information Systems or a closely related field (or foreign education equivalent) and three (3) years of experience as a Principal Site Reliability Engineer (or closely related occupation) designing and automating container and Cloud-based platform products and infrastructure solutions within a production environment.
Skills and Knowledge:
Candidate must also possess:
- Demonstrated Expertise (DE) developing and designing products created around Site Reliability Engineering pillars to improve stability and improve platform availability for containerized workloads and on-premises services using Kubernetes; managing and interpreting datasets using query languages; and developing reports using PowerBi and Grafana.
- DE managing Cloud and on-premises systems using infrastructure-as-code tools -- Azure ARM and Terraform; and building, operating, monitoring, logging, and alerting services of distributed systems at scale and utilizing modern monitoring tools Datadog and Splunk.
- DE creating solutions that support DevOps practice for delivery and operations of services using Jenkins and Azure DevOps, Team Foundation Version Control, and Cloud Formation Template.
- DE maintaining scalability and resiliency of applications deployed on Amazon Web Services (AWS) and Azure using LAMBDA, API Gateway, Fault Injection Service (FIS), and Azure Chaos Studio; and developing scripts and applications to automate repeatable business processes and providing solutions to program needs for applications hosted on Windows and Linux using Python.
#PE1M2
#LI-DNI
Fidelitys Onsite Working Model
Fidelity is transitioning to a full-time onsite working model through a phased rollout across regions and roles. Currently, some roles and locations require 100% onsite presence, while others require less. Onsite expectations are likely to evolve as the rollout continues. This transition does not apply to fully remote roles.
Certifications:
Category:
Information TechnologyPlease be advised that Fidelitys business is governed by the provisions of the Securities Exchange Act of 1934, the Investment Advisers Act of 1940, the Investment Company Act of 1940, ERISA, numerous state laws governing securities, investment and retirement-related financial activities and the rules and regulations of numerous self-regulatory organizations, including FINRA, among others. Those laws and regulations may restrict Fidelity from hiring and/or associating with individuals with certain Criminal Histories.
Apply
All fields are required. Candidates should limit the number of roles they apply to at any given time.
Benefits that balance life and work
From our fully paid parent leave to our on-site health and wellness centers, our benefits support the belief that more balance you have, the better you can achieve your goals.
Company overview
Company overview
At Fidelity, we are passionate about making our financial expertise broadly accessible and effective in helping people live the lives they want. We are a privately held company that places a high degree of value in creating and nurturing a work environment that attracts the best talent and reflects our commitment to our associates. We are proud of our diverse and inclusive workplace where we respect and value our associates for their unique perspectives and experience.
Reasonable accommodations
Fidelity will reasonably accommodate applicants with disabilities who need adjustments to participate in the application or interview process. To initiate a request for an accommodation contact the HR Accommodation Team by sending an email to accommodations@fmr.com, or by calling 800-835-5099, prompt 2, option 3.
Equal opportunity employer
Fidelity Investments is an equal opportunity employer. We believe that the most effective way to attract, develop, and retain a diverse workforce is to build an enduring culture of inclusion and belonging.
Applicant screening
At Fidelity, we value honesty, integrity, and the safety of our associates and customers within a heavily regulated industry. Certain roles may require candidates to go through a preliminary credit check during the screening process. Candidates who are presented with a Fidelity offer will need to go through a background investigation and may be asked to provide additional documentation as requested. This investigation includes but is not limited to a criminal, civil litigations and regulatory review, employment, education, and credit review (role dependent). These investigations will account for 7 years or more of history, depending on the role. Where permitted by federal or state law, Fidelity will also conduct a pre-employment drug screen, which will review for the following substances: Amphetamines, THC (marijuana), cocaine, opiates, phencyclidine.
AI Guidelines
Learn about our guidelines for use of AI when applying for a Fidelity job




