LOG IN
SIGN UP
Canary Wharfian - Online Investment Banking & Finance Community.
Sign In
Forgot password?
Don't have an account?
or
Join Canary Wharfian
By signing up, you agree to our Terms & Conditions and Privacy Policy.
or

Principal AI/ML Engineer

ExperiencedVisa sponsorship available
Fidelity Investments logo

at Fidelity Investments

Asset Management

Posted 6 days ago

No clicks

**Principal AI/ML Engineer** Design and develop advanced, secure AI/ML systems for large-scale enterprise platforms. Key responsibilities include: - Architecting AI/ML infrastructure for model training, inference, observability, and continuous monitoring across distributed cloud environments. - Developing secure and trustworthy AI frameworks, including robust anomaly detection models and model governance mechanisms. - Building and optimizing agentic AI workflows to automate data pipelines, model lifecycle operations, and system self-diagnostics. - Prototyping cutting-edge AI methodologies, such as federated learning, adversarial robustness, and multi-agent safety mechanisms. - Integrating AI safety and model assurance into enterprise architectures, collaborating with cross-functional teams. Requirements: - Bachelor's or Master's degree in Computer Science, Engineering, or a related field. - Five years (or three years with a Master's degree) of experience in a similar role, developing ML platform applications for cloud infrastructures. - Demonstrated expertise (DE) in delivering scalable, secure distributed applications, architecting high-performance big data applications, automating end-to-end application deployment, implementing secure AI, and evaluating ML model inference performance. - Proficient in Python, Java, Go, AWS, Azure, GCP, SageMaker, Kubernetes, and relevant tools.

Compensation
Not specified

Currency: Not specified

City
Not specified
Country
Not specified

Full Job Description

Job Description:

Note: Fidelity will not provide immigration sponsorship for this position.

Position Description:

Designs and develops advanced Machine Learning (ML) and trustworthy AI systems that support large-scale, enterprise-wide platforms. Focuses on secure model development, AI safety, multi-agent orchestration, and cloud-scale ML infrastructure, enabling reliable, transparent, and high-assurance deployment of AI across critical business functions. Builds robust pipelines for model training, inference, and continuous monitoring, ensuring compliance and resilience under real-world conditions. Designs model lifecycle operations (MLOps) infrastructure including SageMaker and Kubernetes for reproducible, scalable, and secure ML deployment. Responsible for research and prototyping of cutting-edge AI methodologies, including federated learning, adversarial robustness, and multi-agent safety mechanisms.

Primary Responsibilities:

  • Architects and implements AI/ML systems that support model training, inference, observability, and continuous monitoring across distributed cloud environments.

  • Develops secure and trustworthy AI frameworks, including adversarial robustness pipelines, anomaly detection models, and model governance mechanisms that ensure compliance, transparency, and risk mitigation.

  • Builds and optimizes agentic AI workflows that automate data pipelines, model lifecycle operations, and system self-diagnostics.

  • Supports research and prototyping of advanced AI/ML methodologies, including hybrid neural architectures, federated learning, adversarial learning, and multi-agent AI safety mechanisms.

  • Conducts performance, reliability, and robustness evaluations of AI systems under real-world constraints, high-throughput workloads, and adversarial conditions.

  • Collaborate with cybersecurity, cloud engineering, and data science teams to integrate AI safety and model assurance into enterprise architectures.

  • Collaborates with cross-functional teams to integrate AI safety and governance into enterprise architectures while mentoring engineering teams on advanced ML algorithms and secure development practices.

  • Delivers technical guidance and mentorship to cross-functional engineering teams on advanced ML algorithms, infrastructure patterns, and secure development practices.

Education and Experience:

Bachelors degree in Computer Science, Engineering, Information Technology, Information Systems, or a closely related field (or foreign education equivalent) and five (5) years of experience as a Principal AI/ML Engineer (or closely related occupation) developing ML platform applications for Cloud infrastructures (Amazon Web Services (AWS), Azure, Google, and IBM) using agile methodologies.

Or, alternatively, Masters degree in Computer Science, Engineering, Information Technology, Information Systems, or a closely related field (or foreign education equivalent) and three (3) years of experience as a Principal AI/ML Engineer (or closely related occupation) developing ML platform applications for Cloud infrastructures (Amazon Web Services (AWS), Azure, Google, and IBM) using agile methodologies.

Skills and Knowledge:

Candidate must also possess:

  • Demonstrated Expertise (DE) delivering scalable, secure, and distributed applications with robust Identity and Access Management (IAM) and encryption (Knowledge Management System (KMS)), by architecting and deploying enterprise-scale AI/ML systems and auto-ML infrastructure through Infrastructure as code (IAC) and using Cloud-native platforms (Terraform, AWS, Azure, or GCP) and container orchestration frameworks (Kubeflow).

  • DE architecting and engineering high-performance big data applications (AWS Glue, EMR, Kinesis, Athena, and Dynamo DB) and autonomous multi-agent systems; designing batch processing jobs and Extract, Transform, Load (ETL) pipelines to support predictive analytics using Hadoop, MongoDB, AWS, and PostgreSQL; developing agentic workflows using frameworks including Strands, CrewAI, LangGraph, and OpenAI Swarm with protocols -- Model Context Protocol (MCP) Server and Accelerated Graphics Port (AGP); and driving context aware predictive analytics and optimization models by implementing multi threaded, asynchronous solutions in Python and Java, supported by short term and long term memory management architectures for Retrieval Augmented Generation (RAG) pipelines using vector databases, OpenSearch, and high performance caching solutions (Redis or Memcached).

  • DE automating end-to-end application deployment and MLOps via workflow orchestration tools (Airflow and AWS Step Functions) and CI/CD pipelines (Jenkins and Git); performing model metadata processing to deliver lineage tracking, version control, auditability, reproducibility, and real time analysis of model performance, parameters, and deployment history across distributed ML workflows, using AWS DynamoDB, AWS RDS, MLflow, and AWS Athena; enabling reproducible builds, integrity checks, and scalable CI/CD by securely handling, versioning, and distributing container images, ML models, and software dependencies using JFrog Artifactory; and accelerating model development, monitoring, drift detection, and interpretability within Agile environments by engineering bridge solutions, using Go and FastAPI.

  • DE implementing secure AI through adversarial robustness, federated learning, and governance across distributed ML systems using AI Generative Adversarial Network (AIGAN), Google Federated Learning Framework, SageMaker Clarify, and MLflow; enhancing low latency, high throughput model serving and maximizing central processing unit (CPU) or graphics processing unit (GPU) utilization through deployment on accelerated inference servers, using Deep Java Library (DJL), Triton, and Flask; and evaluating ML model inference performance using statistical analysis, monitoring tools (CloudWatch, Datadog, and Splunk), and dashboards including Streamlit and Gradio.

[Experience and/or expertise may be gained during doctoral program.]

#PE1M2

#LI-DNI

Fidelitys Onsite Working Model
Fidelity is transitioning to a full-time onsite working model through a phased rollout across regions and roles. Currently, some roles and locations require 100% onsite presence, while others require less. Onsite expectations are likely to evolve as the rollout continues. This transition does not apply to fully remote roles.

Certifications:

Category:

Information Technology

Please be advised that Fidelitys business is governed by the provisions of the Securities Exchange Act of 1934, the Investment Advisers Act of 1940, the Investment Company Act of 1940, ERISA, numerous state laws governing securities, investment and retirement-related financial activities and the rules and regulations of numerous self-regulatory organizations, including FINRA, among others. Those laws and regulations may restrict Fidelity from hiring and/or associating with individuals with certain Criminal Histories.

Apply

All fields are required. Candidates should limit the number of roles they apply to at any given time.

Benefits that balance life and work

From our fully paid parent leave to our on-site health and wellness centers, our benefits support the belief that more balance you have, the better you can achieve your goals.

Benefits

Company overview

Company overview 

At Fidelity, we are passionate about making our financial expertise broadly accessible and effective in helping people live the lives they want. We are a privately held company that places a high degree of value in creating and nurturing a work environment that attracts the best talent and reflects our commitment to our associates. We are proud of our diverse and inclusive workplace where we respect and value our associates for their unique perspectives and experience. 

Reasonable accommodations

Fidelity will reasonably accommodate applicants with disabilities who need adjustments to participate in the application or interview process. To initiate a request for an accommodation contact the HR Accommodation Team by sending an email to accommodations@fmr.com, or by calling 800-835-5099, prompt 2, option 3.

Equal opportunity employer

Fidelity Investments is an equal opportunity employer. We believe that the most effective way to attract, develop, and retain a diverse workforce is to build an enduring culture of inclusion and belonging.

Applicant screening

At Fidelity, we value honesty, integrity, and the safety of our associates and customers within a heavily regulated industry. Certain roles may require candidates to go through a preliminary credit check during the screening process. Candidates who are presented with a Fidelity offer will need to go through a background investigation and may be asked to provide additional documentation as requested. This investigation includes but is not limited to a criminal, civil litigations and regulatory review, employment, education, and credit review (role dependent). These investigations will account for 7 years or more of history, depending on the role. Where permitted by federal or state law, Fidelity will also conduct a pre-employment drug screen, which will review for the following substances: Amphetamines, THC (marijuana), cocaine, opiates, phencyclidine.

AI Guidelines

Learn about our guidelines for use of AI when applying for a Fidelity job

Return to job search

Principal AI/ML Engineer

Compensation

Not specified

City: Not specified

Country: Not specified

Fidelity Investments logo
Asset Management

6 days ago

No clicks

at Fidelity Investments

ExperiencedVisa sponsorship available

**Principal AI/ML Engineer** Design and develop advanced, secure AI/ML systems for large-scale enterprise platforms. Key responsibilities include: - Architecting AI/ML infrastructure for model training, inference, observability, and continuous monitoring across distributed cloud environments. - Developing secure and trustworthy AI frameworks, including robust anomaly detection models and model governance mechanisms. - Building and optimizing agentic AI workflows to automate data pipelines, model lifecycle operations, and system self-diagnostics. - Prototyping cutting-edge AI methodologies, such as federated learning, adversarial robustness, and multi-agent safety mechanisms. - Integrating AI safety and model assurance into enterprise architectures, collaborating with cross-functional teams. Requirements: - Bachelor's or Master's degree in Computer Science, Engineering, or a related field. - Five years (or three years with a Master's degree) of experience in a similar role, developing ML platform applications for cloud infrastructures. - Demonstrated expertise (DE) in delivering scalable, secure distributed applications, architecting high-performance big data applications, automating end-to-end application deployment, implementing secure AI, and evaluating ML model inference performance. - Proficient in Python, Java, Go, AWS, Azure, GCP, SageMaker, Kubernetes, and relevant tools.

Full Job Description

Job Description:

Note: Fidelity will not provide immigration sponsorship for this position.

Position Description:

Designs and develops advanced Machine Learning (ML) and trustworthy AI systems that support large-scale, enterprise-wide platforms. Focuses on secure model development, AI safety, multi-agent orchestration, and cloud-scale ML infrastructure, enabling reliable, transparent, and high-assurance deployment of AI across critical business functions. Builds robust pipelines for model training, inference, and continuous monitoring, ensuring compliance and resilience under real-world conditions. Designs model lifecycle operations (MLOps) infrastructure including SageMaker and Kubernetes for reproducible, scalable, and secure ML deployment. Responsible for research and prototyping of cutting-edge AI methodologies, including federated learning, adversarial robustness, and multi-agent safety mechanisms.

Primary Responsibilities:

  • Architects and implements AI/ML systems that support model training, inference, observability, and continuous monitoring across distributed cloud environments.

  • Develops secure and trustworthy AI frameworks, including adversarial robustness pipelines, anomaly detection models, and model governance mechanisms that ensure compliance, transparency, and risk mitigation.

  • Builds and optimizes agentic AI workflows that automate data pipelines, model lifecycle operations, and system self-diagnostics.

  • Supports research and prototyping of advanced AI/ML methodologies, including hybrid neural architectures, federated learning, adversarial learning, and multi-agent AI safety mechanisms.

  • Conducts performance, reliability, and robustness evaluations of AI systems under real-world constraints, high-throughput workloads, and adversarial conditions.

  • Collaborate with cybersecurity, cloud engineering, and data science teams to integrate AI safety and model assurance into enterprise architectures.

  • Collaborates with cross-functional teams to integrate AI safety and governance into enterprise architectures while mentoring engineering teams on advanced ML algorithms and secure development practices.

  • Delivers technical guidance and mentorship to cross-functional engineering teams on advanced ML algorithms, infrastructure patterns, and secure development practices.

Education and Experience:

Bachelors degree in Computer Science, Engineering, Information Technology, Information Systems, or a closely related field (or foreign education equivalent) and five (5) years of experience as a Principal AI/ML Engineer (or closely related occupation) developing ML platform applications for Cloud infrastructures (Amazon Web Services (AWS), Azure, Google, and IBM) using agile methodologies.

Or, alternatively, Masters degree in Computer Science, Engineering, Information Technology, Information Systems, or a closely related field (or foreign education equivalent) and three (3) years of experience as a Principal AI/ML Engineer (or closely related occupation) developing ML platform applications for Cloud infrastructures (Amazon Web Services (AWS), Azure, Google, and IBM) using agile methodologies.

Skills and Knowledge:

Candidate must also possess:

  • Demonstrated Expertise (DE) delivering scalable, secure, and distributed applications with robust Identity and Access Management (IAM) and encryption (Knowledge Management System (KMS)), by architecting and deploying enterprise-scale AI/ML systems and auto-ML infrastructure through Infrastructure as code (IAC) and using Cloud-native platforms (Terraform, AWS, Azure, or GCP) and container orchestration frameworks (Kubeflow).

  • DE architecting and engineering high-performance big data applications (AWS Glue, EMR, Kinesis, Athena, and Dynamo DB) and autonomous multi-agent systems; designing batch processing jobs and Extract, Transform, Load (ETL) pipelines to support predictive analytics using Hadoop, MongoDB, AWS, and PostgreSQL; developing agentic workflows using frameworks including Strands, CrewAI, LangGraph, and OpenAI Swarm with protocols -- Model Context Protocol (MCP) Server and Accelerated Graphics Port (AGP); and driving context aware predictive analytics and optimization models by implementing multi threaded, asynchronous solutions in Python and Java, supported by short term and long term memory management architectures for Retrieval Augmented Generation (RAG) pipelines using vector databases, OpenSearch, and high performance caching solutions (Redis or Memcached).

  • DE automating end-to-end application deployment and MLOps via workflow orchestration tools (Airflow and AWS Step Functions) and CI/CD pipelines (Jenkins and Git); performing model metadata processing to deliver lineage tracking, version control, auditability, reproducibility, and real time analysis of model performance, parameters, and deployment history across distributed ML workflows, using AWS DynamoDB, AWS RDS, MLflow, and AWS Athena; enabling reproducible builds, integrity checks, and scalable CI/CD by securely handling, versioning, and distributing container images, ML models, and software dependencies using JFrog Artifactory; and accelerating model development, monitoring, drift detection, and interpretability within Agile environments by engineering bridge solutions, using Go and FastAPI.

  • DE implementing secure AI through adversarial robustness, federated learning, and governance across distributed ML systems using AI Generative Adversarial Network (AIGAN), Google Federated Learning Framework, SageMaker Clarify, and MLflow; enhancing low latency, high throughput model serving and maximizing central processing unit (CPU) or graphics processing unit (GPU) utilization through deployment on accelerated inference servers, using Deep Java Library (DJL), Triton, and Flask; and evaluating ML model inference performance using statistical analysis, monitoring tools (CloudWatch, Datadog, and Splunk), and dashboards including Streamlit and Gradio.

[Experience and/or expertise may be gained during doctoral program.]

#PE1M2

#LI-DNI

Fidelitys Onsite Working Model
Fidelity is transitioning to a full-time onsite working model through a phased rollout across regions and roles. Currently, some roles and locations require 100% onsite presence, while others require less. Onsite expectations are likely to evolve as the rollout continues. This transition does not apply to fully remote roles.

Certifications:

Category:

Information Technology

Please be advised that Fidelitys business is governed by the provisions of the Securities Exchange Act of 1934, the Investment Advisers Act of 1940, the Investment Company Act of 1940, ERISA, numerous state laws governing securities, investment and retirement-related financial activities and the rules and regulations of numerous self-regulatory organizations, including FINRA, among others. Those laws and regulations may restrict Fidelity from hiring and/or associating with individuals with certain Criminal Histories.

Apply

All fields are required. Candidates should limit the number of roles they apply to at any given time.

Benefits that balance life and work

From our fully paid parent leave to our on-site health and wellness centers, our benefits support the belief that more balance you have, the better you can achieve your goals.

Benefits

Company overview

Company overview 

At Fidelity, we are passionate about making our financial expertise broadly accessible and effective in helping people live the lives they want. We are a privately held company that places a high degree of value in creating and nurturing a work environment that attracts the best talent and reflects our commitment to our associates. We are proud of our diverse and inclusive workplace where we respect and value our associates for their unique perspectives and experience. 

Reasonable accommodations

Fidelity will reasonably accommodate applicants with disabilities who need adjustments to participate in the application or interview process. To initiate a request for an accommodation contact the HR Accommodation Team by sending an email to accommodations@fmr.com, or by calling 800-835-5099, prompt 2, option 3.

Equal opportunity employer

Fidelity Investments is an equal opportunity employer. We believe that the most effective way to attract, develop, and retain a diverse workforce is to build an enduring culture of inclusion and belonging.

Applicant screening

At Fidelity, we value honesty, integrity, and the safety of our associates and customers within a heavily regulated industry. Certain roles may require candidates to go through a preliminary credit check during the screening process. Candidates who are presented with a Fidelity offer will need to go through a background investigation and may be asked to provide additional documentation as requested. This investigation includes but is not limited to a criminal, civil litigations and regulatory review, employment, education, and credit review (role dependent). These investigations will account for 7 years or more of history, depending on the role. Where permitted by federal or state law, Fidelity will also conduct a pre-employment drug screen, which will review for the following substances: Amphetamines, THC (marijuana), cocaine, opiates, phencyclidine.

AI Guidelines

Learn about our guidelines for use of AI when applying for a Fidelity job

Return to job search