LOG IN
SIGN UP
Canary Wharfian - Online Investment Banking & Finance Community.
Sign In
or continue with e-mail and password
Forgot password?
Don't have an account?
Join Canary Wharfian
or continue with e-mail and password
By signing up, you agree to our Terms & Conditions and Privacy Policy.

SRE Software Engineer

ExperiencedNo visa sponsorship
Capgemini logo

at Capgemini

Consultancies

Posted 12 days ago

No clicks

**SRE Software Engineer**: M 많았다Years of experience in SRE, DevOps, or similar roles needed for this Linux-admin focused position. Manage and support Kubernetes, and work with AWS, Azure, or GCP. Proficient in monitoring tools like Prometheus and Grafana, Terraform, and scripting with Bash, Python, or Go. Troubleshoot issues, conduct Root Cause Analysis, and drive preventive improvements. Maintain high availability, optimize cloud infrastructure, and automate processes. Collaborate with teams to enhance CI/CD pipelines. Remote work available.

Compensation
Not specified

Currency: Not specified

City
Not specified
Country
Not specified

Full Job Description

Job Description

Your Profile

  • 4+ years of experience in Site Reliability Engineering, DevOps, Cloud Operations, or Infrastructure Engineering.
  • Strong hands-on experience with Linux administration, troubleshooting, and production support.
  • Experience managing and supporting Kubernetes and containerized workloads (Docker/OpenShift is a plus).
  • Solid knowledge of AWS, Azure, or GCP cloud environments.
  • Experience with monitoring and observability tools such as Prometheus, Grafana, Datadog, Splunk, ELK, or CloudWatch.
  • Experience with Infrastructure as Code (Terraform preferred) and CI/CD pipelines.
  • Ability to troubleshoot complex production issues, perform Root Cause Analysis (RCA), and drive preventive improvements.
  • Working knowledge of automation and scripting using Bash, Python, or Go.
  • Intermediate to advanced English (B2+).

Responsibilities

  • Support and maintain business-critical production environments, ensuring high availability and system reliability.
  • Monitor infrastructure and applications, proactively identifying and resolving issues before they impact users.
  • Participate in incident response activities, troubleshooting production outages and coordinating recovery efforts.
  • Perform RCA and contribute to postmortems, corrective actions, and continuous improvement initiatives.
  • Manage and optimize Kubernetes clusters and cloud infrastructure.
  • Develop and maintain monitoring dashboards, alerts, and observability solutions.
  • Automate operational processes and infrastructure deployments using IaC and scripting.
  • Collaborate with engineering and product teams to improve scalability, performance, and operational excellence.
  • Support and enhance CI/CD pipelines to ensure reliable and efficient software delivery

#LI-NO1

#LI-Remote

SRE Software Engineer

Compensation

Not specified

City: Not specified

Country: Not specified

Capgemini logo
Consultancies

12 days ago

No clicks

at Capgemini

ExperiencedNo visa sponsorship

**SRE Software Engineer**: M 많았다Years of experience in SRE, DevOps, or similar roles needed for this Linux-admin focused position. Manage and support Kubernetes, and work with AWS, Azure, or GCP. Proficient in monitoring tools like Prometheus and Grafana, Terraform, and scripting with Bash, Python, or Go. Troubleshoot issues, conduct Root Cause Analysis, and drive preventive improvements. Maintain high availability, optimize cloud infrastructure, and automate processes. Collaborate with teams to enhance CI/CD pipelines. Remote work available.

Full Job Description

Job Description

Your Profile

  • 4+ years of experience in Site Reliability Engineering, DevOps, Cloud Operations, or Infrastructure Engineering.
  • Strong hands-on experience with Linux administration, troubleshooting, and production support.
  • Experience managing and supporting Kubernetes and containerized workloads (Docker/OpenShift is a plus).
  • Solid knowledge of AWS, Azure, or GCP cloud environments.
  • Experience with monitoring and observability tools such as Prometheus, Grafana, Datadog, Splunk, ELK, or CloudWatch.
  • Experience with Infrastructure as Code (Terraform preferred) and CI/CD pipelines.
  • Ability to troubleshoot complex production issues, perform Root Cause Analysis (RCA), and drive preventive improvements.
  • Working knowledge of automation and scripting using Bash, Python, or Go.
  • Intermediate to advanced English (B2+).

Responsibilities

  • Support and maintain business-critical production environments, ensuring high availability and system reliability.
  • Monitor infrastructure and applications, proactively identifying and resolving issues before they impact users.
  • Participate in incident response activities, troubleshooting production outages and coordinating recovery efforts.
  • Perform RCA and contribute to postmortems, corrective actions, and continuous improvement initiatives.
  • Manage and optimize Kubernetes clusters and cloud infrastructure.
  • Develop and maintain monitoring dashboards, alerts, and observability solutions.
  • Automate operational processes and infrastructure deployments using IaC and scripting.
  • Collaborate with engineering and product teams to improve scalability, performance, and operational excellence.
  • Support and enhance CI/CD pipelines to ensure reliable and efficient software delivery

#LI-NO1

#LI-Remote