Senior Site Reliability Engineer

Brak informacji o wynagrodzeniu
SeniorFull-time·B2B
#413303·Dodano 2 dni temu·0
Źródło: IT Performance
Aplikuj teraz

Tech Stack / Keywords

ArchitectureDevOpsAWSCloudKubernetes

Firma i stanowisko

We are looking for experienced candidates for the position of Senior Site Reliability Engineer. This job is dedicated to one of the largest pharmaceutical companies.

Wymagania

  • 10+ years of hands-on experience in Software Engineering, DevOps, or Systems Infrastructure.
  • 3+ years of advanced AWS experience, including designing, operating, and troubleshooting cloud environments.
  • 3+ years of production-level Kubernetes experience, including cluster management, scaling, and provisioning.
  • Strong understanding of site reliability, automation, cloud infrastructure, and production-grade systems.
  • Proven ability to work with a high degree of autonomy, making sound architectural decisions and delivering production-ready solutions without day-to-day technical supervision.
  • Strong experience in designing scalable and automated infrastructure.
  • Excellent technical communication skills and the ability to collaborate effectively with Technical Leads and cross-functional engineering teams.
  • Ability to remain effective and structured when dealing with high-pressure production incidents.

Obowiązki

  • Design and drive the reliability, scalability, and performance of the multi-cloud provisioning platform across production environments.
  • Design and implement end-to-end automation pipelines to eliminate manual processes and reduce technical toil.
  • Define, monitor, and continuously improve key reliability metrics and SLIs/SLOs, including latency, throughput, error rates, and capacity utilization.
  • Collaborate with product and engineering teams to integrate reliability and security considerations early in the software development lifecycle.
  • Own and lead incident response, including detection, triage, critical incident management, Root Cause Analysis (RCA), and implementation of long-term preventive solutions.
  • Identify platform bottlenecks, eliminate single points of failure, and continuously simplify complex systems.
  • Take ownership of architectural decisions and establish reliability standards across the platform.
  • Provide technical guidance and mentoring to other engineers within the team.

Benefity

  • B2B Contract
  • Remote work with occasional travel within Europe
IT Performance

IT Performance

29 aktywnych ofert

Zobacz wszystkie oferty
Aplikuj teraz