Senior Site Reliability Engineer
Brak informacji o wynagrodzeniu
SeniorFull-time·B2B
#413303·Dodano 2 dni temu·0
Źródło: IT PerformanceTech Stack / Keywords
ArchitectureDevOpsAWSCloudKubernetes
Firma i stanowisko
We are looking for experienced candidates for the position of Senior Site Reliability Engineer. This job is dedicated to one of the largest pharmaceutical companies.
Wymagania
- 10+ years of hands-on experience in Software Engineering, DevOps, or Systems Infrastructure.
- 3+ years of advanced AWS experience, including designing, operating, and troubleshooting cloud environments.
- 3+ years of production-level Kubernetes experience, including cluster management, scaling, and provisioning.
- Strong understanding of site reliability, automation, cloud infrastructure, and production-grade systems.
- Proven ability to work with a high degree of autonomy, making sound architectural decisions and delivering production-ready solutions without day-to-day technical supervision.
- Strong experience in designing scalable and automated infrastructure.
- Excellent technical communication skills and the ability to collaborate effectively with Technical Leads and cross-functional engineering teams.
- Ability to remain effective and structured when dealing with high-pressure production incidents.
Obowiązki
- Design and drive the reliability, scalability, and performance of the multi-cloud provisioning platform across production environments.
- Design and implement end-to-end automation pipelines to eliminate manual processes and reduce technical toil.
- Define, monitor, and continuously improve key reliability metrics and SLIs/SLOs, including latency, throughput, error rates, and capacity utilization.
- Collaborate with product and engineering teams to integrate reliability and security considerations early in the software development lifecycle.
- Own and lead incident response, including detection, triage, critical incident management, Root Cause Analysis (RCA), and implementation of long-term preventive solutions.
- Identify platform bottlenecks, eliminate single points of failure, and continuously simplify complex systems.
- Take ownership of architectural decisions and establish reliability standards across the platform.
- Provide technical guidance and mentoring to other engineers within the team.
Benefity
- B2B Contract
- Remote work with occasional travel within Europe
IT Performance
29 aktywnych ofert