Site Reliability Engineering Manager

23.7k - 32k PLN/ mies.UoP
SeniorFull-time·Umowa o pracę
#435039·Dodano miesiąc temu·1
Źródło: CAE
Aplikuj teraz

Tech Stack / Keywords

PostgreSQLOracleSQL ServerAWSKubernetes

Firma i stanowisko

CAE is a company focused on innovation in simulation, training, and mission readiness to support critical operations worldwide, with nearly 80 years of experience.

Wymagania

  • 7+ years of experience in Site Reliability Engineering, Database Engineering or related production engineering roles
  • 2+ years of leadership experience managing engineers or technical teams
  • Strong expertise in relational database platforms such as PostgreSQL, Oracle, or SQL Server
  • Experience operating reliable services in cloud or hybrid environments, including AWS-managed or self-managed platforms
  • Deep analytical, operational, and performance-focused mindset
  • Ability to balance service reliability, engineering velocity, and cost efficiency
  • Strong background in automation, observability, incident response, performance tuning, and high availability/disaster recovery practices
  • Knowledge of Kubernetes, containerized workloads, and platform engineering practices
  • Experience defining SLOs, SLIs, alerting strategies, and operational readiness standards
  • Strong communication and cross-functional collaboration skills across engineering and business stakeholders

Obowiązki

  • Lead and develop a team of reliability engineers across database and platform domains
  • Define the reliability strategy, operating model, and roadmap for critical production services
  • Build a culture of automation, ownership, resilience, and continuous improvement
  • Own the availability, performance, scalability, and resilience of production databases and dependent platform services
  • Define and enforce SLIs, SLOs, and error budgets for critical services and database platforms
  • Lead incident management, postmortems, and remediation efforts to reduce recurrence and improve recovery
  • Define observability strategy across databases, services, infrastructure, and dependencies
  • Ensure effective metrics, logs, traces, dashboards, and alerting for proactive detection and response
  • Drive Infrastructure as Code, operational automation, and self-healing mechanisms
  • Reduce manual toil through standardized workflows for provisioning, scaling, backup, recovery, and routine operations
  • Improve system and database performance through tuning, capacity planning, and architecture reviews
  • Partner on cost optimization initiatives across cloud infrastructure, database services, storage, and licensing
  • Own RTO/RPO alignment, resilience planning, backup strategy, and disaster recovery exercises
  • Improve production readiness for releases, migrations, and major operational events
  • Partner with Development, Platform Engineering, Infrastructure, Security, and Product teams to improve service reliability end to end
  • Act as a technical leader and escalation point for reliability concerns across the stack

Benefity

  • Multisport
  • Private Medical Care - LuxMed
  • Travel Insurance (available with Private Medical Care – LuxMed)
  • Life Insurance – UNUM
  • Employee Assistance Program (EAP)
  • Employee Stock Purchase Plan (ESPP)
  • Employee Referral Program
  • Eyeglasses Reimbursement
  • LinkedIn Learning
  • Bereavement leave
  • Volunteer leave
  • Maternal allowance supplement for 10 weeks in the first year after birth or adoption
  • Employee Capital Plans 1,5% (PPK)
  • Tax-deductible costs
Karta sportowa
Opieka zdrowotna
Ubezpieczenie
Udziały pracownicze

Inne informacje

CAE is committed to providing equal opportunities to all applicants regardless of protected characteristics. Reasonable accommodations are available upon request. The recruitment process may use AI-supported tools with human decision-making at every step.

CAE

CAE

3 aktywne oferty

Zobacz wszystkie oferty
Aplikuj teraz