Senior SRE

Brak informacji o wynagrodzeniu
SeniorFull-time
#452714·Dodano 3 dni temu·0
Źródło: LinkGroup
Aplikuj teraz

Tech Stack / Keywords

TerraformPulumiAWSKubernetesPrometheusDatadogGrafanaDynatraceAIFinOps

Firma i stanowisko

LinkGroup is the employer of this position located in Warsaw.

Wymagania

  • Bachelor's degree in Computer Science or 10+ years of experience in infrastructure engineering, DevOps, or SRE with technical leadership.
  • Mastery of public cloud environments, specifically AWS, and IaC automation at enterprise scale.
  • Profound understanding of Kubernetes architecture and operational experience with distributed, high-availability container platforms.
  • History of building and scaling automated delivery pipelines.
  • Expert-level knowledge of modern observability stacks and enterprise telemetry standards.
  • Experience embedding AI or machine learning tools into operational workflows.
  • Strong mentorship and architectural direction capabilities across multiple teams.
  • Proven ability to design robust backends for data-intensive and high-concurrency microservices.
  • Deep understanding of cloud networking, FinOps/cost tuning, and large-scale distributed system security.
  • Excellent communication skills translating business objectives to technical realities.
  • Motivation to innovate in a rapidly expanding and agile organization valuing continuous learning and cross-functional teamwork.

Obowiązki

  • Blueprint and enforce enterprise-wide SRE methodologies, ensuring peak performance and high availability.
  • Drive infrastructure-as-code (IaC) vision utilizing tools like Terraform or Pulumi to manage large-scale cloud (AWS) deployments.
  • Take full ownership of containerization strategy and roadmap for expansive Kubernetes clusters.
  • Collaborate to revamp continuous integration and deployment (CI/CD) workflows promoting automation.
  • Design monitoring and telemetry ecosystems using tools such as Prometheus, Datadog, Grafana, or Dynatrace.
  • Command incident response, root cause analysis, and post-event reviews.
  • Orchestrate disaster recovery protocols and capacity forecasting.
  • Mentor staff-level peers and integrate AI-assisted monitoring and auto-remediation tools.
  • Strengthen cloud security postures and produce architectural documentation for distributed engineering hubs.
Link Group

Link Group

492 aktywne oferty

Zobacz wszystkie oferty
Aplikuj teraz