XM
XM
New

Site Reliability Engineer (SRE)

Brak informacji o wynagrodzeniu
SeniorFull-time
#402943·Dodano wczoraj·0
Źródło: nofluffjobs.com
Aplikuj teraz

Tech Stack / Keywords

AWSSREMonitoringAPMObservabilityIncident ManagementChange ManagementTroubleshootingCI/CDIaCKafkaRDBMSPythonGo

Wymagania

  • BSc/MSc degree in Computer Science or related field
  • 5+ years of cloud services experience, with at least 3 years on AWS cloud
  • 3+ years of experience in SRE or a similar role
  • Experience with monitoring, APM, logging, and notification tools
  • Familiarity with incident, problem and change management procedures and practices
  • Advanced knowledge of SRE practices and methods
  • Understanding and practice of Service Levels
  • Strong troubleshooting skills and the ability to mentor others
  • Extensive experience with Kubernetes and related technologies, services, and ecosystem
  • Advanced knowledge of CI/CD, Infrastructure as Code (IaC) concepts and tools, especially HCL Terraform and AWS CloudFormation
  • Experience with versioning tools like Git
  • Strong organizational and documentation skills
  • Exceptional time management and research abilities
  • Advanced Linux, networking, and scripting skills

Nice to have:

  • Experience with platforms like Kafka (MSK)
  • Experience with RDBMSs, particularly Postgres and MySQL
  • Knowledge of scripting languages such as Python or Go

Obowiązki

  • Honor and practice the Resiliency pillar of the Well Architected Framework in all tasks and responsibilities
  • Conduct Chaos Engineering experiments and relevant exercises to improve resiliency and fault-tolerance
  • Research workloads for migrating to the cloud with minimal disruption and impact
  • Monitor cloud migration projects to ensure seamless transitions
  • Design, consult, re-platform, and re-factor the observability of current cloud infrastructure
  • Coordinate with other IT departments and teams regarding observability for both individual and organizational needs
  • Regularly assess cloud deployments for compliance with the company’s standards and best practices
  • Investigate and correct areas where observability is lagging
  • Stay up to date and provide training on new and current technologies, services, tools, methodologies, and practices
  • Occasionally participate in service capacity planning, software performance analysis, and system tuning
  • Mentor colleagues in technical skills and knowledge
  • Analyze, oversee, and remediate the company’s resiliency
  • Participate in on-call support 24/7 based on a rotation schedule

Benefity

  • Attractive remuneration package and perks
  • Intellectually stimulating work environment
  • Continuous personal development and international training opportunities

Inne informacje

All applications will be treated with strict confidentiality!

XM

XM

7 aktywnych ofert

Zobacz wszystkie oferty
Aplikuj teraz