Senior Site Reliability Engineer

53.3k - 119.9k USD/ rok.UoP
SeniorFull-time·Umowa o pracę
#413596·Dodano miesiąc temu·0
Źródło: Remote
Aplikuj teraz

Tech Stack / Keywords

GoAIArchitectureSecurityDevOpsKubernetesDockerCloud

Firma i stanowisko

Remote is solving modern organizations’ biggest challenge – navigating global employment compliantly with ease. They enable businesses to recruit, pay, and manage international teams. Remote operates fully remote, with team members working asynchronously across six continents. Innovation, including Automation and AI capabilities, is integrated into every role.

Wymagania

  • Solid professional experience in SRE, DevOps, or Platform Engineering.
  • Hands-on experience with Kubernetes, operating and scaling production clusters.
  • Experience with container tooling such as Docker.
  • Experience building and managing cloud infrastructure on AWS or similar.
  • Strong infrastructure-as-code skills with Terraform.
  • Experience with reliability frameworks: SLOs, SLIs, error budgets, alerting strategies.
  • Observability background with tools like OpenTelemetry, Grafana, Prometheus.
  • Proficiency with CI/CD tools such as GitLab CI, GitHub Actions, or similar.
  • Comfortable with Golang, Bash/scripting; broader programming is a plus.
  • Practical, embedded use of AI in infrastructure, operations, or development workflows.
  • Clear and thoughtful communication in asynchronous, global settings.
  • Proactive, curious, and ownership-driven mindset.
  • Collaborative and respectful across cultures and time zones.

Nice to have:

  • Experience with one back-end programming language (Elixir, Node.js, Python, etc.).
  • Experience running and configuring Linux systems in non-cloud environments.
  • Security knowledge from both defensive and offensive perspectives.

Obowiązki

  • Lead solution discovery and delivery for complex reliability and infrastructure problems autonomously.
  • Contribute to platform architecture, tooling, and roadmap; influence team priorities and technical initiatives.
  • Define and operate reliability practices: SLOs, SLIs, error budgets, alerting, and observability.
  • Resolve cross-team issues by creating reusable fixes and runbooks.
  • Operationalize AI for the team with reusable prompts, tooling, and secure, observable agentic workflows.
  • Mentor less-senior engineers and participate in hiring, onboarding, and RFC discussions.
  • Collaborate with Security on platform hardening and threat mitigation.
  • Participate in incident response and on-call rotations to maintain system reliability.

Benefity

  • Work from anywhere.
  • Flexible paid time off.
  • Flexible working hours with an async-first culture.
  • 16 weeks paid parental leave.
  • Budget towards co-working spaces, learning, wellness, and gym memberships.
  • Mental health support services.
  • Stock options.
  • Home office budget and IT equipment.
Elastyczne godziny
Płatny urlop
Udziały pracownicze

Inne informacje

This position prioritizes candidates located in Europe due to diversity and timezone requirements. Applications and interviews are conducted in English. Remote is an equal opportunity employer encouraging applications from all backgrounds and offers accommodations during the hiring process. The company embraces AI as a tool while prioritizing human creativity and authenticity.

Remote

Remote

10 aktywnych ofert

Zobacz wszystkie oferty
Aplikuj teraz