Senior Engineer - SRE & Infrastructure Services

Brak informacji o wynagrodzeniu
SeniorFull-time
#394673·Dodano wczoraj·0
Źródło: EPAM Systems
Aplikuj teraz

Tech Stack / Keywords

Web ServicesGrafanaJenkinsKubernetesPrometheusGoGoogle Cloud PlatformJava

Firma i stanowisko

EPAM is global leader in AI transformation engineering and integrated consulting, serving Forbes Global 2000 companies and ambitious startups. With over thirty years of expertise in custom software, product and platform engineering, we empower our clients to become AI-Native enterprises, driving measurable value from innovation and digital investments.

Wymagania

  • Minimum 7 years of professional experience in software development, DevOps, and/or Site Reliability Engineering
  • Minimum 3 years of experience building and maintaining CI/CD pipelines and SRE automation within cloud environments at scale
  • Experience with monitoring and alerting platforms such as PagerDuty, Prometheus, and Grafana
  • Hands-on experience deploying and managing cloud infrastructure
  • Experience with Amazon Web Services (e.g., EC2, Elasticsearch, Lambda, CloudFormation)
  • Working experience with at least one additional cloud provider (Google Cloud Platform, Microsoft Azure, or OCI)
  • Experience with CI/CD toolchains (Jenkins, Kubernetes, Flux)
  • Proficiency in one or more of the following languages: Python, Go, Java, or C
  • Minimum English language level of B1+

Nice to have:

  • Experience applying Generative AI or ML-based tooling within an operations context
  • Experience with DevSecOps practices and security automation
  • Background in architecting monitoring and management systems for enterprise SaaS products

Obowiązki

  • Design and implement CI/CD pipelines leveraging Kubernetes, Flux, and related cloud-native technologies
  • Implement and maintain monitoring and management solutions for cloud-based products using a combination of commercial off-the-shelf (COTS) and in-house tooling
  • Collaborate with architects and development teams on standardized, scalable approaches for log management, service components, and infrastructure elements
  • Evaluate and integrate AI-enabled tooling across observability, pipeline efficiency, and SRE troubleshooting workflows
  • Develop DevOps tooling that reduces manual toil, strengthens security posture, and minimizes human error
  • Build and maintain resilient, self-scaling systems that minimize customer impact while supporting a sustainable operational environment
  • Participate in incident response, root cause analysis, and post-incident review processes

Benefity

  • Opportunity to work remotely within Poland or in hybrid mode
  • Chance to work abroad for up to 60 days annually
  • Business-driven relocation opportunities
  • Career development programs including certification (Anthropic, Gemini, GCP, Azure, AWS)
  • English classes
  • Stable pay
  • Participation in the Employee Stock Purchase Plan with a 15% discount
  • Benefits package including health insurance, multisport, shopping vouchers
  • Referral bonuses up to $2,000
  • Offices with entertainment and relaxation zones, table tennis, football, free snacks and coffee
  • Corporate, social, and well-being events
Elastyczne godziny
Dofinansowanie szkoleń
Budżet konferencyjny
Kursy językowe
Opieka zdrowotna
Karta sportowa
Ubezpieczenie
Udział w zyskach
Premie
Napoje w biurze
Darmowe przekąski
Spotkania integracyjne
EPAM Systems

EPAM Systems

279 aktywnych ofert

Zobacz wszystkie oferty
Aplikuj teraz