Principal DevOps Engineer
Brak informacji o wynagrodzeniu
SeniorFull-time
#440665·Dodano wczoraj·0
Źródło: CommITTech Stack / Keywords
AWSTerraformKubernetesAmazon EKSEKSEC2RDSAuroraVPCTransit Gateway Route 53 IAM AWS Organizations Argo CD Prometheus Grafana OpenTelemetry CloudWatch Python Go Bash Kafka MSK Vault
Firma i stanowisko
CommIT is building a new gaming platform designed for sportsbook peak traffic with millions of players, focusing on highly available and compliant cloud infrastructure.
Wymagania
- 10+ years of DevOps/platform/infrastructure engineering, including 3+ years at staff/principal level
- Deep hands-on AWS experience: EKS, EC2, RDS/Aurora, networking (VPC, Transit Gateway, Route 53), IAM at scale, multi-account architectures
- Expert-level infrastructure as code with Terraform
- Strong Kubernetes operational expertise: day-2 operations, upgrades, capacity, cost
- Proven track record building CI/CD and developer platform engineering teams
- Observability and SLO/error-budget practice (Prometheus/Grafana, OpenTelemetry, CloudWatch)
- Experience with high-availability, high-throughput production systems and incident playbooks
- Strong scripting/programming skills (Python, Go, or Bash); comfortable reading application code
- Experience designing backup, disaster recovery, and business continuity for critical systems
- Excellent written and spoken English
Nice to have:
- Hands-on use of AI coding tools (Claude Code, Codex) for infrastructure and ops automation
- Security engineering: cloud security posture management, secrets management (e.g. Vault), vulnerability management, SAST/DAST/SCA, ISO 27001/SOC 2/PCI DSS
- Experience in iGaming, sports betting, fintech or other regulated high-transaction domains
- Migration from third-party vendor platforms
- Event-streaming infrastructure (Kafka/MSK) and large-scale database operations
- Knowledge of Polish and/or Spanish
Obowiązki
- Design the AWS foundation from scratch: multi-account architecture (AWS Organizations), landing zone, VPC and networking, IAM strategy, cost governance
- Build and operate the container platform — Amazon EKS, service mesh, autoscaling tuned for spiky sportsbook load, multi-AZ and multi-region resilience
- Define infrastructure as code: Terraform for all infrastructure, GitOps delivery (Argo CD or similar), paved-road CI/CD pipelines
- Establish the observability stack — metrics, logging, tracing, alerting (Prometheus/Grafana, OpenTelemetry, CloudWatch), and drive an SLO-based reliability practice with error budgets
- Own production readiness: incident response, on-call design, runbooks, chaos/load testing, blameless postmortems
- Build the infrastructure side of migration off the current third-party platform: dual-running environments, data migration pipelines, cutover mechanics
- Embed compliance into the platform: audit trails, environment segregation, backup/DR, and controls for gaming regulators
- Partner with the AI coding platform team to provision and operate infrastructure behind AI-assisted development
- Mentor engineers across teams on cloud-native and operational best practices; set organization-wide standards
CommIT
21 aktywnych ofert