Senior Site Reliability Engineer – Cloud and Java Systems
~19.5k - 23.5k PLN19 500 - 23 500 PLN/ mies.B2B
Brak informacji o wynagrodzeniu
SeniorFull-time·B2B·Umowa o pracę
#445378·Dodano wczoraj·0
Źródło: ITDSTech Stack / Keywords
AnsibleDockerGCPGrafanaInfluxDBJavaJenkinsLokiOraclePostgreSQLPrometheusSplunk
Firma i stanowisko
This position is for a leading international bank focused on financial technology solutions facilitating global banking operations and risk management.
Wymagania
- 4+ years of experience in developing and supporting distributed Java systems.
- Strong understanding of application lifecycle management tools: JIRA, Confluence, Ansible, CI/CD pipelines.
- Hands-on experience with modern observability tools: Grafana, InfluxDB, Prometheus, Splunk, Loki.
- Experience managing large cross-platform environments troubleshooting complex applications.
- Basic Cloud knowledge (GCP preferred), with familiarity in automation tools like Jenkins and Ansible.
- Good command of English (Communicative).
- Proven ability to lead technical discussions and work across regions.
- Core Java development knowledge and experience with relational databases such as Oracle or PostgreSQL.
- Methodical troubleshooting approach with disaster recovery expertise.
Nice to have:
- Experience with cloud platforms (GCP), automation, and containerization technologies like Docker.
- Knowledge of Unix/Linux environments and scripting.
- Additional certifications in Cloud, SRE, or related fields.
Obowiązki
- Manage application support operations focusing on resiliency, availability, and system health monitoring.
- Coordinate resolution of production incidents, conducting root cause analysis and post-mortem reviews.
- Develop observability tools and techniques for monitoring, alerting, incident detection, and capacity management.
- Apply SRE principles to improve platform reliability, reduce toil, and optimize performance.
- Implement and manage logging, monitoring, and alerting frameworks across hybrid cloud environments.
- Lead technical discussions and collaborate across regional teams for incident resolution.
- Maintain comprehensive documentation of recovery steps and contribute to knowledge sharing.
- Participate in on-call rotations, supporting critical systems during weekends and evenings.
Benefity
- Hybrid work model (remote and office).
- Opportunity to work for a leading international bank on enterprise-scale systems.
- Participation in on-call rotations supporting critical systems.
Inne informacje
Only candidates with an existing legal right to work in the European Union will be considered for this role.
ITDS
411 aktywnych ofert