Senior Site Reliability Engineer
28k - 32k PLN28 000 - 32 000 PLN/ mies.B2B
28k - 32k PLN28 000 - 32 000 PLN/ mies.UoP
SeniorFull-time·B2B·Umowa o pracę
#393100·Dodano 2 dni temu·1
Źródło: nofluffjobs.comTech Stack / Keywords
DockerDatadogObservabilityCosmos DBMySQLCI/CDGitHub ActionsELK StackGrafanaPrometheus.NET
Firma i stanowisko
OneRail Poland Sp. z o.o. is hiring a Senior Site Reliability Engineer to ensure the availability, performance, scalability, observability, and resilience of their SaaS platform. The company values deep expertise in cloud-native platforms, distributed systems, observability, and performance optimization, working closely with Engineering, Product, and Operations teams to improve platform reliability and operational efficiency.
Wymagania
- Bachelor’s degree in Computer Science, Engineering, or a related technical field.
- 5+ years of experience in Site Reliability Engineering, Platform Engineering or a related role.
- Strong experience with cloud-native architectures and distributed systems.
- Hands-on experience implementing observability solutions, including metrics, logging, tracing, and application performance monitoring.
- Experience designing scalable backend architecture using Node.js, TypeScript, and .NET.
- Strong knowledge of database administration, performance tuning, and optimization, including Azure Cosmos DB and MySQL.
- Experience building automation frameworks, scripting solutions, and operational tooling.
- Familiarity with CI/CD pipelines and continuous integration practices using GitHub Actions or similar platforms.
- Experience with containerization technologies such as Docker.
- Strong understanding of event-driven architectures, messaging systems, and real-time data processing platforms.
- Experience with monitoring and observability tools such as Datadog, OpenTelemetry, ELK Stack, App Insights, Grafana, or Prometheus.
- Excellent troubleshooting, analytical, and problem-solving skills.
- Strong written and verbal communication skills with ability to collaborate across technical and business teams.
- Advanced proficiency in English and Polish (B2+).
Obowiązki
- Serve as the platform subject matter expert, mentoring engineering teams on reliability, scalability, security, and operational best practices.
- Design, implement, and maintain observability solutions covering logs, metrics, traces, APM, and alerting across all platform services.
- Benchmark, analyze, and optimize application performance, cloud services, integrations, databases, and distributed systems.
- Build and maintain performance, load, chaos, and resilience testing frameworks to proactively identify system weaknesses.
- Perform advanced analysis, tuning, and optimization of structured and unstructured databases to improve performance and scalability.
- Develop automation frameworks and operational tooling that eliminate manual processes and reduce human error.
- Lead incident response efforts, conduct root cause analysis, and drive continuous improvement through postmortem reviews and reliability initiatives.
- Optimize the performance of databases, caches, streaming platforms, message brokers, and backend services.
- Collaborate closely with Engineering, Product, and Operations teams to deliver highly available, production-grade platform solutions.
- Establish and maintain monitoring, alerting, and reliability standards across the platform.
- Create and maintain technical documentation, runbooks, architectural diagrams, and operational procedures.
- Evaluate emerging technologies and recommend solutions that improve platform reliability, scalability, observability, and operational efficiency.
Benefity
- Sport subscription
- Private healthcare
- Flat structure
- Small teams
- Free coffee
- Playroom
- Free snacks
- Free beverages
- Free lunch
- Bike parking
- Free parking
- In-house trainings
- In-house hack days
- Modern office
- Startup atmosphere
- No dress code
Karta sportowa
Opieka zdrowotna
Szkolenia wewnętrzne
Napoje w biurze
Darmowe przekąski
OneRail Poland Sp. z o.o.
20 aktywnych ofert