Senior Data Engineer / Tech Lead (Spark)
Tech Stack / Keywords
Firma i stanowisko
Addepto is a leading AI consulting and data engineering company that builds scalable, ROI-focused AI solutions for some of the world's largest enterprises and pioneering startups, including Rolls Royce, Continental, Porsche, ABB, and WGU. The company focuses exclusively on Artificial Intelligence and Big Data, helping organizations unlock the full potential of their data through systems designed for measurable business impact and long-term growth. Addepto is part of KMS Technology, a US-based global technology group, combining AI specialization with enterprise-scale delivery capabilities.
Wymagania
- At least 4 years of commercial experience in Data Engineering or Big Data projects
- Strong programming skills in Python
- Very good knowledge of SQL
- Strong hands-on experience with Apache Spark and distributed data processing
- Practical experience with workflow orchestration tools, preferably Apache Airflow
- Experience with Apache Iceberg and/or modern lakehouse architectures
- Experience working with Docker and Kubernetes
- Good understanding of data modelling, data pipelines, and data processing architectures
- Experience working with relational databases and/or query engines such as PostgreSQL or Trino
- General understanding of cloud environments and cloud-based data solutions
- Ability to take ownership of technical solutions and contribute to architectural and technical decisions
- Experience supporting other engineers through technical guidance, knowledge sharing, or mentoring
- Ability to work independently and take ownership of project deliverables
- Strong communication and collaboration skills
- Fluent English (at least C1 level)
- Bachelor’s degree in technical or mathematical studies
Nice to have:
- Experience with Kafka or other streaming technologies
- Familiarity with modern data modelling, transformation, and ingestion tools (e.g. dbt, dlt, Airbyte)
- Experience with CI/CD and DevOps tools (e.g. GitHub Actions, ArgoCD)
- Familiarity with monitoring and visualization tools (e.g. Grafana, Power BI)
- Experience with Databricks and cloud-based data environments
Obowiązki
- Design, develop, and maintain scalable data pipelines using Python, SQL, Spark, and Airflow
- Build and improve data processing solutions based on Apache Iceberg and modern lakehouse architectures
- Work with technologies such as PyArrow, PyIceberg, PostgreSQL, and Trino
- Take technical ownership of selected areas and contribute to technical and architectural decisions
- Combine hands-on development with technical leadership and knowledge sharing within the team
- Build and maintain containerized workloads using Docker and Kubernetes
- Ensure data pipelines are scalable, reliable, maintainable, and production-ready
- Collaborate with engineers and other stakeholders to understand business and technical requirements and translate them into effective solutions
- Contribute to engineering best practices, including code quality, testing, CI/CD, and automation
- Monitor, troubleshoot, and continuously improve existing data pipelines and platform components
- Use modern development practices, including AI-assisted development, to improve engineering efficiency
Benefity
- Work in a supportive team of passionate enthusiasts of AI and Big Data
- Engage with top-tier global enterprises and cutting-edge startups on international projects
- Flexible work arrangements allowing to work remotely or from modern offices and coworking spaces
- Career paths, knowledge-sharing initiatives, language classes, and sponsored training or conferences
- Partnership with Databricks and Anthropic for training materials and certifications
- Team-building events and integration budget
- Celebrations for work anniversaries, birthdays, and milestones
- Access to medical and sports packages, eye care, and well-being support services including psychotherapy and coaching
- Full work equipment including a laptop and other necessary devices
- Opportunities to boost personal brand by speaking at conferences, writing for the blog, or participating in meetups
- Smooth onboarding with a dedicated buddy and a friendly, supportive, autonomous culture
Inne informacje
Informujemy, że administratorem danych jest Addepto sp. z o.o. z siedzibą w Warszawie, ul. Swieradowska 47, 02-662. Podanie danych obowiązkowych jest wymagane do realizacji procesu rekrutacji. Dane będą przetwarzane do czasu zakończenia postępowania rekrutacyjnego oraz przez okres możliwości dochodzenia roszczeń. Zgoda na przetwarzanie danych osobowych może zostać wycofana w dowolnym momencie.
Addepto
92 aktywne oferty