Data Engineer (Spark)
15.1k - 21k PLN15 120 - 21 000 PLN/ mies.B2B
MidFull-time·B2B
#408965·Dodano 2 dni temu·2
Źródło: nofluffjobs.com⚠️Uwaga: ta oferta może już nie być aktualna. Sprawdź na stronie pracodawcy, czy rekrutacja jest nadal otwarta.
Tech Stack / Keywords
PythonSQLSparkAirflowClouderaCI/CDKubernetesKafkaNiFiJavaScalaDockerDatabricksMLOpsDevOpsIceberg
Firma i stanowisko
Addepto is a leading AI consulting and data engineering company that builds scalable, ROI-focused AI solutions for some of the world's largest enterprises and pioneering startups. The company specializes exclusively in Artificial Intelligence and Big Data. It has developed its own product called ContextClue and contributes open-source solutions to the AI community. Addepto is part of KMS Technology, a US-based global technology group, which combines AI specialization with enterprise-scale delivery capabilities.
Wymagania
- At least 3 years of commercial experience implementing, developing, or maintaining Big Data systems, data governance, and data management.
- Strong programming skills in Python (or Java/Scala) with clean code and OOP design.
- Hands-on experience with Big Data technologies like Spark, Cloudera, Data Platform, Airflow, Kafka, NiFi, Docker, and Iceberg.
- Excellent understanding of dimensional data and data modeling techniques.
- Experience deploying solutions in cloud environments.
- Consulting experience with strong communication and client management skills.
- Ability to work independently and take ownership of project deliverables.
- Fluent in English (at least C1 level).
- Bachelor’s degree in technical or mathematical studies.
Nice to have:
- Experience with MLOps frameworks such as Kubeflow or MLflow.
- Familiarity with Databricks and/or dbt.
Obowiązki
- Develop and maintain a high-performance data processing platform for automotive data, ensuring scalability and reliability.
- Design and implement data pipelines processing large data volumes in streaming and batch modes.
- Optimize data workflows for efficient ingestion, processing, and storage with technologies like Spark, Cloudera, and Airflow.
- Work with data lake technologies such as Iceberg to manage structured and unstructured data.
- Collaborate with cross-functional teams to integrate diverse data sources.
- Monitor and troubleshoot the platform to ensure high availability, performance, and data accuracy.
- Leverage cloud services including AWS for infrastructure management and scaling.
- Write and maintain high-quality code in Python (or Java/Scala) for data processing and automation tasks.
Benefity
- Work in a supportive team passionate about AI and Big Data.
- Engage with top-tier global enterprises and startups on international projects.
- Flexible work arrangements including remote work or modern offices.
- Professional growth through career paths, knowledge sharing, language classes, and sponsored trainings and conferences.
- Partnership with Databricks and Anthropic for training materials and certifications.
- Team-building events and integration budget.
- Celebration of work anniversaries, birthdays, and milestones.
- Access to medical and sports packages, eye care, psychotherapy, and coaching.
- Full work equipment including laptop and necessary devices.
- Opportunities to boost personal brand by speaking at conferences, writing blogs, or participating in meetups.
- Smooth onboarding with a dedicated buddy.
- Friendly, supportive, and autonomous culture.
Elastyczne godziny
Kursy językowe
Budżet konferencyjny
Dofinansowanie szkoleń
Spotkania integracyjne
Opieka zdrowotna
Karta sportowa
Płatny urlop
Premie
Napoje w biurze
Addepto
39 aktywnych ofert