Lead Data Platform Engineer | Remote
Tech Stack / Keywords
Firma i stanowisko
A large-scale e-commerce marketplace operating in Central and Eastern Europe, serving millions of users. The company runs a highly distributed, cloud-native technology stack and processes high traffic and transaction volumes. Engineering teams focus on scalability, reliability, and automation, with strong emphasis on platform engineering, data-driven decision making, and continuous delivery in a fast-growing product environment.
Wymagania
- Strong commercial experience with Python development.
- Hands-on experience with Google Cloud Platform (GCP).
- Practical experience with BigQuery and cloud-based data platforms.
- Experience with Apache Airflow, preferably Cloud Composer.
- Experience building platform engineering, developer tooling, or internal self-service solutions.
- Knowledge of Infrastructure as Code practices and Terraform.
- Experience with GitOps principles and modern software delivery practices.
- Good understanding of data engineering concepts and Data Product lifecycle management.
- Experience with PySpark and/or Apache Spark ecosystems.
- Familiarity with FastAPI, Pydantic, and modern Python tooling.
- Knowledge of CI/CD processes and source control best practices.
- Experience working in Agile development environments.
- Strong problem-solving skills and ability to work independently.
- Effective communication skills and ability to collaborate with cross-functional teams.
- Professional proficiency in Polish and English.
Nice to have:
- Experience with Vertex AI or other GenAI platforms.
- Hands-on experience with GitHub Copilot, Copilot extensions, plugins, or AI-assisted development solutions.
- Experience developing Agentic AI workflows or AI automation capabilities.
- Knowledge of Backstage plugin development.
- Experience with Dataproc.
- Familiarity with dbt and Kedro.
- Experience with ArgoCD and Crossplane.
- Experience building observability, telemetry, or platform analytics solutions.
- Experience working in large-scale data environments supporting analytics and Machine Learning workloads.
Obowiązki
- Design and develop tools that standardize, automate, and simplify Data Platform operations.
- Build and maintain internal CLI tools for common platform and Data Product management tasks.
- Develop and evolve Data Product templates based on modern data engineering practices.
- Design and implement visualizations and operational insights within Backstage.
- Collect and analyze platform usage metrics, audit data, and adoption statistics.
- Design and implement Agentic AI workflows and developer-assistance capabilities.
- Develop reusable AI skills and automation components that standardize work with data products and data assets.
- Create mechanisms for distribution, lifecycle management, and monitoring of AI skills.
- Maintain and enhance a shared Apache Airflow platform based on Cloud Composer.
- Build reusable libraries, operators, and common components for Airflow DAG development.
- Develop and maintain infrastructure-as-code assets and Terraform modules.
- Contribute to GitOps adoption initiatives using tools such as Backstage, GitHub, ArgoCD, and Crossplane.
- Collaborate with platform, data engineering, analytics, and machine learning teams to improve platform usability and engineering efficiency.
Benefity
- Stable long-term cooperation with opportunity for extension.
- Opportunity to work with modern technologies in an international environment.
- Professional growth and continuous development opportunities.
- Participation in challenging and impactful projects.
Inne informacje
Informujemy, że administratorem danych jest z siedzibą w , ul.(dalej jako "administrator"). Masz prawo do żądania dostępu do swoich danych osobowych, ich sprostowania, usunięcia lub ograniczenia przetwarzania, prawo do wniesienia sprzeciwu wobec przetwarzania, a także prawo do przenoszenia danych oraz wniesienia skargi do organu nadzorczego. Dane osobowe przetwarzane będą w celu realizacji procesu rekrutacji. Podanie danych w zakresie wynikającym z ustawy z dnia 26 czerwca 1974 r. Kodeks pracy jest obowiązkowe. W pozostałym zakresie podanie danych jest dobrowolne. Odmowa podania danych obowiązkowych może skutkować brakiem możliwości przeprowadzenia procesu rekrutacji. Administrator przetwarza dane obowiązkowe na podstawie ciążącego na nim obowiązku prawnego, zaś w zakresie danych dodatkowych podstawą przetwarzania jest zgoda. Dane osobowe będą przetwarzane do czasu zakończenia postępowania rekrutacyjnego i przez okres możliwości dochodzenia ewentualnych roszczeń, a w przypadku wyrażenia zgody na udział w przyszłych postępowaniach rekrutacyjnych - do czasu wycofania tej zgody. Zgoda na przetwarzanie danych osobowych może zostać wycofana w dowolnym momencie. Odbiorcą danych jest serwis Just Join IT oraz inne podmioty, którym powierzyliśmy przetwarzanie danych w związku z rekrutacją.
DCV Technologies
66 aktywnych ofert