Data Engineer (Tax Care)
Brak informacji o wynagrodzeniu
MidFull-time
#443347·Dodano dziś·0
Źródło: XTBTech Stack / Keywords
PythonREST APISQLT-SQLMS SQL ServerApache SparkPySparkdbtDatabricksAirflow
Firma i stanowisko
XTB is a global company from the financial industry, focusing on online trading of financial instruments. It is the largest FinTech in Poland and a leader in Central and Eastern Europe, with operations in several countries including Asia and South America. The company invests in employee development and offers various training and development programs.
Wymagania
- Minimum 3 years experience as a Data Engineer or related software engineering role
- Strong Python skills, producing clean, testable production code including backend services (e.g., FastAPI, Flask) exposing REST APIs
- Advanced SQL skills on large datasets and practical experience with T-SQL / MS SQL Server, including legacy stored procedures
- Hands-on experience with Apache Spark (PySpark)
- Experience with dbt including models, tests, macros, and layered warehouse design
- Strong understanding of data warehouse design principles and data modeling
- Practical knowledge of ETL/ELT processes
- Experience analyzing and migrating legacy code to newer technologies
- Familiarity with data quality, monitoring, and pipeline reliability
- Experience with workflow orchestration tools (e.g., Airflow, Dagster)
- Knowledge of CI/CD pipeline building and maintenance
- Understanding of REST standards for system integration
- Strong analytical problem-solving skills and attention to detail, able to work under strict monthly accounting deadlines
- Effective communication skills to collaborate with engineers and non-technical stakeholders
- Openness to learning and exploring new technologies and methodologies
Nice to have:
- Experience with Databricks or other lakehouse platforms (Delta Lake, Unity Catalog)
- Experience with Snowflake analytical platform
- Basic front-end skills (e.g., React, Streamlit, Dash)
- Experience with SSIS and SQL Server Agent (useful context)
- Experience in financial, accounting, or tax domains or with reconciliation-heavy data
- Understanding of gRPC for system integration
- Experience with event streaming systems such as Kafka or Pub/Sub
- Experience building near-real-time or real-time pipelines
- Experience with Infrastructure-as-Code tools
- Understanding of CDC (Change Data Capture) and integration from transactional systems
Obowiązki
- Designing, building and maintaining data pipelines feeding accounting and tax systems using SQL, Python, dbt, and Apache Spark / PySpark
- Migrating data platform to Databricks lakehouse, rewriting legacy MS SQL Server stored procedures into Spark / Databricks SQL
- Building and maintaining a Python backend exposing data over a REST API with web UI, covering data model, API design, and deployment
- Working with raw data and designing integration methods for the data platform
- Developing CI/CD pipelines for data engineering solutions
- Creating and maintaining Infrastructure as Code solutions
- Integrating the data platform with other systems and applications
- Participating in architectural decision-making and defining engineering standards
- Conducting code reviews, documenting solutions, and sharing knowledge with team members
Benefity
- Real influence on company and product development
- Work with an experienced team eager to share knowledge
- Clear development vision with regular feedback and career paths
- Regular team-building meetings
- Training budget for courses and conferences
- Extra day off on birthday and additional day off for parents
- Equipment tailored to individual needs
- Private medical care and group insurance
- Access to e-learning platform for English and benefits platform
- Access to wellbeing platform with workshops and private therapy sessions
- Flexible work location: remote, office in Warsaw, or coworking space in candidate's city
Dofinansowanie szkoleń
Płatny urlop
Opieka zdrowotna
Ubezpieczenie
XTB
40 aktywnych ofert