Ontology Engineer
Tech Stack / Keywords
Firma i stanowisko
Andersen is a global technology provider delivering data analytics and AI-powered solutions that help organizations better understand customer behavior, optimize marketing performance, and make informed business decisions. The project focuses on building a large-scale Knowledge Graph and Identity platform that models relationships between audiences, content, brands, devices, and behavioral data, combining graph technologies, big data processing, and AI-powered enrichment to support identity resolution and advanced analytics.
Wymagania
- Experience implementing RDF/RDFS/OWL ontologies in graph databases
- Proficiency writing SPARQL queries and graph traversals
- Skilled in PySpark and Databricks for large-scale event data pipelines
- Strong programming skills in Python with clean, tested, and documented code
- Knowledge of SHACL validation for graph consistency and data quality
- Experience applying embedding models and vector search for entity matching
- Ability to support ontology versioning and changelog documentation
- Hands-on knowledge of semantic web standards and entity resolution techniques
Nice to have:
- Experience with Amazon Neptune, Stardog, or RDF-native triplestores
- Familiarity with data virtualization technologies
- Experience with LLM APIs or RAG approaches for information extraction
- Domain knowledge in media, entertainment, ad tech, TV viewership, or audience data
- Exposure to identity resolution, probabilistic record linkage, or device graphs
Obowiązki
- Implement and extend Samba's RDF/RDFS/OWL ontology schemas in the graph database under direction
- Build and maintain SHACL validation shapes for consistency checks and data quality
- Support ontology versioning, changelog documentation, and consistency checking
- Write efficient SPARQL queries and graph traversals for data science and product use cases
- Contribute to event-to-ontology data transformation with PySpark/Databricks pipelines
- Implement derivation logic validated against SHACL before graph load
- Support incremental graph refresh aligned with batch cadence
- Write production-quality, well-tested, and documented Python code
- Use PySpark and Databricks to process and transform high-volume data
- Apply embedding-based approaches to entity matching and ontology alignment
- Contribute to tooling, documentation, and reusable graph components
- Collaborate with data engineering on pipeline design and incremental ingestion
- Participate in ontology design reviews and cross-functional working groups
- Work with product and operations teams to translate use cases into graph schema updates
- Develop expertise in W3C semantic web standards, RDF-native graph databases, and entity resolution
Benefity
- Opportunity to work with leaders in FinTech, Healthcare, Retail, Telecom and other industries
- Possibility to change projects and develop domain expertise
- Professional, financial, and career growth with mentoring and adaptation systems
- Annual bonus up to EUR 1,000 depending on expertise level
- Access to corporate training portal with extensive knowledge base
- Bright corporate life including parties, pizza days, entertainment, and snacks
- Certification compensation for AWS, PMP, and others
- Referral program
- Private health insurance and sports compensation depending on employment type
Inne informacje
Informujemy, że administratorem danych jest Andersen Soft UAB z siedzibą w Krakow, ul. Al. Pokoju 18, 31 - 564 dalej jako "administrator"). Masz prawo do żądania dostępu do swoich danych osobowych, ich sprostowania, usunięcia lub ograniczenia przetwarzania, prawo do wniesienia sprzeciwu wobec przetwarzania, a także prawo do przenoszenia danych oraz wniesienia skargi do organu nadzorczego. Dane osobowe przetwarzane będą w celu realizacji procesu rekrutacji. Podanie danych w zakresie wynikającym z ustawy z dnia 26 czerwca 1974 r. Kodeks pracy jest obowiązkowe. W pozostałym zakresie podanie danych jest dobrowolne. Odmowa podania danych obowiązkowych może skutkować brakiem możliwości przeprowadzenia procesu rekrutacji. Administrator przetwarza dane obowiązkowe na podstawie ciążącego na nim obowiązku prawnego, zaś w zakresie danych dodatkowych podstawą przetwarzania jest zgoda. Dane osobowe będą przetwarzane do czasu zakończenia postępowania rekrutacyjnego i przez okres możliwości dochodzenia ewentualnych roszczeń, a w przypadku wyrażenia zgody na udział w przyszłych postępowaniach rekrutacyjnych - do czasu wycofania tej zgody. Zgoda na przetwarzanie danych osobowych może zostać wycofana w dowolnym momencie. Odbiorcą danych jest serwis Just Join IT oraz inne podmioty, którym powierzyliśmy przetwarzanie danych w związku z rekrutacją.
Andersen
60 aktywnych ofert