Data Engineer with Databricks

Xebia is a global AI-first, digital transformation, and engineering partner. With over 25 years of experience and a team of 5,000 professionals across 16 countries, we help organizations design and build scalable products, platforms, and data-driven solutions.

We specialize in Artificial Intelligence, Data and Cloud, Intelligent Automation, and Digital Products, combining deep technical expertise with a strong focus on engineering excellence and a people-first culture.

In the CEE region, we’re a team of nearly 1,000 experts delivering modern applications, data platforms, and AI solutions for clients such as McLaren, Aviva, Deloitte, Spotify, Disney, ING, UPS, Tesco, Truecaller, AllSaints, Volotea, Schmitz Cargobull, Allegro, InPost, and many, many more. We work with leading technologies including AWS, Azure, GCP, Databricks, and Snowflake, and combine strong engineering culture with a consulting mindset and a continuous focus on growth and knowledge sharing.



You will be:

    • build, maintain, and optimize data pipelines using Python, Databricks, and PySpark for batch and real-time processing,
    • integrate data from multiple internal and external sources into the organization’s analytical platform,
    • maintain, enhance, and support the analytical database and the surrounding data ecosystem,
    • collaborate with business stakeholders and cross-functional teams to deliver data products that meet business needs and timelines,
    • identify technical debt and continuously improve platform reliability, maintainability, and performance,
    • build and manage Databricks Workflows for large-scale data orchestration,
    • develop, deploy, and maintain PyFunc models within production environments,
    • implement secure secrets and configuration management using Azure Key Vault,
    • automate data flows and business processes using Azure Logic Apps,
    • design, develop, and optimize ETL and ELT pipelines following data engineering best practices,
    • govern and manage data assets using Unity Catalog and Data Lakehouse principles,
    • implement messaging and event-driven workflows using Azure Service Bus,
    • collaborate closely with architects, analysts, data scientists, and engineering teams to deliver end-to-end data solutions,
    • contribute to CI/CD processes and deployment automation using Azure DevOps,
    • ensure high quality through testing, monitoring, troubleshooting, and continuous improvement activities,
    • depending on seniority, take ownership of technical areas, support architectural decisions, and mentor less experienced team members,

Your profile:

  • minimum 3 years of commercial experience in software engineering, data engineering, or related fields,
  • experience working in senior engineering, technical leadership, or highly autonomous delivery roles,
  • strong hands-on knowledge of Python as the primary development language,
  • strong experience with Databricks and PySpark in production environments,
  • proven experience building and supporting modern data analytics platforms,
  • practical experience working within Azure cloud environments or similar cloud ecosystems,
  • experience designing modular, reusable, scalable, and maintainable system components,
  • strong understanding of data engineering best practices and modern data platform architectures,
  • familiarity with Medallion Architecture or equivalent data modeling and design patterns,
  • ability to gather, clarify, and translate business requirements into technical solutions,
  • strong problem-solving, communication, and organizational skills,
  • proactive approach to identifying issues, opportunities, and improvement areas,
  • experience working with Azure DevOps and CI/CD practices,
  • excellent communication skills and ability to collaborate with technical and business stakeholders,
  • experience working within Agile delivery environments and engineering best practices,
  • quick learner with a strong interest in new technologies and continuous professional development,
  • Practical experience using AI-powered assistants (e.g. Claude Code, GitHub Copilot, Cursor) to improve productivity, quality, or decision-making in software delivery.

  • Work from the European Union region and a work permit are required.

Nice to have:

  • Databricks certification,
  • experience with Infrastructure as Code technologies such as Terraform or ARM templates,
  • experience with MLOps practices and automation of machine learning and data pipelines,
  • experience working with Kubernetes or other container orchestration platforms,
  • requirements engineering experience and business analysis awareness,
  • familiarity with event-driven architectures and distributed data processing systems,
  • Experience applying GenAI in a more structured way within the SDLC, including defined workflows, prompt patterns, or tool integrations embedded into daily work.

  • Interest in and familiarity with emerging AI-driven practices (e.g. agent-based workflows, automation patterns, AI-augmented development), with a willingness to explore and experiment beyond standard approaches.

Recruitment Process:

CV review – HR call – InterviewClient Interview – Decision

What this application asks

greenhouse

First Name, Last Name, Email, Phone, Resume/CV, Cover Letter

  • LinkedIn Profile optional
  • Website optional
  • Where did you find this job offer? choose one
  • What is your notice period? choose one
  • What is your preferred form of cooperation? choose one
  • Based on your preferred form of cooperation (per hour or monthly) what are your financial expectations?
  • What country do you currently reside in? choose one
  • Do you have documents entitling you to work in the European Union (valid work permits to work in the EU)? choose one
  • Do you speak English at a minimum B2 level? choose one
  • I declare that I agree to the processing of my Personal Data contained in the content of documents sent in response to the job/cooperation offer, and Personal Data collected during a possible recruitment interview, in order to participate in future recruitment processes conducted by the Administrator, i.e. Xebia sp. z o.o. with its registered office in Wrocław. choose one
  • I declare that I agree to sending to my e-mail address indicated in the content of recruitment documents, any information about recruitment processes conducted by the Administrator, i.e. Xebia sp. z o.o. with its registered office in Wrocław. choose one
  • The administrator of Personal Data is Xebia sp. z o.o. with its registered office in Wrocław, ul. Sucha 3, 50-086 Wrocław, KRS: 0000978067, NIP: 8971719181, REGON: 020363023 with a share capital of PLN 37 168 600.00. Your data contained in the CV will be processed only for recruitment purposes. The legal basis for the processing of your personal data is art. 221 cl. 1 of the Labour Code. If you provide separate consent, we will process your personal data also for future recruitment purposes. You have the right to access your personal data, to correct them, to remove them, to restrict their processing, to transfer your data, to submit an objection, to withdraw consent to data processing any time without affecting the lawfulness of processing carried out on the basis of the consent before it was withdrawn. In order to exercise the abovementioned rights, please send an e-mail with your request to: [email protected]. If you believe that your data are processed illegally, you can submit a complaint to the supervisory body with its registered office in ul. Stawki 2, Warsaw. We may only disclose your personal data if you provide consent thereto or to authorised bodies, when necessary. choose one

See also

要針對這個職缺調整履歷嗎?

目前無法檢查您與這個職缺的符合程度;請先將履歷加入個人檔案,下次即可查看。

A new version of freehire is available