Who We Are
Role Description
Project description
- development and optimization of a production Lakehouse platform on Databricks at enterprise scale for an international client
- working in an international team focused on large-scale data processing, process automation, and solving advanced analytical challenges using Delta Lake, Apache Spark, and Python
- design and development of data integrations, CI/CD pipelines, and performance and cost optimization for end-to-end data solutions
- option to work fully remote (for candidates based in the Czech Republic, a hybrid model with 2 days onsite in Prague is available, allowing involvement in other local projects)
Project requirements
- advanced experience with:
- Databricks and Delta Lake
- Apache Spark (development of data transformations and integrations)
- experience with:
- Python and SQL at an advanced level
- cloud platforms (Azure, AWS, or GCP)
- Git version control and CI/CD processes
- advanced knowledge of:
- English (C1/B2) for day-to-day communication in an international team
- advantage:
- optimization of parallel data processing and performance tuning on Delta Lake
- data warehousing (DWH), SCD2, CDC, Data Quality, and BI
We Expect You to Have:
Oops! Something went wrong while submitting the form.
.png)

