Data Engineer
You will build and operate production-grade batch and real-time data pipelines on GCP. You will implement data quality controls and observability, maintain connectors and integrations, establish engineering standards, support feature engineering and machine-learning data pipelines, and promote reliable, governed data delivery.
Responsibilities
- Design and build batch and real-time data pipelines on GCP
- Build Kafka-based event streams
- Create BigQuery transformations
- Orchestrate workflows with Airflow
- Define and implement data contracts and quality checks
- Instrument pipeline observability
- Maintain in-house connectors and third-party integrations
- Ensure ingestion resilience and monitoring
- Contribute to code reviews and engineering standards
- Develop dbt modelling patterns
- Maintain CI/CD pipeline hygiene
- Support change control for shared datasets
- Build feature engineering and model data pipelines
- Enable lineage, reproducibility, and freshness for machine-learning workloads
Requirements
- Hands-on experience building and operating production-grade data pipelines on GCP
- BigQuery
- Cloud Composer or Airflow
- Google Cloud Storage
- Cloud Run
- SQL
- Python
- Apache Kafka or equivalent streaming technologies
- Event-driven ingestion patterns
- dbt
- Medallion architecture patterns
- Data contracts
- Data lineage
- Observability tooling
- Pipeline SLO monitoring
- Data governance in regulated financial services
- Access control
- Audit trails
- Change control
Benefits
- Hybrid working model with 3 days in the office
- Home office equipment reimbursement
- Performance-related bonus
- Private medical cover for the employee and family or partner
- Multikafeteria system with multisport cards and vouchers
- Life insurance
- Employee share plans
- Well-being events
- Employee Assistance Programme
- Summer picnic and New Year party
- Three additional days off annually for a birthday and voluntary work
- App-based parking spot booking
- Stretching sessions