Principal Data Engineer
- Location
- Greater London, England, United Kingdom
platforms capable of processing billions of events per day and managing terabytes scale datasets. Develop modern Data Lakehouse platforms leveraging S3-compatible object storage, Apache Iceberg, Spark, and Trino. Build real-time streaming integration using Apache Kafka and Apache Flink for low-latency event processing, enrichment … OpenShift, Terraform, CI/CD, GitOps, and Infrastructure as Code. Lead metadata, lineage, and data discovery capabilities using technologies such as DataHub, OpenMetadata, or Apache Atlas to improve platform transparency and usability. Ensure platform security and regulatory compliance through fine-grained access controls, encryption, secrets management, auditing, and secure ...