1 to 25 of 202 Remote Apache Spark Jobs

Data Engineer

Location
Newcastle upon Tyne, England, United Kingdom
United Kingdom A hands-on data engineering role within a large-scale cloud data programme, responsible for building, maintaining, and troubleshooting data pipelines using Apache Spark, PySpark, Apache Airflow, and a broad suite of AWS services. You will apply strong analytical and engineering skills to deliver trusted … complex, cloud-based data programme - designing, building, and maintaining data pipelines that process large volumes of data across a modern AWS-native stack. Using Apache Spark and PySpark for distributed data processing, Apache Airflow for orchestration, and a range of AWS services for storage, compute, and analytics ...

Software Engineer – Query Engines

Location
Greater London, England, United Kingdom
computations are represented, how query plans are optimized, how operators execute, and how data is read and written efficiently. We build on and extend Apache Spark, Apache DataFusion, and Apache Iceberg, bringing advances in open-source engines and table formats into the demanding environments our customers … want to advance the capabilities of modern data systems and apply that work to consequential problems. Technologies We Use Java, Scala, Rust, and Python Apache Spark, Apache DataFusion, Apache Comet, and Velox for data processing and query execution Apache Iceberg for table management and catalog ...

Staff Backend Engineer - Data Platform

Location
Greater London, England, United Kingdom
platform at scale, ideally in a consumer or marketplace environment. Deep understanding of distributed systems and modern data ecosystems — including extensive experience with Databricks, Apache Spark, Apache Kafka, Apache Flink, Apache Beam and DBT. Demonstrated success in designing, building, and operating data platforms at scale ...

(INV) Senior Manager, Databricks AI & Data Client Solutions Lead, TC UKI

Location
Greater London, England, United Kingdom
intersection of consulting, technology, innovation, and strategic alliances and you will bring deep expertise in the Databricks Data Intelligence Platform, Lakehouse architecture, Delta Lake, Apache Spark, Unity Catalog, MLflow, Mosaic AI, data engineering, machine learning, GenAI and cloud-native delivery. As a key leader within our AI & Data … translate business needs into Databricks-led solution designs, build prototypes and proof points. Shape end-to-end solutions using Databricks Intelligence Platform, Delta Lake, Apache Spark, Unity Catalog, Databricks Workflows, Delta Live Tables/Lakeflow-style pipelines, MLflow, Feature Store, Model Serving and Mosaic AI capabilities. Build, lead ...

(INV) Senior Manager, Databricks AI & Data Client Solutions Lead, TC UKI

Hiring Organisation
EY UK
Location
Manchester, UK
Employment Type
Full-time
intersection of consulting, technology, innovation, and strategic alliances and you will bring deep expertise in the Databricks Data Intelligence Platform, Lakehouse architecture, Delta Lake, Apache Spark, Unity Catalog, MLflow, Mosaic AI, data engineering, machine learning, GenAI and cloud-native delivery. As a key leader within our AI & Data … translate business needs into Databricks-led solution designs, build prototypes and proof points. • Shape end-to-end solutions using Databricks Intelligence Platform, Delta Lake, Apache Spark, Unity Catalog, Databricks Workflows, Delta Live Tables/Lakeflow-style pipelines, MLflow, Feature Store, Model Serving and Mosaic AI capabilities. Team Leadership ...

(INV) Senior Manager, Databricks AI & Data Client Solutions Lead, TC UKI

Hiring Organisation
EY UK
Location
Southwark, Greater London, UK
Employment Type
Full-time
intersection of consulting, technology, innovation, and strategic alliances and you will bring deep expertise in the Databricks Data Intelligence Platform, Lakehouse architecture, Delta Lake, Apache Spark, Unity Catalog, MLflow, Mosaic AI, data engineering, machine learning, GenAI and cloud-native delivery. As a key leader within our AI & Data … translate business needs into Databricks-led solution designs, build prototypes and proof points. • Shape end-to-end solutions using Databricks Intelligence Platform, Delta Lake, Apache Spark, Unity Catalog, Databricks Workflows, Delta Live Tables/Lakeflow-style pipelines, MLflow, Feature Store, Model Serving and Mosaic AI capabilities. Team Leadership ...

Lead Java Developer (VP)

Location
Belfast City District, Northern Ireland, United Kingdom
application components, microservices, and data integration layers to support high-throughput regulatory reporting. Big Data Processing: Develop and optimize distributed data processing jobs using Apache Spark, Hadoop, and Hive to aggregate and transform large-scale financial datasets. Real-Time Streaming: Build and maintain real-time data ingestion … streaming pipelines using Apache Kafka. Database Management: Design and optimize data models in MongoDB and relational databases, ensuring high performance for complex queries and data storage. Agile Execution: Drive sprint planning, task estimation, and daily progress within an Agile/Scrum framework to ensure timely delivery of project milestones. ...

Data Engineer

Location
Birmingham, England, United Kingdom
should have specialist experience in one or more of the following technologies**Azure Databricks*** Design and build high-performance data pipelines: Utilize Databricks and Apache Spark to extract, transform, and load data into Azure Data Lake Storage and other Azure services.* Experience of Databricks ML and Azure … develop predictive models and drive business insights.* Proven expertise in Databricks, Apache Spark, and data pipeline development and strong understanding of data warehousing concepts and practices.* Experience with Microsoft Azure cloud platform, including Azure Data Lake Storage, Databricks and Azure Data Factory.* Azure Data Engineer Associate and Databricks ...

Software Engineer - Data, Lakehouse and AI Data Platform Engineer - Analyst/Associate - London

Location
Greater London, England, United Kingdom
root‐cause analysis. Experience building or supporting production data pipelines in a collaborative engineering environment. Experience working with distributed data processing frameworks such as Apache Spark . Working knowledge of common data formats such as JSON , Avro and Parquet . For More Experienced Candidates Stronger ownership of technical … should bring relevant experience and the ability to work across comparable technologies. Examples of technologies in scope include: Data processing and logic: ANSI SQL, Apache Spark, Kafka Platforms and storage: Snowflake, Apache Iceberg, Databricks, Hadoop ecosystem technologies, Sybase IQ Engineering and deployment: CI/CD tooling, containerised ...

Senior Data Engineer

Location
Horsham, England, United Kingdom
help design and deliver cutting-edge data solutions for customers operating in some of the UK's most complex and challenging environments. Using Databricks, Apache Spark and cloud-native technologies, you’ll develop scalable Lakehouse architectures and production-grade data pipelines that enable organisations to trust, share … opportunity to grow your career while helping customers unlock the power of their data. Responsibilities Design, build and optimise scalable data pipelines using Databricks, Apache Spark, PySpark and Delta Lake. Develop and implement modern Lakehouse architectures, applying Medallion (Bronze, Silver, Gold) design principles to support trusted and reusable ...

Platform Engineer

Hiring Organisation
Gattaca
Location
Surrey, United Kingdom
Employment Type
Permanent
Salary
£75000 - £85000/annum
Responsibilities: Design, build and enhance a secure, scalable and reliable Databricks platform Develop and maintain batch and real-time data pipelines using Databricks, Apache Spark and Delta Lake Contribute to platform architecture, engineering standards and best practices across the data estate Work with Product, Analytics and Data Science … continuous improvement across the engineering team Job Requirements: Significant experience building and supporting scalable, production-grade data platforms Strong hands-on expertise with Databricks, Apache Spark, Delta Lake, Python and SQL Good understanding of cloud-based data engineering, ideally within Microsoft Azure environments Experience with Infrastructure as Code ...

Platform Engineer

Location
Greater London, England, United Kingdom
Responsibilities: Design, build and enhance a secure, scalable and reliable Databricks platform Develop and maintain batch and real-time data pipelines using Databricks, Apache Spark and Delta Lake Contribute to platform architecture, engineering standards and best practices across the data estate Work with Product, Analytics and Data Science … continuous improvement across the engineering team Job Requirements: Significant experience building and supporting scalable, production-grade data platforms Strong hands-on expertise with Databricks, Apache Spark, Delta Lake, Python and SQL Good understanding of cloud-based data engineering, ideally within Microsoft Azure environments Experience with Infrastructure as Code ...

Lead Data Engineer - Founding Member (Contract) [HR208] (UK)

Location
Greater London, England, United Kingdom
storage architectures. Strong understanding of data processing performance and optimisation. Experience with some of the following data frameworks and infrastructure technologies is highly desirable: Apache Spark, Apache Airflow, Kafka, and Elasticsearch or OpenSearch. Experience with relevant database technologies is highly desirable, including PostgreSQL, MongoDB, and vector databases ...

Sr. Software Engineer - Data Query Platform - London (Hybrid)

Location
Greater London, England, United Kingdom
frameworks, tools and applications that make that data available to other teams and systems. Your primary toolset in your work will be Java microservices, Spark/Scala data processing (also some Flink), Kubernetes and AWS native tooling. What You'll Do Write highly fault-tolerant Java code within Apache Spark to produce platform products used by our customers to query our event pipelines/ingestion for insight into active threat trends and related analytics Design, develop, and maintain ultra-high-scale data platforms that process petabytes of data Participate in technical reviews of our products and help ...

Sr. Software Engineer - Data Query Platform - London (Hybrid)

Hiring Organisation
CrowdStrike
Location
London, UK
Employment Type
Full-time
frameworks, tools and applications that make that data available to other teams and systems. Your primary toolset in your work will be Java microservices, Spark/Scala data processing (also some Flink), Kubernetes and AWS native tooling. What You'll Do: Write highly fault-tolerant Java code within Apache Spark to produce platform products used by our customers to query our event pipelines/ingestion for insight into active threat trends and related analyticsDesign, develop, and maintain ultra-high-scale data platforms that process petabytes of dataParticipate in technical reviews of our products and help us develop ...

Mid/Senior Data Engineer

Location
Greater London, England, United Kingdom
knowledge of optimisation techniques for large‐scale data processing Strong proficiency in SQL and Python for handling complex data problems Hands‐on experience with Apache Spark (PySpark or Spark SQL) Experience with the Azure data stack Knowledge of workflow orchestration tools like Azure Data Factory or Apache … cost optimisation strategies for cloud data platforms Experience with data quality frameworks and implementation Experience with data visualisation tools like Power BI or Apache Superset Experience with other cloud data platforms like AWS, GCP or Oracle Experience with modern unified data platforms like Databricks or Microsoft Fabric Experience with ...

Principal Data Engineer

Hiring Organisation
Raytheon
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
Experience of working in an Agile delivery team Strong analytic skills related to working with unstructured datasets Python (PySpark, Pandas, PyArrow) Distributed data processing (Apache Spark) Data ETL (Apache Airflow, AWS Step Functions, Apache NiFi) Cloud services (AWS, Azure or GCP) Messaging/Streaming (Kafka ...

Senior Data Engineer

Location
Greater London, England, United Kingdom
Experience of working in an Agile delivery team Strong analytic skills related to working with unstructured datasets Python (PySpark, Pandas, PyArrow) Distributed data processing (Apache Spark) Data ETL (Apache Airflow, AWS Step Functions, Apache NiFi) Cloud services (AWS, Azure or GCP) Messaging/Streaming (Kafka ...

Principal Data Engineer

Location
Greater London, England, United Kingdom
Experience of working in an Agile delivery team Strong analytic skills related to working with unstructured datasets Python (PySpark, Pandas, PyArrow) Distributed data processing (Apache Spark) Data ETL (Apache Airflow, AWS Step Functions, Apache NiFi) Cloud services (AWS, Azure or GCP) Messaging/Streaming (Kafka ...

Data Technical Lead

Hiring Organisation
PA Consulting
Location
London, UK
Employment Type
Full-time
Modern data platforms: Databricks, Snowflake and cloud-native data services Data architecture: Lakehouse, medallion, data mesh and domain-oriented patterns Data processing: Apache Spark, SQL and distributed processing technologies Data storage: Delta Lake, Apache Iceberg and open table formats Data pipelines: Enterprise-scale ETL/ELT, batch … streaming patterns Transformation and quality: dbt, Great Expectations and equivalent tooling Orchestration: Apache Airflow and cloud-native orchestration services Event streaming: Kafka and cloud-native event-streaming technologies Data modelling: Dimensional, analytical and domain-oriented modelling Data governance: Catalogue, metadata, lineage, quality, security, privacy and access controls Data products ...

Data Architect - W2 Only

Hiring Organisation
Enormous Enterprise LLC
Location
Jersey City, New Jersey, United States
Employment Type
Any
Salary
USD Annual
enterprise cloud strategies. Evaluate and recommend technologies for data storage, processing, streaming, governance, and observability. Provide technical leadership and mentorship to engineering teams implementing Spark, Kafka, Databricks, and cloud-native services. Drive performance optimization, cost management, and operational excellence across the data platform. Support executive stakeholders with migration roadmaps … distributed systems. Demonstrated experience leading enterprise-scale cloud migration initiatives involving hundreds of terabytes or petabytes of data. Strong experience with: o Databricks o Apache Spark o Apache Kafka o AMPS, MQ, Hadoop, or experience with other enterprise messaging systems o AWS cloud services o Large-scale ...

Senior Data Engineer

Hiring Organisation
GCS
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£600 - £850/day Inside IR35
Senior Data Engineer Role (Python/Spark/Databricks) - Hybrid - Contract - Banking Role - Senior Data Engineer Rate - £850 p/d (Inside IR35) Duration - 6 months with very likely extension Location - Hybrid/Liverpool Street (London) - 3 days per week in a Liverpool Street office About the Role Global … delivering critical data platforms and analytics capabilities across the business. This is a genuinely hands-on engineering role, requiring strong expertise in Python, Databricks, Spark/PySpark and SQL. The successful candidate will be actively involved in the design, development, optimisation and support of enterprise-scale data platforms, working ...

Data Engineer - Databricks

Location
Bristol, England, United Kingdom
associated Azure technologies. Build scalable ETL/ELT pipelines to ingest, transform, and load data from multiple enterprise systems. Develop data transformations using PySpark, Spark SQL, SQL, and Python. Create and maintain curated datasets that support business intelligence, reporting, and analytics. Implement data quality, reconciliation, validation, and monitoring processes. … engineering or business intelligence solutions. Strong proven commercial experience with Databricks. Strong experience developing ETL/ELT pipelines and data integration solutions. Experience using Apache Spark, PySpark, and Spark SQL. Strong SQL development and optimisation skills. Experience working with large and complex datasets from multiple source systems. ...

Data Engineer

Hiring Organisation
Hackajob Ltd
Location
Newcastle Upon Tyne, Tyne and Wear, North East, United Kingdom
Employment Type
Permanent, Work From Home
engineers where needed. Core Data Engineering Strong programming proficiency in Java (preferred) or Python. Hands-on experience with at least one of: Kafka, Flink, Spark (Flink/Kafka preferred for streaming). Solid understanding of stream processing concepts (e.g., event time, state, backpressure). Understanding of software engineering best ...

Business Analyst

Hiring Organisation
McGregor Boyall Associates Limited
Location
Belfast, County Antrim, Northern Ireland, United Kingdom
Employment Type
Contract
capital markets. Exposure to Python for more complex data extraction and analysis. Familiarity with the Hadoop and big data ecosystem, including Hive, Impala, HDFS, Apache Spark, Spark-SQL, UDF and Sqoop. Experience with Tableau and reporting best practices. Previous project management or project delivery experience. Exposure ...