1 to 25 of 59 Apache Hive Jobs

Data Lineage & Governance Analyst

Location
Greater London, England, United Kingdom
provenance. Hands‐on experience with data lineage/metadata tooling in enterprise environments (e.g., Collibra, Alation, Informatica EDC/IDMC, IBM Infosphere, Microsoft Purview, Apache Atlas, Amundsen, DataHub or similar). Proven ability to build lineage for complex platforms: data lakes, warehouses, marts, and distributed processing (Spark‐based pipelines … patterns, data modeling (dimensional + normalized), and batch scheduling dependencies. Ability to interpret data transformation logic from pipelines (Spark SQL/PySpark/Hive queries/orchestration configs). Strong documentation capability: source‐to‐target mappings, lineage diagrams, data dictionaries, metadata standards, and control evidence packs. Technical Skills Strong ...

“Techno-Functional” Analyst– Capital Markets Data Transformation

Location
Greater London, England, United Kingdom
pipelines and data quality frameworks (rules, thresholds, exception handling). Working knowledge of scheduling/orchestration tools such as Autosys and/or Apache Airflow (monitoring schedules, reruns, failure triage). Experience with CI/CD and release controls (Git, Harness, UrbanCode Deploy (UCD), Red Hat OpenShift or equivalent … such as Tableau dashboards (data validation, extract refresh checks, reconciliation to source). Nice-to-Have Experience with tools such as PySpark, Spark SQL, Hive, Impala, HDFS, Parquet, and Oracle databases. Exposure to data governance concepts: critical data elements (CDEs), lineage, data quality dimensions, audit frameworks. Experience running delivery ...

Systems Engineer - Data Engineering

Hiring Organisation
Hackajob Ltd
Location
Edinburgh, Midlothian, Scotland, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
/CD) pipelines, containerisation, and workflow orchestration. Familiar with ETL/ELT frameworks, and experienced with Big Data Processing Tools (e.g. Spark, Airflow, Hive, etc.) Knowledge of programming languages (e.g. Java, Python, SQL) Hands-on experience with SQL/NoSQL database design. Degree in STEM, or similar field ...

Systems Engineer - Data Engineering

Hiring Organisation
Leonardo DRS
Location
Edinburgh, UK
Employment Type
Full-time
/CD) pipelines, containerisation, and workflow orchestration. Familiar with ETL/ELT frameworks, and experienced with Big Data Processing Tools (e.g. Spark, Airflow, Hive, etc.)Knowledge of programming languages (e.g. Java, Python, SQL)Hands-on experience with SQL/NoSQL database design. Degree in STEM, or similar field ...

Systems Engineer - Data Engineering

Location
City of Edinburgh, Scotland, United Kingdom
/CD) pipelines, containerisation, and workflow orchestration. Familiar with ETL/ELT frameworks, and experienced with Big Data Processing Tools (e.g. Spark, Airflow, Hive, etc.) Knowledge of programming languages (e.g. Java, Python, SQL) Hands-on experience with SQL/NoSQL database design. Degree in STEM, or similar field ...

Engineering Manager - Data Platform (Hybrid, GBR)

Hiring Organisation
CrowdStrike
Location
London, UK
Employment Type
Full-time
software development lifecycleSolid background in Java/Scala and a scripting language like Python. Experience building large scale data pipelinesStrong familiarity with the Apache Hadoop ecosystem including : Spark, Kafka, Flink, Iceberg/Delta Lake/Hive, Apache Presto/Trino, etc. Experience with relational SQL and NoSQL ...

Technical Analyst – Data Governance, Controls & Traceability

Location
Greater London, England, United Kingdom
. Strong proficiency in Python for data engineering and automation. Hands‐on experience with PySpark and Spark SQL in production environments. Solid knowledge of Hive, Impala, HDFS, and Parquet. Advanced SQL skills; experience with Oracle databases. Experience designing and supporting ETL/ELT pipelines and data quality frameworks (rules … thresholds, exception handling). Working knowledge of Autosys and Apache Airflow (monitoring schedules, reruns, failure triage). Experience with CI/CD and release controls (Git, Harness, UrbanCode Deploy (UCD), Red Hat OpenShift or equivalent). Familiarity with AWS S3 for large‐scale data storage and dataset movement patterns. ...

Banking Data Quality Analyst

Location
Greater London, England, United Kingdom
validate transformations and trace data across platforms. Tooling/Delivery Methods Working knowledge of scheduling/orchestration tools such as Autosys and/or Apache Airflow (monitoring schedules, reruns, failure triage). Experience with CI/CD and release controls (Git, Harness, UrbanCode Deploy (UCD), Red Hat OpenShift … such as Tableau dashboards (data validation, extract refresh checks, reconciliation to source). Nice-to-Have Experience with tools such as PySpark, Spark SQL, Hive, Impala, HDFS, Parquet, and Oracle databases. Hands-on exposure to DCRM tooling and operational exception management processes. Experience with governance/catalog tools ...

data engineer for customer experience

Location
Greater London, England, United Kingdom
including complex joins, window functions, aggregation, and query optimization; Experience with a major data warehouse or big data technology such as BigQuery, Snowflake, Redshift, Hive, Spark, Vertica, or StarRocks; Experience with ETL/ELT tools or orchestration frameworks such as Airflow, dbt, or internal frameworks; Proficiency in at least ...

Senior Azure Data Engineer (Databricks)

Hiring Organisation
Capco
Location
London, UK
Employment Type
Full-time
improve productivity while maintaining appropriate human oversight, governance, and solution quality. Bonus Points ForDevelopment experience using Scala or Java. Familiarity with Cloudera, Hadoop, Hive, Spark, and related distributed data ecosystems. Experience working with sensitive data and applying data privacy regulations, including GDPR.Curiosity to explore new technologies, including responsible ...

Senior Data Engineer - AWS

Hiring Organisation
Capco
Location
London, UK
Employment Type
Full-time
Bonus Points ForExpertise in data modelling, schema design, and working with structured and semi-structured datasets. Experience with distributed technologies including Hadoop, Spark, HDFS, Hive, Databricks, or AWS Lake Formation. Knowledge of automating cloud-based ingestion and transformation frameworks. Experience delivering data engineering solutions within highly regulated industries such ...

Principal Data Engineer - AWS

Hiring Organisation
Capco
Location
London, UK
Employment Type
Full-time
Bonus Points ForExpertise in data modelling, schema design, and working with structured and semi-structured datasets. Experience with distributed technologies including Hadoop, Spark, HDFS, Hive, Databricks, or AWS Lake Formation. Knowledge of automating ingestion and transformation frameworks within cloud-native data platforms. Experience delivering solutions within highly regulated industries ...

AI Strategist, Generative AI Innovation Center

Hiring Organisation
AmazonWebServices
Location
London, UK
Employment Type
Full-time
professional or military settings- Bachelor's degree, or experience as technical specialist in design and architecture- Experience in Database like NoSQL, or experience in Hive/Spark/Hbase/Yarn and experience in any Bigdata architecture- Bachelor's degree in business administration, finance, economics, computer science, data science ...

Business Systems Analyst

Location
Greater London, England, United Kingdom
future needs of the department. This is an exciting opportunity to work on cutting edge technologies including Big Data and NoSQL databases (Hadoop, HBase, Hive, Spark and MongoDB) to allow the business to gain advanced insight into their portfolios and valuation metrics. Key responsibilities include:* Lead business requirements elicitation ...

Technical Banking Analyst

Location
Greater London, England, United Kingdom
programs. Strong proficiency in Python for data engineering and automation. Hands‐on experience with PySpark and Spark SQL in production environments. Solid knowledge of Hive, Impala, HDFS, and Parquet . Advanced SQL skills; experience with Oracle databases . Experience designing and supporting ETL/ELT pipelines and data quality … frameworks . Working knowledge of Autosys& Apache Airflow Experience with CI/CD tools (Git, Harness, UrbanCode Deploy (UCD), Red Hat OpenShift) Familiarity with AWS S3 for large-scale data storage. Experience supporting Tableau dashboards. Nice-to-Have Experience in regulated or enterprise data environments (banking, compliance). Exposure ...

Senior Data Engineer - AWS

Location
Glasgow, Scotland, United Kingdom
Expertise in Data Modelling, schema design, and handling both structured and semi-structured data. Familiarity with distributed systems such as Hadoop, Spark, HDFS, Hive, Databricks. Exposure to AWS Lake Formation and automation of ingestion and transformation layers. Background in delivering solutions for highly regulated industries. Passion for mentoring ...

Solution Architect

Location
Greater London, England, United Kingdom
software vendor, SaaS provider, or professional services organisation. Familiarity with Ataccama or similar enterprise data platforms. Exposure to Hadoop/Spark, NoSQL, XSLT, Hive, Kafka, Snowflake, or Databricks. Python or other programming languages. Familiarity with SDLC, code management and upgrades, and Unix/Linux. Perks & Benefits Long-Term Incentive ...

Data Scientist II - QuantumBlack, AI by McKinsey

Hiring Organisation
McKinsey & Company
Location
London, UK
Employment Type
Full-time
setsProgramming experience (focus on machine learning): SQL and Python's Data Science stack; good knowledge of at least one big data framework (e.g., PySpark, Hive, Hadoop) is a plus. R, SPSS, and SAS are considered nice-to-haveStrong understanding of machine learning methods and experience applying them to complex ...

Data & AI Science Consultant

Location
Greater London, England, United Kingdom
languages eg. Python, R, Scala, etc.; (Python preferred) Proficiency in database technologies eg. SQL, ETL, No‐SQL, DW, and Big Data technologies e.g. pySpark, Hive, etc. Experienced working with structured and also unstructured data eg. Text, PDFs, jpgs, call recordings, video, etc. Knowledge of machine learning modelling techniques ...

Quantitative Engineer

Location
Orpington, England, United Kingdom
automation Strong Python development skills (including Pandas and related data-processing libraries). Experience with big data technologies such as Spark, PySpark, Hadoop, and Hive Exposure to quantitative modelling or financial modelling is a plus but not required. Skills that will help: Global Risk Management experience Benefits of working ...

Data Science Manager

Location
London, United Kingdom
eg.?Python, R, Scala, etc.(Python preferred) Strong proficiency in database technologies eg.?SQL, ETL, No-SQL, DW, and Big Data technologies eg.?pySpark, Hive, etc. Experienced working with structured and also unstructured data eg.?Text, PDFs, jpgs, call recordings, video,?etc. Knowledge of machine learning modelling techniques ...

Data Science Manager

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
eg.?Python, R, Scala, etc.(Python preferred) Strong proficiency in database technologies eg.?SQL, ETL, No-SQL, DW, and Big Data technologies eg.?pySpark, Hive, etc. Experienced working with structured and also unstructured data eg.?Text, PDFs, jpgs, call recordings, video,?etc. Knowledge of machine learning modelling techniques ...

Senior Data Engineer, Python, Spark

Location
Manchester, England, United Kingdom
scripting language — Python is required . Proficiency in at least one object-oriented language. Experience with big data technologies such as HDFS, YARN, MapReduce, Hive, Kafka, Spark, Airflow, or Presto. Experience with AWS, GCP, or Looker (advantageous but not essential). Solid background in data modelling, including the design ...

Senior Data Engineer, Python, Spark

Location
Cambridge, England, United Kingdom
scripting language — Python is required . Proficiency in at least one object-oriented language. Experience with big data technologies such as HDFS, YARN, MapReduce, Hive, Kafka, Spark, Airflow, or Presto. Experience with AWS, GCP, or Looker (advantageous but not essential). Solid background in data modelling, including the design ...

Fullstack Python Engineer

Location
Greater London, England, United Kingdom
/Azure also relevant). Containerisation & orchestration: Kubernetes, Docker. Strong CICD exposure: GitHub, Terraform, automated pipelines. Familiarity with SQL/NoSQL databases (e.g. Bigtable, Hive). Experience working in Agile environments (Scrum, Kanban, or stream-aligned team models). Nice to Have ✔️ Experience with predictive analytics/big data ...