Jithesh Gopinathan

Senior Data Engineer, UK Civil Service · Spark, Python, Cloud

Leeds, United Kingdom 14+ years in data engineering

Jithesh Gopinathan, in a grey suit and open-collar white shirt, looking directly at the camera

Background

I’m a data engineer with 14+ years building and operating large-scale data platforms, most recently as a Senior Data Engineer with the UK Civil Service in Leeds. My background spans Java and big data development — designing ETL and streaming pipelines, building Hadoop and Spark ecosystems from the ground up, and working across AWS, Azure, Databricks, and Snowflake.

As a consultant with BJSS/CGI, I delivered data platforms for clients across aviation, energy, e-commerce, utilities, and public health: migrating data capabilities onto a modern, decentralised architecture for a major UK airline; building data ingestion and billing pipelines for a UK water services provider; turning a national energy-sector organisation’s usage data into information other teams could use; and managing national vaccination data for a UK public health body. Earlier, I built a big data ecosystem from a vanilla Hadoop installation and worked closely with data scientists across the full machine learning pipeline.

I hold the AWS Certified Data Analytics – Specialty and AWS Certified Developer – Associate certifications, along with a Machine Learning Nanodegree from Udacity. I’m currently exploring how agentic AI tools fit into data engineering practice, and I welcome conversations about scalable, decentralised data architecture.

Where I’ve worked

UK Civil Service

– Present · Leeds, United Kingdom

Senior Data Engineer

– Present

Azure Fabric Delta Lake Informatica

BJSS/CGI

· Leeds, United Kingdom · Remote

Senior Data Engineer

Client: a UK water services provider

Built the client’s data ingestion and billing pipelines on Azure.

  • Built ingestion pipelines drawing customer, usage, and market data from multiple channels into a Delta Lake landing layer using Azure Data Factory.
  • Built the billing data pipelines in Azure Data Factory, using ingested data to generate customer billing.
Azure Data Factory Delta Lake Azure Pipelines Azure Functions

Senior Data Engineer

Client: a major UK airline

Led the initial migration of the client’s data capabilities onto a data mesh architecture.

  • Led the initial phase of moving data capabilities onto a modern, decentralised data mesh architecture, improving data accessibility across the organisation.
Python Spark Databricks Delta Lake Azure Fabric Azure Pipelines

Lead Data Engineer

Client: a UK e-commerce retailer

Led the data engineering workstream collecting and loading the client’s business data into Delta tables.

  • Built real-time and batch ingestion from multiple disjoint data endpoints into a Delta Lake landing layer on Azure Fabric.
  • Worked directly with business analysts and stakeholders to translate requirements into pipeline design, and unblocked cross-team dependencies within the sprint.
Python Spark Databricks Delta Lake Azure Fabric Azure Pipelines

Senior Data Engineer

Client: a UK energy-sector organisation

Built the pipelines that transformed the client’s national energy usage data for downstream teams.

  • Applied complex business logic across large national energy datasets and stored the results in Delta tables for efficient downstream access.
  • Collected real-time data from multiple energy data endpoints and transformed it into the formats each downstream team needed.
Python Spark Databricks Delta Lake Azure Azure Pipelines

Data Engineer

Client: a UK public health body

Ran ETL processes and maintained the data warehouse behind the client’s national vaccination data.

  • Ran ETL processes over the national vaccination dataset, including SNOMED code updates, and maintained the data warehouse behind it.
Python Spark Databricks Jenkins Git

The Oakland Group

· Leeds, United Kingdom

Data Engineer

Client: a major UK radio broadcaster

Built the Customer Data Platform pipelines for the client’s entire user base.

  • Built the streaming and batch ingestion pipelines feeding mParticle, transforming data into the CDP’s required format across the full user base.
Java Python Spark Airflow dbt Snowflake AWS EC2 AWS EMR AWS Kinesis AWS Lambda GCP BigQuery GCP Dataproc GCP Data Fusion

ST Engineering

· Singapore

Assistant Principal Engineer

Built a big data and machine learning platform from the ground up.

  • Built the entire Hadoop ecosystem from a bare installation, standing up the big data infrastructure the platform ran on.
  • Worked closely with data scientists across the full ML lifecycle — data collection, transformation, and model pipelines — and automated model deployment.
Java Python Hadoop Spark Hive Airflow MLflow AWS EC2 AWS EMR AWS RDS AWS Kinesis HBase Kafka

Crédit Agricole CIB

· Singapore

Big Data Developer

Built the streaming integration layer for a payment system processing millions of transactions.

  • Built a Kafka-to-Kafka streaming application, processing payment events through Spark and writing results back for downstream systems handling millions of transactions.
Hadoop Spark Yarn HBase Kafka Maven

Java Developer

Developed backend functionality for the bank’s KYC application, supporting Investment Bank counterparts.

  • Implemented change requests from business analysts on the KYC data-collection application, working directly with the Paris team.
Java J2EE JSP Spring Hibernate jQuery MS SQL

Allianz Cornhill Information Services

· Trivandrum, India

Senior Software Engineer

Built an online insurance application for Australian home, motor, and landlord customers.

  • Delivered well-tested, production-ready code across the full development lifecycle as an offshore engineer working directly with Australian clients.
Java J2EE JSP Spring Boot Hibernate Angular JMS DB2

D+H Solutions India

· Trivandrum, India

Software Engineer

Built account-opening and lending tools for US banking clients.

  • Designed and built the Newday configuration tool for D+H’s uOpen platform, cutting a 5–6 day bank-configuration process down to one day.
  • Received the company’s ‘Go Extra Mile’ award, March 2014.
Java J2EE JSP JavaScript jQuery AJAX Spring Hibernate MS SQL

Tools and technologies

Languages & runtimes

Java Python

Frameworks & platforms

Spark Hadoop Databricks Airflow MLflow dbt Spring Spring Boot Hibernate J2EE Angular

Cloud & infrastructure

AWS EC2 AWS EMR AWS RDS AWS Kinesis AWS Lambda Azure Data Factory Azure Fabric Azure Pipelines Azure Functions GCP BigQuery GCP Dataproc GCP Data Fusion Docker Kafka IBM MQ

Data & integration

Delta Lake Snowflake Redshift HBase Hive ETL/ELT pipelines Data mesh Data fabric

Delivery & collaboration

Cross-functional stakeholder delivery Offshore/onshore delivery Agile delivery Code review & mentoring Agentic AI tooling

Get in touch

The fastest way to reach me is email.