GG
GSSTech Group
Data Engineer

Data Engineer - ETL/PySpark (Banking Domain)

On-siteSeniorData Engineerposted 1mo ago
✦Role summaryAI-generated

Design and maintain PySpark‑based ETL pipelines and data marts for a banking environment, handling the full SDLC from development through production support. Requires 5+ years of experience with Python, Oracle, CI/CD, and data engineering best practices.

Skills required

About this role

Role Summary

We are looking for a hands-on Data Engineer with strong PySpark and Python expertise to build and maintain data marts and ETL pipelines within a banking environment. The ideal candidate owns the full SDLC, from build through UAT, bug fixing, production deployment, and postproduction support.

Key Responsibilities

  • Design, build, and maintain ETL pipelines and data marts using PySpark and Python
  • Write clean, maintainable, and robust production-grade code
  • Own end to end SDLC activities: build, UAT, UAT bug fixes, production deployment, and postproduction support
  • Perform Oracle query analysis and PySpark code debugging
  • Work across structured, semi structured, and unstructured data sources
  • Apply software engineering best practices to production pipelines
  • Support CI/CD processes and data testing/validation activities

Required Skills & Experience

  • 5+ years commercial experience in a data-driven role
  • Hands-on experience building data marts and ETL pipelines
  • Expert level Python for ETL scripting
  • Strong PySpark experience
  • Analytical expertise in Oracle SQL and data analysis
  • Understanding of software engineering concepts and best practices for production pipelines
  • Banking client or banking domain knowledge
  • Strong Data Warehousing fundamentals
  • Familiarity with query languages and both SQL and NoSQL database technologies
  • CI/CD exposure, including testing and validation of data pipelines

Daily Tech Stack

  • Python
  • Spark / PySpark
  • Jupyter
  • SQL and NoSQL DBMS
  • Hadoop / MapReduce / Hive
  • Pandas
✕ position closed

This role is no longer accepting applications. It’s kept here for reference — check out the similar open roles below.

Similar open roles

TF
NEW

Senior Staff Software Engineer – Data Platform

Thermo Fisher Scientific·Remote, Indiana, United States of America; Remote, Arizona, United States of America; Remote, Colorado, United States of America; Remote, Georgia, United States of America; Remote, Idaho, United States of America; Remote, Kansas, United States of America; Remote, Michigan, United States of America; Remote, Minnesota, United States of America; Remote, Nebraska, United States of America; Remote, Nevada, United States of America; Remote, New Mexico, United States of America; Remote, North Carolina, United States of America; Remote, Ohio, United States of America; Remote, Oregon, United States of America; Remote, Pennsylvania, United States of America; Remote, South Carolina, United States of America; Remote, Tennessee, United States of America; Remote, Texas, United States of America; Remote, Utah, United States of America; Remote, Wisconsin, United States of America
RemoteStaffData Engineer
$136k – $204k USD
yesterday
JC
NEW

Sr Lead Software Engineer - Data Engineering

JPMorgan Chase·Wilmington, DE, United States
On-siteSeniorData Engineer
yesterday
JC
NEW

Lead Software Engineer- Data Engineer/Pyspark/Databricks

JPMorgan Chase·Houston, TX, United States
On-siteSeniorData Engineer
yesterday
Data Engineer - ETL/PySpark (Banki… at GSSTech Group — India