NucleusTeq
|Data Engineer II
Indore, Madhya Pradesh, India
Summary
As Data Engineer II, implemented and optimized ELT pipelines and Databricks workflows to process millions of records daily, delivering critical enterprise tables for business intelligence and maintaining robust data flow.
Highlights
Engineered and deployed ELT pipelines, ingesting over 1 million records daily from diverse sources (SOAP APIs, S3 feeds, JDBC) into AWS S3 using XML and Parquet formats.
Developed and optimized Databricks workflows to transform raw data into a structured Bronze, Silver, and Gold medallion architecture within the Databricks catalog, enhancing data processing efficiency.
Consolidated and integrated data from multiple sources utilizing PySpark and SQL to create and deliver critical enterprise tables, supporting robust reporting and analytics for business stakeholders.
Proactively monitored and maintained Databricks jobs, swiftly resolving data flow issues to ensure consistent data availability and integrity.