Design, develop, and maintain scalable batch and streaming data pipelines using Apache Spark, cloud-native technologies, and modern data platforms Build and optimize ETL/ELT frameworks to ingest, transform, validate, and curate data from diverse enterprise source systems