Benefits
Design, develop, and maintain ETL workflows and data pipelines for large-scale data processing.Write and optimize medium-complexity SQL queries on Hive, Spark, and Impala for use case-driven scenarios.Develop and maintain scripts using HIVE, PySpark, Spark, and Scala for data transformation and processing.Implement and troubleshoot Unix Shell scripts for automation and system tasks.Collaborate with cross-functional teams to ensure data integrity and performance.Work in alignment with U.S. business hours for meetings and deliverables.Contribute to cloud-based data solutions and ensure best practices in data engineering.Strong SQL skills with Hive, Spark, and Impala.Proficiency in HIVE / PySpark / Spark / Scala scripting.Good understanding of Unix Shell scripting.Conceptual knowledge of cloud ecosystems (AWS preferred).