A Data Engineer with 2+ years of experience specializing in designing, building, and maintaining scalable ETL/ELT data pipelines across cloud and big data ecosystems. The role involves working with large-scale distributed systems like Hadoop and Apache Spark, optimizing data warehouses in Snowflake, implementing rigorous data modeling strategies, and orchestrating complex workflows using Apache Airflow. The candidate is expected to be proficient in Python, PySpark, Advanced SQL, and shell scripting, experienced in managing structured and semi-structured formats within cloud platforms (AWS, Azure, or GCP), and capable of ensuring data quality, performance optimization, and seamless version control using Git and CI/CD frameworks.