DEEPESH KUMAR

DataEngineer|BigData Engineer| Bigdata Analyst| Bigdata Developer | Works at EY |HDFS|Sqoop|Hive|MySQL|ShellScripting|Python|Pyspark|ScalaSpark|SparkSQL|AWS|S3|Glue|Lambda|Redshift|EMR|SNS|Snowflake

Gurugram, Haryana, India

About

Possessing strong hands-on experience in Big Data technologies, I bring in-depth knowledge of Hadoop and its internals, with proven expertise in handling complex data processing and distributed systems. ✍️ My expertise includes data ingestion and integration using tools like Sqoop, enabling efficient and reliable data movement across systems. ✍️ Adept in working with data warehouses such as Hive and Snowflake, along with query engines like Impala and tools like Hue, I have a solid understanding of scalable data storage and optimized data retrieval. ✍️ I have strong proficiency in Spark with Scala, working extensively with DataFrames, Spark SQL, and performance optimization techniques within distributed environments. ✍️ Experienced in working with core AWS services including EMR, S3, Lambda, SNS, Glue, and Redshift, I have built and deployed scalable data pipelines and cloud-based data solutions. ✍️ Proficient in Python, SQL (MySQL), and Linux, enabling end-to-end development, data processing, and system-level operations. ✍️ I actively engage in discussions around Data Engineering, Big Data, and SQL, contributing insights and staying aligned with industry trends. ✍️ My professional journey includes designing and implementing robust data solutions across diverse projects. I specialize in building scalable pipelines, optimizing performance, and troubleshooting complex data workflows to deliver reliable and efficient solutions. Open to opportunities in Big Data Engineering, Data Engineering, and Cloud Data Platforms.

Experience

  • Big Data engineer at EY
    Dec 2021 - Present · 4 yrs 8 mos