Job Summary
Experience: 4+ YearsAs a Big Data Engineer, will be responsible for designing, developing, and optimizing big data solutions using Apache Hive within the Hadoop ecosystem. You will work with large datasets, write complex HiveQL queries, and collaborate with data engineers, analysts, and scientists to ensure efficient data processing and accessibility.ResponsibilitiesDesign and implement data processing solutions using Hive and Hadoop tools.Write and optimize HiveQL queries for data extraction, transformation, and analysis.Develop ETL pipelines and data ingestion workflows.Implement partitioning, bucketing, and indexing strategies for performance tuning.Collaborate with cross-functional teams to understand data requirements.Monitor and troubleshoot Hive jobs and Hadoop cluster performance.Maintain documentation for data models, workflows, and best practices.Ensure data quality, security, and compliance standards are met.Proven experience with Hive and Hadoop ecosystem.Strong proficiency in SQL and HiveQL.Experience with data modelling and schema design.Familiarity with tools like Spark, Pig, Oozie, and Flume is a plus.Programming skills in Java, Python, or Scala.Knowledge of data warehousing and ETL concepts.Good understanding of performance tuning and optimization techniques.Cloud computing knowledge is a plus.Bachelor’s degree in Computer Science, Engineering, or related field