12 Sep
|
Cloudxtreme
|
Pune
Role & responsibilities
Qualifications
Bachelor's degree in computer science, Information Technology, or a related quantitative field. Master's degree preferred.
8-10+ years of hands-on experience in data modeling, database design, and data architecture, with a primary focus on Hadoop/Hive ecosystems and big data technologies.
Expert-level proficiency in HiveQL and deep understanding of Hive architecture.
Solid hands-on experience with Apache Hadoop (HDFS, YARN) and related components.
Mandatory proficiency in PySpark for data processing, transformation, and optimization.
Demonstrable experience in designing and implementing dimensional models, 3NF, Data Vault, or other relevant data modeling techniques for large-scale data warehouses/lakes.
Experience with other big data technologies such as Kafka,
Spark Streaming, HBase, Impala, or Presto is highly desirable.
Strong understanding of data partitioning, bucketing, and file formats (Parquet, ORC, Avro) within Hadoop for performance optimization.
Familiarity with cloud-based big data platforms (e.g., AWS EMR, Azure HDInsight, Google Cloud Dataproc) is a significant plus.
Excellent analytical, problem-solving, and communication skills, with the ability to articulate complex technical concepts to non-technical stakeholders.
Proven ability to lead technical initiatives, manage complex projects, and drive architectural decisions.
📌 Data Architect- Hadoop (9+ Years) (Pune)
🏢 Cloudxtreme
📍 Pune