12 Aug
|
Data-Virtuo
|
Hyderabad
12 Aug
Data-Virtuo
Hyderabad
We are seeking a skilled Data Engineer to design, build, and maintain scalable data pipelines and platforms. In this role, you will work on processing large volumes of data using Python and PySpark, integrating data from various sources, and ensuring high data quality and performance across our big data ecosystem.
Key Responsibilities:
- Design, develop, and optimize ETL/ELT pipelines using Python and PySpark.
- Build and maintain scalable big data platforms (e.g., Hadoop, Spark, Kafka).
- Integrate data from various structured and unstructured sources.
- Ensure data quality, integrity, and performance in data processing.
- Collaborate with data scientists, analysts, and other stakeholders to understand data requirements.
- Monitor system performance and troubleshoot issues in production.
- Implement best practices in coding, testing, and documentation.
Requirements
Requirements:
- Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field.
- 4+ years of experience in data engineering or a similar role.
- Proficiency in Python programming.
- Hands-on experience with PySpark and distributed data processing.
- Familiarity with big data frameworks and tools (e.g., Hadoop, Kafka, Hive).
- Solid understanding of database concepts (SQL and NoSQL).
- Experience with cloud platforms (e.g., AWS, Azure, GCP) is a plus.
- Robust problem-solving and communication skills.
📌 Cloud Data Engineer (Hyderabad)
🏢 Data-Virtuo
📍 Hyderabad