Location: Mumbai / Hyderabad / Bengaluru / Pune (Hybrid/Onsite)
Employment Type: Full-Time
Job Summary
We are seeking an experienced GCP Data Engineer with 5+ years of hands-on experience in designing, developing, and maintaining scalable data pipelines and modern data platforms on Google Cloud Platform (GCP) . The ideal candidate should have expertise in GCP data services, ETL/ELT development, cloud data warehousing, and SQL/Python programming. Experience with big data technologies and CI/CD practices will be an added advantage.
Key Responsibilities
- Design, develop, and maintain scalable ETL/ELT pipelines on Google Cloud Platform.
- Build and optimize batch and real-time data ingestion pipelines.
- Develop and manage cloud-native data solutions using GCP services.
- Design data models and implement data warehouse solutions using BigQuery.
- Perform data cleansing, transformation, validation, and optimization.
- Integrate data from multiple on-premise and cloud-based data sources.
- Ensure data quality, governance, security, and compliance standards are met.
- Monitor and troubleshoot production data pipelines.
- Optimize SQL queries and improve overall pipeline performance.
- Collaborate with Data Architects, Data Scientists, Analysts, and business stakeholders.
- Implement CI/CD pipelines and Infrastructure as Code where applicable.
- Participate in code reviews and follow Agile development methodologies.
Mandatory Skills
- 5+ years of experience as a Data Engineer.
- Strong hands-on experience with Google Cloud Platform (GCP) .
- Expertise in BigQuery .
- Experience with Cloud Storage (GCS) .
- Hands-on experience with Cloud Composer (Apache Airflow) .
- Experience with Dataflow (Apache Beam) .
- Knowledge of Pub/Sub for streaming data pipelines.
- Experience with Cloud Functions .
- Robust SQL programming skills.
- Proficiency in Python .
- Experience in ETL/ELT development.
- Experience in data modeling and data warehousing concepts.
- Hands-on experience with Git and version control.
- Strong debugging and performance tuning skills.
Preferred Skills
- Experience with Dataproc , Spark, or Hadoop.
- Knowledge of Kafka or other streaming platforms.
- Experience with Terraform for Infrastructure as Code.
- Familiarity with Docker and Kubernetes.
- Exposure to Looker or Looker Studio.
- Knowledge of DevOps and CI/CD pipelines.
- Experience working with REST APIs.
- Understanding of data governance and security best practices.