07 Aug
|
Tata Consultancy Services
|
India
07 Aug
Tata Consultancy Services
India
Very good experience in implementing data platform modernization using GCP data services. Good data modelling skills on SQL/NoSQL based data platforms.
Key Skills: GCP Data Engineer, Pyspark, Dataflow, Dataproc, BigQuery, Airflow
Must Have Skills:
- Experience working in GCP based Big Data deployments (Batch/Realtime) leveraging components like GCP Big Query, air flow, Google Cloud Storage, Data fusion, Data flow, Data Proc etc.
- Positive skills in Python Language, PYSPARK
- Good Skills in Linux
- Exposure to creation of CI-CD pipelines for promoting big data release deployments and designing log monitoring features.
- Excellent written and verbal communication skills in English
Good to Have Skills: Experience in On Premise Hadoop technologies (Hive, Sqoop, SPARK, Kafka) Hadoop, Spark, Apache Beam
Responsibilities or Expectations from the Role:
- Data modelling/Data warehouse modernization/Cloud based data lakes
- Create and maintain optimal data pipeline architecture
- Define technical roadmap for data platform modernization and analysis and assessment of the existing data platforms
- Identify, design, and implement internal process improvements: automating manual processes, optimizing data delivery, re-designing infrastructure for greater scalability, etc.
- Build the infrastructure required for optimal extraction, transformation, and loading of data from a wide variety of data sources using SQL and Google cloud big data’ technologies.
- Work with stakeholders including the Executive, Product, Data and Design teams to assist with data-related technical issues and support their data infrastructure needs.
- Define KPIs for successful modernization of data platform and measure the improvement against the KPIs.
📌 GCP Data Engineer+ Pyspark (Anywhere in India)
🏢 Tata Consultancy Services
📍 India