20 Aug
|
Capgemini
|
Gurugram
20 Aug
Capgemini
Gurugram
Key Responsibilities:
Must have working knowledge in GCP Dataflow, BigQuery, Pub/Sub, Cloud Composer, and Cloud Storage.
Hands-on experience with CI/CD tools and DevOps practices.
Collaborate with data scientists, analysts, and application developers to deliver high-quality data solutions.
Ensure data quality, governance, and security across all pipelines and storage layers.
Write productive and maintainable code in Python and PySpark.
Develop and optimize SQL queries for data extraction and transformation.
Experience with big data tools: Hadoop, Spark, Kafka, etc.
Experience with relational databases such as Microsoft SQL Server, MySQL, PostGreSQL, Oracle and NoSQL databases such as Hadoop, Cassandra, Mongo dB
Solid analytic skills related to working with structured, semi structured, unstructured datasets.
Build processes supporting data transformation, data structures, metadata, dependency and workload management.
Implement data quality checks and monitoring,
ensuring data accuracy and reliability.
Working knowledge of message queuing, stream processing, and highly scalable big data’ data stores.
Experience/Skills:
GCP Dataflow, BigQuery, Pub/Sub, Cloud Composer, and Cloud Storage.
Proficiency in Python and PySpark.
Robust SQL skills, including query optimization and performance tuning.
Familiarity with data warehousing, data lake concepts and ETL processes.
Experience with big data tools: Hadoop, Spark, Kafka, etc.
Experience with relational databases such as Microsoft SQL Server, MySQL, PostgreSQL, Oracle, and NoSQL databases such as Hadoop, Cassandra, MongoDB.
Excellent problem-solving skills with an emphasis on sustainable and reusable development.
Strong communication and collaboration skills
📌 Gcp Data Engineer/architect Gurugram
🏢 Capgemini
📍 Gurugram