19 Aug
|
Capgemini
|
Gurugram
19 Aug
Capgemini
Gurugram
Key Responsibilities:
- Must have working knowledge in GCP Dataflow, BigQuery, Pub/Sub, Cloud Composer, and Cloud Storage.
- Hands-on experience with CI/CD tools and DevOps practices.
- Collaborate with data scientists, analysts, and application developers to deliver high-quality data solutions.
- Ensure data quality, governance, and security across all pipelines and storage layers.
- Write productive and maintainable code in Python and PySpark.
- Develop and optimize SQL queries for data extraction and transformation.
- Experience with big data tools: Hadoop, Spark, Kafka, etc.
- Experience with relational databases such as Microsoft SQL Server, MySQL, PostGreSQL, Oracle and NoSQL databases such as Hadoop, Cassandra, Mongo dB
- Strong analytic skills related to working with structured, semi structured, unstructured datasets.
- Build processes supporting data transformation, data structures, metadata, dependency and workload management.
- Implement data quality checks and monitoring,
ensuring data accuracy and reliability.
- Working knowledge of message queuing, stream processing, and highly scalable big data’ data stores.
Experience/Skills:
- GCP Dataflow, BigQuery, Pub/Sub, Cloud Composer, and Cloud Storage.
- Proficiency in Python and PySpark.
- Strong SQL skills, including query optimization and performance tuning.
- Familiarity with data warehousing, data lake concepts and ETL processes.
- Experience with big data tools: Hadoop, Spark, Kafka, etc.
- Experience with relational databases such as Microsoft SQL Server, MySQL, PostgreSQL, Oracle, and NoSQL databases such as Hadoop, Cassandra, MongoDB.
- Excellent problem-solving skills with an emphasis on sustainable and reusable development.
- Strong communication and collaboration skills
📌 GCP Data Engineer/Architect (Gurugram)
🏢 Capgemini
📍 Gurugram