07 Aug
|
Advance Career Solutions
|
Pune
07 Aug
Advance Career Solutions
Pune
Job Title
Data Engineer
Job Summary
We are looking for a skilled and detail-oriented Data Engineer to design, build, and maintain scalable data pipelines and infrastructure that support analytics, reporting, and machine learning initiatives. The ideal candidate has experience working with large datasets, cloud platforms, ETL/ELT processes, and modern data engineering tools. You will collaborate with data analysts, data scientists, software engineers, and business stakeholders to ensure reliable, high-quality data is available across the organization.
Key Responsibilities
- Design, develop, and maintain scalable ETL/ELT data pipelines.
- Build and optimize data architectures, databases, and data warehouses.
- Develop reliable data ingestion processes from multiple internal and external data sources.
- Ensure data quality, consistency, security, and governance across all data platforms.
- Monitor and troubleshoot data pipelines to ensure high availability and performance.
- Optimize SQL queries, database performance, and storage solutions.
- Collaborate with Data Analysts, Data Scientists, and Business Intelligence teams to support reporting and analytics.
- Implement data validation, monitoring, and automated testing for data workflows.
- Work with cloud-based data platforms such as AWS, Azure, or Google Cloud Platform.
- Maintain documentation for data models, data pipelines, and technical processes.
- Implement CI/CD practices for data engineering workflows.
- Support real-time and batch data processing requirements.
- Ensure compliance with organizational security and data privacy standards.
Required Qualifications
- Bachelor's degree in Computer Science, Information Technology, Data Science, Engineering, or a related field.
- 3+ years of experience as a Data Engineer or in a similar role.
- Strong proficiency in SQL and database optimization.
- Experience with Python, Scala, or Java.
- Hands-on experience with ETL/ELT tools and frameworks.
- Experience working with relational and NoSQL databases.
- Strong understanding of data modeling principles.
- Experience with cloud data services (AWS, Azure, or GCP).
- Knowledge of distributed data processing frameworks such as Apache Spark.
- Experience with version control systems such as Git.
Preferred Qualifications
- Experience with modern data warehouse platforms such as Snowflake, BigQuery, Redshift, or Synapse.
- Experience with orchestration tools like Apache Airflow.
- Familiarity with Kafka or other streaming platforms.
- Knowledge of containerization technologies (Docker, Kubernetes).
- Experience with Infrastructure as Code (Terraform, CloudFormation).
- Exposure to machine learning data pipelines.
- Relevant cloud certifications are a plus.
Required Technical Skills
Programming
- Python
- SQL
- Scala (preferred)
- Java (optional)
Databases
- PostgreSQL
- MySQL
- SQL Server
- Oracle
- MongoDB
- Cassandra
Big Data Technologies
- Apache Spark
- Hadoop
- Hive
- Kafka
Data Warehousing
- Snowflake
- Amazon Redshift
- Google BigQuery
- Azure Synapse
ETL/ELT Tools
- Apache Airflow
- dbt
- Informatica
- Talend
- Azure Data Factory
- AWS Glue
Data Engineer Job Description
Job Title
Data Engineer
Job Summary
We are looking for a skilled and detail-oriented Data Engineer to design, build, and maintain scalable data pipelines and infrastructure that support analytics, reporting, and machine learning initiatives. The ideal candidate has experience working with large datasets, cloud platforms, ETL/ELT processes, and modern data engineering tools. You will collaborate with data analysts, data scientists, software engineers, and business stakeholders to ensure reliable, high-quality data is available across the organization.
Key Responsibilities
- Design, develop, and maintain scalable ETL/ELT data pipelines.
- Build and optimize data architectures, databases, and data warehouses.
- Develop reliable data ingestion processes from multiple internal and external data sources.
- Ensure data quality, consistency, security, and governance across all data platforms.
- Monitor and troubleshoot data pipelines to ensure high availability and performance.
- Optimize SQL queries, database performance, and storage solutions.
- Collaborate with Data Analysts, Data Scientists, and Business Intelligence teams to support reporting and analytics.
- Implement data validation, monitoring, and automated testing for data workflows.
- Work with cloud-based data platforms such as AWS, Azure, or Google Cloud Platform.
- Maintain documentation for data models, data pipelines, and technical processes.
- Implement CI/CD practices for data engineering workflows.
- Support real-time and batch data processing requirements.
- Ensure compliance with organizational security and data privacy standards.
Required Qualifications
- Bachelor's degree in Computer Science, Information Technology, Data Science, Engineering, or a related field.
- 3+ years of experience as a Data Engineer or in a similar role.
- Strong proficiency in SQL and database optimization.
- Experience with Python, Scala, or Java.
- Hands-on experience with ETL/ELT tools and frameworks.
- Experience working with relational and NoSQL databases.
- Strong understanding of data modeling principles.
- Experience with cloud data services (AWS, Azure, or GCP).
- Knowledge of distributed data processing frameworks such as Apache Spark.
- Experience with version control systems such as Git.
Preferred Qualifications
- Experience with up-to-date data warehouse platforms such as Snowflake, BigQuery, Redshift, or Synapse.
- Experience with orchestration tools like Apache Airflow.
- Familiarity with Kafka or other streaming platforms.
- Knowledge of containerization technologies (Docker, Kubernetes).
- Experience with Infrastructure as Code (Terraform, CloudFormation).
- Exposure to machine learning data pipelines.
- Relevant cloud certifications are a plus.
Required Technical Skills
Programming
- Python
- SQL
- Scala (preferred)
- Java (optional)
Databases
- PostgreSQL
- MySQL
- SQL Server
- Oracle
- MongoDB
- Cassandra
Big Data Technologies
- Apache Spark
- Hadoop
- Hive
- Kafka
Data Warehousing
- Snowflake
- Amazon Redshift
- Google BigQuery
- Azure Synapse
ETL/ELT Tools
- Apache Airflow
- dbt
- Informatica
- Talend
- Azure Data Factory
- AWS Glue
Cloud Platforms
- Amazon Web Services (AWS)
- Microsoft Azure
- Google Cloud Platform (GCP)
DevOps & Version Control
- Git
- GitHub
- GitLab
- Jenkins
- Docker
- Kubernetes
DevOps & Version Control
- Git
- GitHub
- GitLab
- Jenkins
- Docker
- Kubernetes
Preferred candidate profile
📌 Data Engineer (Pune)
🏢 Advance Career Solutions
📍 Pune