24 Sep
|
Horizontal Integration
|
Bengaluru
24 Sep
Horizontal Integration
Bengaluru
Senior Data Engineer
Experience: 7 - 12 years
Location: Manyata, Bangalore
Notice: Immediate Joiner or 15 Days Less
Interview: 1st Round (Virtual) and 2nd Round (Face to face) Only
Bangalore Candidates Preferred.
We are looking for an experienced Senior Data Engineer who can quickly onboard and contribute to the design, development, optimization, and support of large-scale data platforms. The ideal candidate should be hands-on, comfortable working with complex data pipelines and distributed systems, and able to independently troubleshoot and deliver solutions in a fast-paced environment.
Key Responsibilities
- Design, develop, and optimize large-scale data pipelines using Spark and Scala.
- Work with high-volume batch and streaming workloads across multiple data sources.
- Develop and support data processing solutions using Hive, Kafka, BQ (GCP services)
- Troubleshoot complex data, performance, and production issues with minimal supervision.
- Contribute to cloud migration and platform modernization initiatives.
- Build and maintain ETL workflows and data orchestration pipelines.
- Perform performance tuning and optimization of Spark jobs and large datasets.
- Develop reusable, maintainable, and production-ready data solutions.
- Implement and maintain CI/CD pipelines and engineering automation.
- Collaborate with engineering, product, analytics, and platform teams to deliver end-to-end solutions.
- Participate in code reviews, testing, production deployments, and operational support.
- Quickly understand existing systems and take ownership of assigned deliverables.
Required Skills
- 7+ years of experience in Data Engineering / Big Data development.
- Strong hands-on expertise in Apache Spark and Scala.
- Strong experience with Hive and distributed data processing.
- Experience with Kafka and event/streaming-based architectures.
- Robust programming experience in Scala and/or Python.
- Experience designing and troubleshooting large-scale ETL/data pipelines.
- Strong understanding of data modeling, data quality, performance tuning, and distributed systems.
- Experience with Git and CI/CD practices.
Cloud Experience Strong hands-on experience with at least one major cloud platform:
GCP, Azure, AWS
Experience working across multiple cloud environments is an advantage.
ETL & Orchestration with any workflows
Oozie
CI/CD & DevOps
- Jenkins
- Git
- Automated build, testing, and deployment pipelines
Preferred / Added Advantage
- Experience with Google Cloud Platform (GCP).
- Exposure to BigQuery, GCS, Dataproc, and Pub/Sub.
- Experience migrating on-premise Spark/Hive workloads to cloud platforms.
- Experience with large-scale data migration and reconciliation.
- Knowledge of data quality and observability frameworks.
- Experience working in environments involving multiple upstream and downstream integrations.
What We Are Looking For Able to ramp up quickly and contribute with minimal handholding. We are specifically looking for someone who demonstrates:
- Strong problem-solving and debugging skills.
- Ability to understand complex existing systems quickly.
- Hands-on technical ownership.
- Ability to work independently while collaborating effectively with the broader team.
- Proactive identification of risks and solutions.
- Strong communication and documentation skills.
- Ability to drive assigned work from discovery through development, testing, and production deployment.
📌 GCP Data Engineer (Bengaluru)
🏢 Horizontal Integration
📍 Bengaluru