14 Sep
|
Halcer
|
Bengaluru
About Halcer Halcer is a leading IT services, consulting, and staffing organization committed to delivering innovative technology solutions and top-tier technical talent to global enterprises.
nJob Overview We are seeking an experienced Data Engineer (6+ Years of Experience) to build and maintain scalable, high-performance data pipelines and infrastructure for our next-generation data platform. The platform ingests and processes real-time and historical data from diverse industrial sources such as airport systems, sensors, cameras, and APIs. You will work closely with AI/ML engineers, data scientists, and DevOps teams to enable reliable analytics, forecasting, and anomaly detection use cases.
nKey Responsibilities
n
n
- Design and implement real-time (Kafka, Spark/Flink) and batch (Airflow, Spark) pipelines for high-throughput data ingestion, processing, and transformation.
n
- Develop data models and manage data lakes and warehouses (Delta Lake, Iceberg, etc.) to support both analytical and ML workloads.
n
- Integrate data from diverse sources: IoT sensors, databases (SQL/NoSQL), REST APIs, and flat files.
n
- Ensure pipeline scalability, observability, and data quality through monitoring, alerting, validation, and lineage tracking.
n
- Collaborate with AI/ML teams to provision clean and ML-ready datasets for training and inference.
n
- Deploy, optimize, and manage pipelines and data infrastructure across on-premise and hybrid environments.
n
- Participate in architectural decisions to ensure resilient, cost-effective, and secure data flows.
n
- Contribute to infrastructure-as-code and automation for data deployment using Terraform, Ansible, or similar tools.
n
nQualifications &
Experience Requirements
n
n
- Total Experience: 6+ years of total experience in core data engineering roles.
n
- Streaming & Real-Time Experience: 2+ years of direct experience building and maintaining real-time or streaming pipelines.
n
- Education: Bachelor's or Master's degree in Computer Science, Engineering, or a related field.
n
nRequired Skills
n
n
- Strong programming proficiency in Python or Java, alongside expert-level SQL.
n
- Hands-on experience with Apache Kafka, Apache Spark, or Apache Flink for real-time and batch processing.
n
- Proficiency with workflow orchestration tools like Airflow, dbt, or similar technologies.
n
- Deep familiarity with data modeling (OLAP/OLTP), schema evolution, and file formats (Parquet, Avro, ORC).
n
- Hands-on experience with hybrid/on-premise and cloud platform deployments (AWS, GCP, or Azure).
n
- Proven track record working with data lakes and modern data warehouses (Snowflake, BigQuery, Redshift, or Delta Lake).
n
- Strong working knowledge of DevOps practices, Docker, Kubernetes, and IaC tools like Terraform or Ansible.
n
- Knowledge of data observability, data cataloging, and quality frameworks (e.g., Outstanding Expectations, OpenMetadata).
n
nGood-to-Have Skills
n
n
- Experience with time-series databases (e.g., InfluxDB, TimescaleDB) and processing high-frequency sensor data.
n
- Prior domain experience in aviation, manufacturing, or logistics.
n
📌 Data Engineer (Bengaluru)
🏢 Halcer
📍 Bengaluru