14 Sep
|
Tanla Platforms
|
Hyderabad
14 Sep
Tanla Platforms
Hyderabad
Job Description
About the Role:
n
We are looking for a highly skilled Data Engineer with solid experience in building scalable ETL/ELT pipelines, distributed data systems, and modern lakehouse architectures. The ideal candidate will work on large-scale telecom and CPaaS datasets, including Call Detail Records (CDR), enabling real-time analytics and business intelligence across the platform ecosystem.
n
What you'll be Responsible for?
n
n
- Design and implement scalable ETL/ELT pipelines for large-scale analytics, data processing, and migration workloads.
n
- Build modern data lakehouse platforms using Iceberg, Delta Lake, or Hudi with catalog services like Nessie, AWS Glue, or Hive Metastore.
n
- Develop and optimize high-performance SQL queries and distributed data processing jobs using Spark (PySpark), Hadoop, and Kafka.
n
- Design and manage data warehouses and analytical platforms using Snowflake, ClickHouse, Dremio, Redshift, Trino, or Presto.
n
- Build ingestion and transformation pipelines using object storage systems such as Amazon S3, Azure Data Lake, GCS, or Nutanix Object Storage.
n
- Process and transform telecom datasets and Call Detail Records (CDR) efficiently at scale.
n
- Implement orchestration workflows using Airflow, Kestra, or similar workflow engines.
n
- Ensure data quality, governance, lineage, observability, scalability, and cost optimization across distributed systems.
n
- Build reusable frameworks for bulk data movement, ingestion acceleration, and transformation at scale.
n
n
What you'd have?
n
n
- 7 –10 years of experience in Data Engineering or related roles.
n
- Strong expertise in Advanced SQL including query optimization, partitioning, indexing, and performance tuning.
n
- Hands-on experience with Apache Spark (PySpark), Hadoop, Kafka, and distributed data processing systems.
n
- Strong expertise in lakehouse technologies such as Iceberg, Delta Lake, or Hudi.
n
- Experience with metadata/catalog systems including Nessie, Glue, or Hive Metastore.
n
- Knowledge of analytical engines such as ClickHouse, Dremio, Trino, or Presto.
n
- Strong understanding of Parquet, ORC, and Avro data formats.
n
- Experience with object storage systems like S3, ADLS, GCS, or Nutanix Object Storage.
n
- Strong programming skills in Python / PySpark / Scala.
n
- Experience with Airflow, Kestra, or similar orchestration tools.
n
- Hands-on exposure to AWS, Azure, or GCP cloud platforms.
n
- Experience in telecom data systems or CDR processing is highly preferred.
n
n
Why join us?
n
n
- Impactful Work: Build large-scale data platforms and analytics systems that power real-time communication products used by millions globally.
n
- Tremendous Growth Opportunities: Accelerate your career by solving complex engineering challenges in a fast-growing CPaaS and product-driven environment.
n
- Innovative Environment: Work alongside world-class engineers building cutting-edge distributed data systems, lakehouse architectures, and cloud-native platforms.
n
n
Tanla is an equal opportunity employer. We champion diversity and are committed to creating an inclusive environment for all employees.
n
www.Karix.com
📌 Senior Data Engineer (Hyderabad)
🏢 Tanla Platforms
📍 Hyderabad