31 Jul
|
Infosys
|
Telangana
Educational Requirements
Bachelor of Engineering
Service Line
Data Analytics Unit
Responsibilities
Data Engineering DevelopmentDesign, develop, and maintain Spark-based batch processing pipelines using Scala for large datasets.
Implement effective transformations, aggregations, and joins, ensuring correctness and scalability.
Write optimized SQL for data extraction, validation, and reconciliation across sources and targets.
Performance, Quality ReliabilityTune Spark jobs (partitioning, caching, shuffles, memory/executor settings) to improve runtime and cost efficiency.
Build data quality checks and validations to ensure accuracy, completeness, and consistency of outputs.
Troubleshoot production issues, perform root-cause analysis, and implement preventive fixes.
Collaboration DeliveryWork with stakeholders to understand data requirements and translate them into technical solutions.
Participate in code reviews, follow engineering best practices, and contribute to reusable components.
Document pipelines, logic, and operational runbooks for maintainability and onboarding.
Additional Responsibilities:
Bachelors degree in Computer Science, Engineering, or a related field (or equivalent practical experience).
59 years of overall experience with strong hands-on development in Spark and Scala.
Solid experience writing and optimizing SQL for analytics and data processing use cases.
Robust understanding of distributed processing concepts, data transformations, and performance considerations.
Ability to debug and resolve issues in data pipelines with a focus on reliability and quality.
Technical and Skilled Requirements:
Technology->Big Data - Data Processing->Spark,Technology->Java->Apache->Scala
Preferred Skills:
Technology->Java->Apache->Scala
Technology->Big Data - Data Processing->Spark
📌 Spark Scala Developer Telangana
🏢 Infosys
📍 Telangana