13 Sep
|
Synechron
|
Bengaluru
13 Sep
Synechron
Bengaluru
Job Summary /n We are seeking an experienced Data Engineer (PySpark) to join our Data Engineering team. The ideal candidate will be responsible for designing, developing, and maintaining large-scale data pipelines, data marts, and analytical solutions. The role requires strong expertise in Python, PySpark, SQL, Data Warehousing, and Big Data technologies, along with hands-on experience across the end-to-end Software Development Life Cycle (SDLC). /n /n Key Responsibilities /n /n
- Design, develop, and maintain scalable ETL pipelines and data processing frameworks.
/n
- Build and support Data Mart solutions to meet business and analytical requirements.
/n
- Develop high-quality, maintainable, and efficient code using Python and PySpark.
/n
- Participate in the complete Software Development Life Cycle (SDLC), including:
/n
- Requirement Analysis
/n
- Development
/n
- Unit Testing
/n
- UAT Support
/n
- Defect Fixing
/n
- Production Deployment
/n
- Post-Production Support
/n
- Perform complex data analysis and troubleshoot data-related issues.
/n
- Debug and optimize PySpark applications for performance and reliability.
/n
- Write and optimize complex Oracle SQL queries.
/n
- Work with structured, semi-structured, and unstructured datasets.
/n
- Implement data quality, validation, and monitoring processes.
/n
- Collaborate with cross-functional teams to resolve dependencies and ensure timely project delivery.
/n
- Follow software engineering best practices, coding standards, testing methodologies, and CI/CD processes.
/n /n /n Required Skills & Experience /n Data Engineering & Big Data /n /n
- PySpark
/n
- Apache Spark
/n
- Hadoop Ecosystem
/n
- MapReduce
/n
- Hive
/n
- ETL Development
/n
- Data Pipeline Development
/n
- Data Mart Development
/n
- Data Warehousing
/n /n Programming /n /n
- Python
/n
- Pandas
/n
- Jupyter Notebook
/n /n Database Technologies /n /n
- Oracle SQL
/n
- SQL
/n
- NoSQL Databases
/n
- Query Optimization
/n
- Data Analysis
/n /n Workflow & Automation /n /n
- Apache Airflow
/n
- Oozie
/n
- Jenkins
/n
- Git
/n
- CI/CD Pipelines
/n /n Data Engineering Techniques /n /n
- Data Cleansing
/n
- Data Linking
/n
- Feature Engineering
/n
- Data Imputation Techniques
/n
- Data Validation and Quality Checks
/n /n /n Preferred Experience /n /n
- Banking or Financial Services domain experience.
/n
- Experience supporting enterprise-scale data platforms.
/n
- Exposure to production support and deployment activities.
/n
- Experience working in Agile environments.
/n /n /n Desired Competencies /n /n
- Strong analytical and problem-solving skills.
/n
- Excellent debugging and troubleshooting capabilities.
/n
- Ability to lead technical initiatives with ownership and accountability.
/n
- Solid collaboration and stakeholder management skills.
/n
- Excellent verbal and written communication skills.
/n
- Ability to communicate effectively with both technical and non-technical stakeholders.
/n
- Ability to prioritize tasks and perform effectively in a fast-paced environment.
/n /n
📌 Pyspark Data Engineer (Bengaluru)
🏢 Synechron
📍 Bengaluru