Role Proficiency This role requires proficiency in developing data pipelines including coding and testing for ingesting wrangling transforming and joining data from various sources The ideal candidate should be adept in ETL tools like Informatica Glue Databricks and DataProc with strong coding skills in Python PySpark and SQL This position demands independence and proficiency across various data domains Expertise in data warehousing solutions such as Snowflake BigQuery Lakehouse and Delta Lake is essential including the ability to calculate processing costs and address performance issues A solid understanding of DevOps and infrastructure needs is also required Outcomes Act creatively to develop pipelines applications by selecting appropriate technical options optimizing application development maintenance and performance through design patterns and reusing proven solutions Support the Project Manager in day-to-day project execution and account for the developmental activities of others Interpret requirements create optimal architecture and design solutions in accordance with specifications Document and communicate milestones stages for end-to-end delivery Code using best standards debug and test solutions to ensure best-in-class quality Tune performance of code and align it with the appropriate infrastructure understanding cost implications of licenses and infrastructure Create data schemas and models effectively Develop and manage data storage solutions including relational databases NoSQL databases Delta Lakes and data lakes Validate results with user representatives integrating the overall solution Influence and enhance customer satisfaction and employee engagement within project teams Measures of Outcomes TeamOne s Adherence to engineering processes and standards TeamOne s Adherence to schedule timelines TeamOne s Adhere to SLAs where applicable TeamOne s of defects post delivery TeamOne s of non-compliance issues TeamOne s Reduction of reoccurrence of known defects TeamOne s Quickly turnaround production bugs Completion of applicable technical domain certifications Completion of all mandatory training requirementst Efficiency improvements in data pipelines e g reduced resource consumption faster run times TeamOne s Average time to detect respond to and resolve pipeline failures or data issues TeamOne s Number of data security incidents or compliance breaches Outputs Expected Code Develop data processing code with guidance ensuring performance and scalability requirements are met Define coding standards templates and checklists Review code for team and peers Documentation Create review templates checklists guidelines and standards for design process development Create review deliverable documents including design documents architecture documents infra costing business requirements source-target mappings test cases and results Configure Define and govern the configuration management plan Ensure compliance from the team Test Review create unit test cases scenarios and execution Review test plans and strategies created by the testing team Provide clarifications to the testing team Domain Relevance Advise data engineers on the design and development of features and components leveraging a deeper understanding of business needs Learn more about the customer domain and identify opportunities to add value Complete relevant domain certifications Manage Project Support the Project Manager with project inputs Provide inputs on project plans or sprints as needed Manage the delivery of modules Manage Defects Perform defect root cause analysis RCA and mitigation Identify defect trends and implement proactive measures to improve quality Estimate Create and provide input for effort and size estimation and plan resources for projects Manage Knowledge Consume and contribute to project-related documents SharePoint libraries and client universities Review reusable documents created by the team Release Execute and monitor the release process Design Contribute to the creation of design HLD LLD SAD architecture for applications business components and data models Interface with Customer Clarify requirements and provide guidance to the Development Team Present design options to customers Conduct product demos Collaborate closely with customer architects to finalize designs Manage Team Set RAPID goals and provide feedback Understand team members aspirations and provide guidance and opportunities Ensure team members are upskilled Engage the team in projects Proactively identify attrition risks and collaborate with BSE on retention measures Certifications Obtain relevant domain and technology certifications Skill Examples Proficiency in SQL Python or other programming languages used for data manipulation Experience with ETL tools such as Apache Airflow Talend Informatica AWS Glue Dataproc and Azure ADF Hands-on experience with cloud platforms like AWS Azure or Google Cloud particularly with data-related services e g AWS Glue BigQuery Conduct tests on data pipelines and evaluate results against data quality and performance specifications Experience in performance tuning Experience in data warehouse design and cost improvements Apply and optimize data models for efficient storage retrieval and processing of large datasets Communicate and explain design development aspects to customers Estimate time and resource requirements for developing debugging features components Participate in RFP responses and solutioning Mentor team members and guide them in relevant upskilling and certification Knowledge Examples Knowledge Examples Knowledge of various ETL services used by cloud providers including Apache PySpark AWS Glue GCP DataProc Dataflow Azure ADF and ADLF Proficient in SQL for analytics and windowing functions Understanding of data schemas and models Familiarity with domain-related data Knowledge of data warehouse optimization techniques Understanding of data security concepts Awareness of patterns frameworks and automation practices Additional Comments Core Areas of Expertise Cloud Data Architecture Design AWS-focused Scalable ETL ELT Pipelines Big Data Engineering Data Lakes Lakehouses and Warehousing AWS Hybrid Distributed Systems Parallel Data Processing DevOps CI CD and Container Orchestration End-to-End Project Ownership From POC to Production Team Leadership Mentorship and Stakeholder Engagement Technical Skills Summary Cloud Infrastructure Cloud Platforms AWS expert Azure intermediate AWS Services extensive hands-on S3 Redshift RDS Athena Glue incl Spark Lambda Step Functions EC2 EMR DMS Data Catalog CloudWatch EKS API Gateway SNS MWAA Azure Services Blob Storage Data Factory VMs App Services Containerization Orchestration Docker EKS Kubernetes ECS Fargate Infrastructure as Code CI CD GitHub Bitbucket Jenkins CircleCI Terraform basic GitOps best practices Data Engineering Big Data Data Pipelines Architected and deployed large-scale pipelines with Spark Glue EMR Python Airflow Dagster Luigi ETL ELT Tools AWS Glue Spark PySpark SSIS Informatica Power Center Streamlined ingestion from on-prem SaaS and RDBMS sources to modern lakehouse architecture Performance Tuning Expert in optimizing queries storage and compute resources across Redshift SQL Server and Snowflake Data Modeling Dimensional modeling Kimball data marts warehouse optimization Data Governance Implemented metadata management quality rules and lineage Glue Catalog custom solutions Programming Scripting Programming Languages Python expert Shell scripting PHP R NET basic JavaScript Data Engineering Tools PySpark Pandas SQL advanced Shell scripts API integrations Automation Monitoring Custom scripts ing log analysis performance dashboards Databases Storage RDBMS PostgreSQL SQL Server MySQL Oracle DB2 SQLite Cloud DWH AWS Redshift Snowflake NoSQL Semi-Structured MongoDB Elasticsearch Cassandra Redis Object Storage S3 Azure Blob Big Data Tools Hadoop Hive Spark Glue EMR Data Visualization BI Dashboards Insights Tableau advanced Power BI SAP HANA Reporting KPI Metric Frameworks Built executive dashboards tied to business metrics and SLAs Experience delivering reporting solutions to support Pharma Insurance Telecom and Publishing sectors Web Backend Development Supportive Skills Backend Web Dev PHP JavaScript NET Excel Macros VBA SharePoint API Integration API Gateway REST services backend data exposure and enrichment Workflow Design Automating tasks and pipelines using scripting Python and platform-native schedulers Skills Data Analysis Data Structures Data Warehousing About Company UST is a global digital transformation solutions provider For more than 20 years UST has worked side by side with the world s best companies to make a real impact through transformation Powered by technology inspired by people and led by purpose UST partners with their clients from design to operation With deep domain expertise and a future-proof philosophy UST embeds innovation and agility into their clients organizations With over 30 000 employees in 30 countries UST builds for boundless impact touching billions of lives in the process
📌 Lead Ii - Data Engineering (Bengaluru)
🏢 UST
📍 Bengaluru
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.