Solid
proficiency in Python (including OOP and data manipulation).
Hands-on
experience with PySpark for distributed data processing.
Proficiency
in building and consuming RESTful
APIs using FastAPI or Flask.
Solid
understanding of AWS
services:
S3, Glue, Lambda, EMR, Athena, and CloudWatch.
Robust
SQL skills and experience with data
modeling and ETL
pipeline development.
Good
to Have Skills :
Experience
with Docker and containerizing APIs or Spark jobs.
Familiarity
with Apache
Airflow,
AWS Step Functions, or similar orchestration tools.
Exposure
to CI/CD
tools like Jenkins, GitHub Actions, or AWS CodePipeline.
API
testing and documentation tools (e.g., Postman, Swagger).
Knowledge
of Delta
Lake, Apache
Hudi,
or Apache
Iceberg for versioned data lakes.
Roles
and Responsibilities :
Design
and implement robust, scalable ETL/ELT
pipelines using Python and PySpark.
Build
and deploy RESTful
APIs using FastAPI or Flask to serve data and services.
Develop
and maintain data
workflows and jobs on AWS services such as Glue, Lambda, and EMR.
Monitor
and troubleshoot data
pipelines and APIs,
ensuring reliability and performance.
Collaborate
with cross-functional teams to understand
data needs,
deliver clean datasets, and optimize data architecture.
Location
:
Hyderabad
CTC
Range :
15
to 20 LPA
Notice
period :
15
DAYS
Shift
Timings :
Mode
of Interview :
FACE
TO FACE
Mode
of Work :
Mode
of Hire :
Note
:
📌 Python+pyspark+aws+api Rest/fast Bengaluru
🏢 Important Business
📍 Bengaluru
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.