Data Engineer (Bengaluru)

Data Engineer (Bengaluru)

11 Sep
|
Amro Partners
|
Bengaluru

11 Sep

Amro Partners

Bengaluru

About AMRO DATA LABS

We are a fast-growing data company focussed on Real Estate as our domain of excellence and with a pan-European capability for supporting data-centric decision-making for our parent group and soon for 3rd party clients.

Our aim is simply to be the best global, intelligent solutions partner for fast-growing commercial real estate companies in the residential and living sectors.

About the Role

We are seeking a proactive Data Engineer to join us on our journey. You will collaborate closely with our London based Data team to build and maintain robust data pipelines that drive insights to support our dynamic pricing models and our market intelligence data products.

Key Responsibilities

- Data Extraction & Scraping: Build, maintain, and continuously improve our Python-based web scrapers (using libraries like Playwright, Selenium etc) capturing daily rent and availability metrics across a wide range of websites. A core focus is making our extraction pipelines faster, more resilient, and as automated as possible.
- AI & LLM Data Engineering: Design and maintain LLM-driven workflows for extracting structured property data from unstructured sources (text descriptions, documents, web content, images). Build automated checks for output hallucination, graceful failure, and prompt injection to keep our AI-dependent ETL pipelines accurate and secure.
- Data Transformation: Write and optimize SQL workflows to transform raw scraped data into useful metrics, enforcing data integrity at the analytical layer.
- Data Quality & Auditing: Implement and maintain testing frameworks and auditing protocols that ensure data integrity across the extraction and transformation layers. Continuously audit real estate data across BigQuery,



Cloud Storage, SharePoint etc., and proactively find and fix anomalies.

Required Qualifications & Experience

- Education: Minimum :Bachelor’s degree in computer science, data science, or a closely related field and preferably a Master’s degree in computer science, data science or a closely related field. Both should be from top universities in India, UK or USA.
- Python: 4+ years of professional, hands-on experience. Strong proficiency in Object-Oriented programming and building ETL/ELT pipelines. Experience building and maintaining production web scrapers using Playwright, Selenium or similar, and hardening pipelines against breakage across many sources. Experience using LLM APIs (e.g., OpenAI API, Gemini API, LangChain etc.), Places API, and libraries like Pandas and Pydantic.
- SQL & NoSQL: 3+ years of professional experience writing advanced, optimized queries for data transformation and analytics. Experience using NoSQL databases.
- AI/LLM: Demonstrable prior work experience integrating LLMs into data pipelines for structured data extraction from unstructured text and building testing/validation suites for LLM outputs.
- Google Cloud Platform (GCP): 3+ years using services like BigQuery, Firestore, Cloud Storage, Dataform, Cloud Run and Compute Engine.
- Quality Assurance / Testing:



Hands-on experience with Pytest/Unittest and data-quality tooling such as Dataplex/Great Expectations, auditing and maintaining large-scale datasets.
- GitHub & DevOps: Experience with GitHub, Terraform, and setting up CI/CD pipelines (e.g., GitHub Actions).
- IDEs: Work experience using LLM-supported coding tools like Claude Code, Codex, Cursor, Antigravity.

Technical Stack & Workplace While you may not use every tool daily, you will be operating within and building on the following stack:

- Programming Languages & Libraries: Python, SQL, NoSQL, YAML, Playwright, Selenium, Crawl4ai, Pydantic, Pandas.
- AI/LLM: OpenAI API, Gemini API, Vertex AI, LangChain, CrewAI, RAG architectures.
- Cloud & Infrastructure (GCP): BigQuery, Cloud Storage, Compute Engine, Cloud Run Jobs, Cloud Run Functions, Cloud Build, Cloud Scheduler, Firestore, Dataform, Artifact Registry, Secrets Manager & Cloud Dataflow.
- DevOps, QA & Tools: GitHub, GitHub Actions (CI/CD), Docker, Terraform (IaC), Linux, Pytest, Pydantic & Jira.

Salary and benefits Commensurate with degree qualifications and prior work experience but not less than 20 LPA.

Application process A current CV

Covering letter explaining your interest in the role, why your application should be considered

Statement of intent: Your current city / location and your notice period with current employer to show your readiness together with a copy of transcripts to confirm degree qualifications

Application deadline

30th September, 2026. Start date is immediate or no later than 15th November 2026.

Pay: ₹1,800,000.00 - ₹2,000,000.00 per year

Work Location: Hybrid remote in Bangalore City, Bengaluru, Karnataka

📌 Data Engineer (Bengaluru)
🏢 Amro Partners
📍 Bengaluru

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: data engineer (bengaluru) / bengaluru