Principal Data Engineer – Driving Data Excellence & BI Transformation
At Siemens, we help organizations transform maintenance and operations through connected insights, AI-powered technology, and intelligent asset management solutions. Our software enables customers to manage the full lifecycle of assets, facilities, and infrastructure while improving efficiency, reducing risk, and optimizing long-term investments. By connecting data, people, and processes, we empower organizations to make smarter decisions, maximize asset performance, and achieve more resilient operations.
Description:
We are seeking a highly skilled and experienced Principal Data Engineer to be a cornerstone of our data team. This pivotal role will focus on architecting, building, and optimizing our data infrastructure, driving BI enablement, and leading critical data transformation and migration projects. You will be instrumental in designing and implementing scalable, efficient, and robust ETL/ELT solutions that power our analytics and business intelligence initiatives.
You’ll make an impact by:
- Leading the design, development, and optimization of large-scale data pipelines and data warehouse/lake solutions.
- Driving data integration (ETL/ELT) strategies and implementations across diverse platforms and technologies.
- Collaborating with BI and analytics teams to ensure data availability, quality, and accessibility for reporting and insights.
- Mentoring junior engineers and contributing to establishing best practices in data engineering and software development.
- Evaluating and implementing new data technologies and approaches to enhance our data ecosystem.
- Championing agentic development across the data team — applying Claude Code or a comparable agentic coding tool (e.g., Snowflake CoCo, Cursor, GitHub Copilot agent mode, Windsurf, Codex CLI, Gemini CLI) to accelerate pipeline development, refactoring, testing, and code review, connecting those agents to governed data context through MCP servers such as the Snowflake and dbt MCP servers, and establishing the guardrails, prompting patterns, and review practices that keep agent-assisted output production-grade.
This is how you'll win us over:
- Bachelor's degree in Engineering or Science, or equivalent practical experience.
- 10+ years of progressive experience in software and data engineering, with at least 8 years dedicated to data-centric roles.
- Mastery of programming languages such as Java, Scala, and/or Python, together with comprehensive knowledge of relational databases and advanced SQL skills to support analytical needs.
- Extensive cloud data experience, with in-depth expertise in AWS-based data services (e.g., Kinesis, Glue, RDS, Athena) and Snowflake Cloud Data Warehouse.
- Extensive experience designing and implementing data integration (ETL/ELT) solutions across diverse platforms using Java, Scala, Python, PySpark, and SparkSQL, with hands-on proficiency in ETL/ELT tools such as dbt and robust data pipeline orchestration skills.
- Demonstrated ability to build, optimize, and maintain production-grade data pipelines supporting batch, replication/CDC, and event streaming patterns for data lakes and warehouses.
- Skilled in data modeling, data migration strategies, and performance tuning for large-scale data systems.
- A proven history of contributing to or leading major initiatives involving the consolidation and rationalization of large-scale data environments, including complex data pipelines and internal/external partner integrations.
- Familiarity with the Software Development Life Cycle (SDLC), source control (e.g., Git), and best practices for ensuring data quality, with adherence to software engineering and Agile development best practices.
- Ability to align work schedule with the US Eastern Standard Time (EST) zone to ensure at least a 6-hour overlap with our US teams.
You'll Thrive Even More If You Also Bring
- Familiarity with leading BI tools such as Power BI and Apache Superset.
- Familiarity with AI-powered coding tools (e.g., GitHub Copilot, Claude Code) to enhance productivity and code quality.
- Hands-on experience with agentic development workflows,
particularly Claude Code or an equivalent agent such as Snowflake CoCo, Cursor, GitHub Copilot agent mode, Windsurf, Codex CLI, or Gemini CLI — decomposing data engineering work into tasks an agent can execute end to end, steering and reviewing agent-generated code and tests, and applying sound judgment about where human review and design ownership remain essential.
- Experience with Snowflake CoCo, Snowflake's data-native coding agent (available in Snowsight, CoCo Desktop, and the CLI), for generating SQL, Python, and pipelines grounded in catalog, lineage, RBAC, and compute context, and for reviewing agent-suggested changes through its diff view before they are applied.
- Practical use of Model Context Protocol (MCP) servers to give agents governed access to the data stack — for example the Snowflake MCP server (Cortex Search, Cortex Analyst, and Cortex Agent tooling) and the dbt MCP server (semantic layer queries, metadata and lineage discovery, and dbt operations such as compiling models and running tests) — along with the authentication and least-privilege considerations their service tokens require.
- A proactive and innovative approach to exploring and implementing cutting-edge technologies and solutions.
- Outstanding written and verbal communication skills, with the ability to articulate complex technical concepts to both technical and non-technical stakeholders.
At Siemens, you'll have the prospect to grow your career while helping organizations operate smarter, safer, and more sustainably. We foster a culture of innovation, collaboration, and continuous learning, where employees are empowered to make a difference every day. If you're excited about solving real-world challenges and shaping the future of asset management, we encourage you to apply.
Our Commitment to Equity and Inclusion in our Diverse Global Workforce:
We value your unique identity and perspective. We are fully committed to providing equitable opportunities and building a workplace that reflects the diversity of society, while ensuring that we attract the best talent based on qualifications, skills, and experiences. We welcome you to bring your authentic self and transform the every day with us.
Siemens maintains a Drug Free workplace in accordance with applicable law.
📌 Principal Data Engineer (Noida)
🏢 Siemens
📍 Noida