09 Sep
|
Willware Technologies
|
India
09 Sep
Willware Technologies
India
Job Type: Full-time
Remote: Hybrid
Principal AI EngineerLocation: India (Bengaluru) - 3(WFO)Employment Type: Full Time (Overlapping EST)Experience Level: Staff/Principal (8–14 years) Working Hours: Comfortable with a daily overlap into US Eastern morning hours (8AM - 12 PM EST)About the RoleTuring is hiring a Staff/Principal AI Engineer to lead enterprise-scale agentic AI implementations for Fortune 500 clients. This is a hands-on engineering role focused on designing and shipping autonomous, tool-calling AI systems — agents that reason over enterprise context, invoke real systems through secure interfaces, and operate reliably at scale under strict latency, cost, and governance constraints.You will own these systems end to end: the data pipelines feeding them, the backend services around them, the agent orchestration layer, the evaluation harness that keeps them honest, and the cloud infrastructure they run on. We are looking for engineers with genuine software engineering and data science depth who have taken agentic systems all the way to production.What We're Looking ForEngineering foundation8–14 years of software engineering experience, with robust hands-on large-scale PythonWorking depth in at least one systems or backend language — Go, Rust, Java, or C/C++ — and the judgment to know when to reach for itStrong data structures and algorithms. Strong understanding of APIs, microservices, and system designHands-on experience building and operating data pipelines and production-grade distributed systems.Agentic AI and LLMs2+ years of hands-on LLM engineering, with at least couple agentic system you designed and took to productionProduction experience with agent frameworks — LangGraph, Google ADK, CrewAI, Claude Agent SDK,
or equivalent — and the fluency to move between them as the ecosystem evolvesExperience building MCP (Model Context Protocol) servers and tool-calling interfacesRAG from first principles: chunking strategy, embeddings, vector and hybrid retrieval, reranking, and response validationStrong experience with vector databases (Milvus, Pinecone, Weaviate, FAISS, etc. or cloud equivalents)Design of guardrails and reliability patterns — validators, policy checks, self-correction loops, deterministic fallbacks, circuit breakers, and rollback pathsOptimizationDeep familiarity with token optimization and context-window management — context shaping, pruning, and compactionLatency and cost optimization through caching, model routing, batching, streaming, and parallel tool callsPerformance testing and tuning systems against defined SLOsEvaluationExperience building evaluation frameworks for LLM systems — offline eval sets, continuous online evaluation, and regression detectionInstrumentation and traceability suitable for regulated enterprise environments using tools like LangSmith, Langfuse, etc.CloudHands-on AWS: containerized services (ECS/EKS), serverless (Lambda), data services (S3, DynamoDB, Redshift) and orchestration (Step Functions)equivalents also valuedFamiliarity with CI/CD pipelines and DevOps practicesInfrastructure as code with Terraform or CloudFormation, and mature CI/CD practiceWorking traitsStrong analytical problem-solving with a bias to ownership and urgencyClear cross-team communication,
working directly with client stakeholders to translate business problems into technical roadmapsAble to work productively in ambiguity from system-level documentation and ramp quickly in unfamiliar codebasesGood to HaveExperience with managed AI platforms — Amazon Bedrock, Vertex AI, Azure AI — paired with fluency in the underlying fundamentalsAzure or GCP Roles & ResponsibilitiesDesign and build agentic systems: Lead the architecture and implementation of tool-calling agents that combine retrieval, structured reasoning, and secure action execution with least-privilege access.Productionize LLM applications: Build retrieval pipelines, prompt synthesis, response validation, and self-correction loops, backed by rigorous evaluation.Own the full stack: Deliver the data pipelines, backend services, distributed compute, and orchestration layer that agentic systems depend on — not only the model invocation.Engineer for reliability and governance: Build validator models, adversarial test suites, and policy checks; enforce deterministic fallbacks and rollback strategies; instrument continuous evaluation.Optimize for cost and latency: Drive measurable improvements in token efficiency, response time, and unit economics against defined SLOs.Codebase ownership: Build, maintain, and review high-quality Python and SQL, with an emphasis on reusable components, scalability, and performance.Cloud integration: Deploy AI applications on AWS, Azure, or GCP with optimized resource usage and robust CI/CD.Cross-functional collaboration: Partner with product owners, data scientists, and business SMEs to define requirements and deliver impactful AI products.Mentoring and technical leadership: Set engineering standards and share knowledge across the team, raising the bar on AI and software engineering practice.
📌 Principal AI Engineer (India)
🏢 Willware Technologies
📍 India