Engineering Manager (India)

Engineering Manager (India)

09 Oct
|
Tekwissen India
|
India

09 Oct

Tekwissen India

India

Overview :

Tek Wissen is a global workforce management provider throughout India and many other countries in the world. The job opportunity described below is for one of our clients, who has developed a core competence in creating and deploying cost-effective capabilities using an offshore-centric business model.

Position: Engineering Manager

Location: Remote

Job Type: Contract
Duration: 6 Months

Work Type: Remote

Job Description:

- Engineering Manager & Principal Systems Architect (Agentic Ops Core)
- We are seeking a hands-on Engineering Manager & Principal Systems Architect to serve as the technical anchor, customer-facing architect, and engineering leader for our enterprise Agentic Ops platform.
- In this player-coach role, you will lead a specialized engineering pod building production grade autonomous runtimes for enterprise IT and Sec Ops. You will own the architectural blueprint, write production code, fine-tune models, and work directly with enterprise customers and technical partners across NVIDIA and Cisco.

Core Responsibilities:

- Architecture & Hands-On Engineering: Lead, design, and personally code an event-driven, high-throughput multi-agent platform in Golang for closed-loop IT self-healing and Sec Ops containment.
- Customer & Partner Advisory: Serve as primary technical lead for enterprise clients and design partners (Cisco, NVIDIA). Lead deep-dive architectural workshops and translate complex operational workflows into multi-agent specifications.
- NVIDIA AI Stack & Inference Acceleration: Architect GPU-accelerated inference pipelines using NVIDIA NIM microservices, NeMo Framework/Guardrails, and TensorRT-LLM across DGX clusters, hybrid cloud, and edge.
- Model Adaptation & Fine-Tuning: Lead data curation and fine-tuning pipelines (PEFT, LoRA/QLoRA, distillation) to adapt open-weight models (Llama, Mistral, Qwen, Deep Seek) for IT/Sec Ops tasks.
- Agent Frameworks & Open-Source Ecosystem: Build durable agent runtimes using Lang Chain, Lang Graph, or custom state machines. Leverage open-source agentic and execution tooling (e.g., Open Claw, Nemoclaw) and local inference engines (vLLM, Ollama).
- Safety, Governance & AI-Native SDLC: Engineer deterministic safety harnesses (dry-run preflights, blast-radius containment, rollbacks) and establish an AI-augmented SDLC (agentic PR reviews, synthetic evals, LLM-as-a-judge).




- Team Leadership & Integration: Recruit, mentor, and lead an elite pod; enforce rigorous code reviews, evals, and Open Telemetry observability while integrating with enterprise substrates (Cisco Cloud Control, Intersight, Nexus, Splunk).

Required Qualifications:

- Experience & Scope: 10+ years architecting scalable distributed systems and backend infrastructure, with 3+ years as a hands-on player-coach leading engineering teams.
- Customer-Facing Acumen: Proven ability to interface with enterprise architects and leaders on technical discovery, system architecture, and complex deployments.
- Core Languages: Deep production mastery of Go (Golang) for high-concurrency, low-latency microservices; strong proficiency in Python for AI/ML pipelines and orchestration.
- NVIDIA AI Stack: Hands-on experience with NVIDIA AI Enterprise, NIMs, NeMo Guardrails, TensorRT-LLM, and Triton Inference Server on modern GPU architectures.
- LLM Engineering & Fine-Tuning: Practical experience with fine-tuning toolchains (Hugging Face, PEFT/LoRA, Axolotl, Deep Speed), alignment (DPO/RLHF), quantization, and agent frameworks (Lang Chain, Lang Graph, DSPy).
- Open-Source Tooling: Experience self-hosting open models (vLLM, Ollama) and integrating open-source agentic/browser automation tools (Open Claw, Browser-Use).
- Distributed Systems & Safety: Strong background in event-driven streaming (Kafka, NATS, gRPC), vector/graph databases (pgvector, Qdrant, Neo4j), and zero-trust safety execution guardrails.
- Domain Fluency: Working knowledge of enterprise IT fabrics (SDN, data center networking), observability platforms (Splunk, Thousand Eyes), and SOC incident lifecycles.

Leadership & Role:

- Engineering Manager, Principal Systems Architect, Player-Coach
- Hands-on Technical Leadership, Team Mentoring, Code Reviews
- Customer-Facing Architect, Technical Discovery, Architectural Workshops

Core Languages & Backend

- Golang (Go), Python
- Distributed Systems, Event-Driven Architecture,



Microservices
- High-Concurrency, Low-Latency, High-Throughput
- gRPC, Kafka, NATS

Agentic AI

- Agentic Ops, Multi-Agent Systems, Autonomous Agent
- Lang Chain, Lang Graph, DSPy, Custom State Machines
- Open Claw, Browser-Use
- Closed-Loop Automation, IT Self-Healing, Sec Ops Containment

NVIDIA AI Stack:

- NVIDIA AI Enterprise, NVIDIA NIM, NeMo Framework, NeMo Guardrails TensorRT-LLM, Triton Inference Server
- DGX, GPU-Accelerated Inference

LLM Engineering & Fine-Tuning

- Fine-Tuning, PEFT, LoRA/QLoRA, Distillation, Quantization
- DPO, RLHF, Alignment
- Hugging Face, Axolotl, Deep Speed
- Open-Weight Models (Llama, Mistral, Qwen, Deep Seek)
- vLLM, Ollama, Self-Hosted LLMs

Data & Storage

- Vector Databases (pgvector, Qdrant), Graph Databases (Neo4j)

Safety & Governance

- Zero-Trust, Safety Guardrails, Dry-Run Preflight
- Blast-Radius Containment, Rollbacks
- LLM-as-a-Judge, Synthetic Evals, Agentic PR Reviews, AI-Native SDLC

Observability & Enterprise Integration

- Open Telemetry, Splunk, Thousand Eyes
- Cisco Cloud Control, Intersight, Nexus
- SDN, Data Center Networking, SOC Incident Lifecycle.

Experience Requirements

- 10+ Years in Distributed Systems / Backend Infrastructure
- 3+ Years Player-Coach / Team Lead
- Enterprise Customer Engagement (Cisco, NVIDIA partners)

Required Competencies

- Must possess excellent communication skills oral and written
- Must possess knowledge of latest technology trends
- Must be a keen learner should be able to drive "Self Learning"
- Must practice principle of "First Time Right"
- Must have an Eye for Details
- Must have high Customer Orientation
- Must be adaptable to working in multiple / matrix work environment
- Must possess good systems thinking
- Must possess positive negotiation, analytical and interpersonal skills. Good leadership & team player qualities.
- High on personal integrity with ability to establish relationships and work in teams and should be
- able to influence stakeholders. Should poses independence, robust ethics and resilience.

Years of Experience:

- 10+ years of relevant work experience with a reputed organization.

Educational Qualification:

- ME (IT, Computer), BE (IT, Computer), MCA, MSC-IT, BCA

Tek Wissen Group is an equal opportunity employer supporting workforce diversity

📌 Engineering Manager (India)
🏢 Tekwissen India
📍 India

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: engineering manager (india) / india