We use cookies for the best user experience on our website, including to personalize content & offerings, to provide social media features and to analyze traffic. By clicking “Accept All Cookies” you agree to our use of cookies. You can also manage your cookies by clicking on the "Cookie Preferences" and selecting the categories you would like to accept. For more information on how we use cookies please visit our and Modify Cookie Preferences
Reject All Cookies Accept All Cookies
Senior Technical Architect
India
Job Description
Senior Technical Architect
Hyderabad, Telangana
Job Summary
Own the on-device AI execution stack for ARM-based SoCs with dedicated neural accelerators - model conversion and quantization, runtime and delegate integration, and deployment of vision, LLM and VLM workloads within edge power and memory budgets
Key Responsibilities
- Deploy and optimize neural network models on NPU, GPU and CPU backends using embedded inference runtimes (LiteRT/TensorFlow Lite, ONNX Runtime, ExecuTorch or vendor SDKs).
- Own the model conversion pipeline - graph capture, operator mapping, quantization (PTQ and QAT support), calibration and accuracy validation against reference.
- Diagnose and resolve unsupported operators, graph partitioning and fallback behaviour; work with vendor toolchains on accelerator limitations.
- Enable on-device LLM and VLM workloads - weight quantization, KV-cache management, prefill/decode optimization, memory footprint and token-throughput tuning.
- Benchmark and optimize inference latency, throughput, memory bandwidth and energy per inference; publish reproducible performance data.
- Integrate inference into real-time media and robotics pipelines with zero-copy tensor/buffer sharing.
- Build model deployment and regression tooling so accuracy and performance are tracked across releases.
- Advise product and data-science teams on model architecture choices suitable for edge accelerators
Skill Requirements
- 8-10 years in software engineering, with 3+ years in on-device / edge AI deployment.
- Strong C++ and Python.
- Hands-on with embedded inference runtimes - LiteRT/TFLite, ONNX Runtime, ExecuTorch, TVM or vendor NPU SDKs - including delegate/execution-provider integration.
- Practical quantization expertise - INT8/INT4, per-channel schemes, calibration, accuracy recovery.
- Model formats and conversion tooling across PyTorch/TensorFlow to deployable graphs.
- Profiling on heterogeneous SoCs and reasoning about memory bandwidth as the dominant constraint.
- Understanding of CNN, transformer and contemporary vision/language model architectures
Other Requirements
- On-device LLM/VLM deployment - llama.cpp class runtimes, speculative decoding, paged KV cache.
- Custom operator or kernel development for DSP/NPU/GPU (OpenCL, Vulkan compute, DSP intrinsics).
- Compiler-level work - MLIR, TVM, graph-level optimization passes.
- Vision pipeline integration and camera-to-inference zero-copy paths.
- MLOps for edge - model versioning, A/B evaluation, field accuracy monitoring
Information at a Glance
Why HCLTech?
At HCLTech, you'll supercharge your potential. You'll find your career. And you'll find your spark. All at a place that knows that helping its customers stay on top starts by putting its people first.
HCLTech is a global technology company, home to more than 223,000 people across 60 countries, delivering industry-leading capabilities centered around digital, engineering, cloud and AI, powered by a broad portfolio of technology services and products. We work with clients across all major verticals, providing industry solutions for Financial Services, Manufacturing, Life Sciences and Healthcare, Technology and Services, Telecom and Media, Retail and CPG, and Public Services. Consolidated revenues as of 12 months ending June 2026 totaled $14.8 billion.
Copyright © 2026 HCL Technologies Limited
×
Cookie Consent Manager
When you visit any website, it may store or retrieve information on your browser, mostly in the form of cookies. Because we respect your right to privacy, you can choose not to allow some types of cookies. However, blocking some types of cookies may impact your experience of the site and the services we are able to offer.
Required Cookies
These cookies are required to use this website and can't be turned off.
Required Cookies
Show More Details
Required Cookies Provider Description Enabled
SAP as service provider
We use the following session cookies, which are all required to enable the website to function:
- "route" is used for session stickiness
- "careerSiteCompanyId" is used to send the request to the correct data center
- "JSESSIONID" is placed on the visitor's device during the session so the server can identify the visitor
- "Load balancer cookie" (actual cookie name may vary) prevents a visitor from bouncing from one instance to another
Cookies from provider SAPasserviceprovider are required and cannot be turned off
Functional Cookies
These cookies provide a better customer experience on this site, such as by remembering your login details, optimizing video performance, or providing us with information about how our site is used. You may freely choose to accept or decline these cookies at any time. Note that certain functionalities that these third-parties make available may be impacted if you do not accept these cookies.
Consent to all Functional Cookies
Show More Details
Functional Cookies Provider Description Enabled
Vimeo
Vimeo is a video hosting, sharing, and services platform focused on the delivery of video. Opting out of Vimeo cookies will disable your ability to watch or interact with Vimeo videos.
Consent to cookies from provider Vimeo
YouTube
YouTube is a video-sharing service where users can create their own profile, upload videos, watch, like, and comment on videos. Opting out of YouTube cookies will disable your ability to watch or interact with YouTube videos.
Consent to cookies from provider YouTube
Advertising Cookies
These cookies serve ads that are relevant to your interests. You may freely choose to accept or decline these cookies at any time. Note that certain functionality that these third parties make available may be impacted if you do not accept these cookies.
Consent to all Advertising Cookies
Show More Details
Advertising Cookies Provider Description Enabled
Google Analytics
Google Analytics is a web analytics service offered by Google that tracks and reports website traffic.
Consent to cookies from provider GoogleAnalytics
Google Tag Manager
Google Tag Manager is a tag management system for conversion tracking, site analytics, remarketing, and more.
Consent to cookies from provider GoogleTagManager
LinkedIn
LinkedIn is an employment-oriented social networking service. We use the Apply with LinkedIn feature to allow you to apply for jobs using your LinkedIn profile. Opting out of LinkedIn cookies will disable your ability to use Apply with LinkedIn.
📌 Senior Technical Architect (India)
🏢 HCLTech
📍 India