AI Software Lead – PyTorch & CUDA Runtime (Next-Gen Accelerator) 10 years (Bengaluru)

AI Software Lead – PyTorch & CUDA Runtime (Next-Gen Accelerator) 10 years (Bengaluru)

06 Aug
|
Sandisk
|
Bengaluru

06 Aug

Sandisk

Bengaluru

Job Description

Role Overview

We are looking for a Software Lead (8+ years’ experience) to own the runtime and neural network (NN) layer of a next-generation AI accelerator platform. This role focuses on designing, optimizing, and implementing NN operators and developing new ops using CUDA/custom runtime APIs to deliver high-performance execution on custom AI hardware.

Key Responsibilities

- Design and optimize NN operators for performance-critical workloads
- Develop recent NN ops using CUDA/custom runtime APIs
- Drive runtime-level optimizations across compute, memory, and scheduling
- Own runtime ↔ NN layer interfaces and execution model
- Implement and optimize operator fusion (e.g., matmul + bias + LayerNorm) for efficient hardware utilization
- Identify and resolve performance bottlenecks across the stack
- Collaborate with compiler, PyTorch framework, and low-level SW teams

Impact

- Own how efficiently AI workloads execute on the platform
- Drive performance, scalability, and hardware utilization through optimized runtime and NN ops design

📌 AI Software Lead – PyTorch & CUDA Runtime (Next-Gen Accelerator) 10 years (Bengaluru)
🏢 Sandisk
📍 Bengaluru

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: ai software lead – pytorch & cuda runtime (next-gen accelerator) 10 years (bengaluru) / bengaluru

Subscribe to this job alert:

Get the latest job offers by email for: ai software lead – pytorch & cuda runtime (next-gen accelerator) 10 years (bengaluru) / bengaluru