Location: Whitefield, Bengaluru (3 days work from office) Job Summary
We are looking for a GPU Computing Engineer with solid expertise in CUDA programming, GPU architecture, and parallel computing. The role involves building and optimizing high-performance solutions for AI, simulation, and large-scale computing workloads. Key Responsibilities
Develop and optimize CUDA-based applications and GPU kernels.
Design scalable parallel computing solutions for compute-intensive workloads.
Optimize application performance through profiling, benchmarking, and tuning.
Build and support multi-GPU and distributed computing environments.
Implement efficient workload scheduling and resource utilization strategies.
Collaborate with AI/ML,
platform, and infrastructure teams.
Required
Skills
CUDA C/C++
NVIDIA GPU Architecture
Parallel Computing & Multithreading
Linux
Python
Performance Optimization
OpenMP, MPI, NCCL
Scheduling Algorithms Preferred Skills
Kubernetes
Slurm Desired Profile
5-12 years of experience in GPU Computing, HPC, AI Infrastructure, or Performance Engineering.
Strong understanding of parallel processing and distributed computing concepts.