Bengaluru, India On-site Full-time

Sandisk is hiring an AI Software Lead – PyTorch & CUDA Runtime (Next-Gen Accelerator) 10+ years

Role Overview We are looking for a Software Lead (8+ years’ experience) to own the runtime and neural network (NN) layer of a next-generation AI accelerator platform. This role focuses on designing, optimizing, and implementing NN operators and developing new ops using CUDA/custom runtime APIs to deliver high-performance execution on custom AI hardware. Key Responsibilities - Design and optimize NN operators for performance-critical workloads - Develop new NN ops using CUDA/custom runtime APIs - Drive runtime-level optimizations across compute, memory, and scheduling - Own runtime ↔ NN layer interfaces and execution model - Implement and optimize operator fusion (e.g., matmul + bias + LayerNorm) for efficient hardware utilization - Identify and resolve performance bottlenecks across the stack - Collaborate with compiler, PyTorch framework, and low-level SW teams Impact - Own how efficiently AI workloads execute on the platform - Drive performance, scalability, and hardware utilization through optimized runtime and NN ops design
Job Details
Location Bengaluru, India
Work mode On-site
Employment Full-time
Department Firmware Engineering
Category other
Posted 2 months ago
or drop your CV first
About company
Sandisk logo
Sandisk innovates in Flash and advanced memory technologies, delivering solutions that enable digital world needs with groundbreaking memory products recognized globally for performance and quality.
All jobs at Sandisk Visit website