





Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Metro location and mid-level (2–6yrs) amplify competition despite niche DSP specialization.
Role requires niche DSP and low-level optimization experience, limiting cross-industry transferability.
Explicit 2–6 year requirement plus specialized DSP, SIMD, and accelerator expertise narrows candidate pool.
Develop and optimize AI inference kernels in C/C++ across various data types such as INT8, INT16, BF16, and FP32.
Perform SIMD/vectorization, Intrinsics use, memory optimization, and detailed performance tuning including profiling and debugging.
Collaborate with hardware and software teams to integrate and enhance AI inference pipeline performance.
2+ years of professional experience in software development or related role.
Proficient in C/C++ programming with practical knowledge of AI/Deep Learning inference concepts.
Experience working with DSPs, hardware accelerators, or similar compute platforms involving SIMD/vectorization and low-level software optimization.
Location initially in Bangalore with hybrid work mode afterward.
Experience specifically in AI/ML inference optimization and kernel performance tuning.
Familiarity with hardware accelerators such as C7x DSP, HWA-MMA or similar platforms.
Strong skills in profiling, debugging, and performance analysis targeting heterogeneous computing and memory optimization.