Lead Software Engineer - Kernels
MatX
<p><strong>What MatX Is Building</strong></p> <p>MatX is on a mission to be the compute platform for AGI. We are developing vertically integrated full-stack solutions from silicon to systems including hardware and software to train and run the largest ML workloads for AGI.&nbsp;</p> <p>We are seeking a hands-on engineering TLM to lead the Kernel team. You will manage, mentor, and grow a team of kernel engineers while setting technical direction and executing. At the same time, you will be an active coding contributor by designing and implementing high-performance compute kernels for specialized AI hardware.</p> <p>You will partner with the ML team to maximize output and improve diagnostic tools, and collaborate with hardware to co-design next-generation AI architectures.</p> <p><strong>What You'll Do Here</strong></p> <ul> <li>Manage, mentor, and build a team of highly specialized kernel engineers and software developers</li> <li>Design and optimize kernels that interface directly with our hardware</li> <li>Work in partnership with our ML Research and Hardware Engineering teams</li> <li>Provide expertise and guidance on hardware architecture from a programmer's perspective, ensuring seamless integration with the software stack</li> </ul> <p><strong>Who You Are</strong></p> <ul> <li>2+ years of engineering team management experience</li> <li>7+ years of working directly within engineering teams experience</li> <li>Bachelor of Computer Science or equivalent degree</li> <li>Experience optimizing software for specialized hardware, employing techniques such as parallelism, SIMD programming, C, assembly-level optimization, or GPU/CUDA programming</li> <li>Language: at least one of assembly, C++, C, Zig, or Rust</li> <li>Ability to navigate ambiguity, manage escalations, and align stakeholders around infrastructure decisions.</li> <li>Excellent communication and collaboration skills, with a track record of translating complex technical trade-offs with the founders and stakeholders</li> </ul> <p><strong>Bonus Points If You Have</strong></p> <ul> <li>Experience implementing kernels for ML models such as Transformers</li> <li>Experience using and implementing distributed parallelism techniques such as AllReduce, AllToAll, data parallelism, tensor parallelism.</li> <li>Familiarity with how compilers work</li> </ul> <p><strong>Compensation</strong></p> <p>The US base salary for this full-time position is determined based on a variety of factors including role, experience, location, job related skills, and relevant education and training. Career length is only a guideline for compensation.</p> <ul> <li>Early Career - $120,000 - $275,000 + equity</li> <li>Mid Career - $175,000 - $400,000 + equity</li> <li>Senior Career - $250,000 - $600,000 + equity</li> </ul> <p><strong>What We Offer</strong></p> <ul> <li><strong>A Stake in our success&nbsp;</strong>A flexible cash equity compensation mix that fits your needs</li> <li><strong>Health &amp; Wellness </strong>Company subsidized Health, Dental, Vision, and Life insurance; Pre-tax Health Savings Accounts with generous company contribution (even if you don’t)</li> <li><strong>Time To Recharge&nbsp;</strong>4 weeks paid time off (accrued), 12 company holidays, and 3 weeks remote/flexible work per year</li> <li><strong>Support to Parents</strong>&nbsp;Up to 12 weeks of paid parental leave, regardless of your path to parenthood</li> &l