Modular Logo

Modular

MAX Kernel Engineering Manager

Posted Yesterday
Be an Early Applicant
Remote
Hiring Remotely in United States
248K-372K Annually
Mid level
Remote
Hiring Remotely in United States
248K-372K Annually
Mid level
Lead and grow a team of approximately 12 kernel engineers developing high-performance inference kernels across NVIDIA, AMD, and ASIC hardware. Own the MAX kernel roadmap, performance optimization, portable kernel architecture, and delivery. Partner with framework, serving, compiler/runtime, hardware enablement, and Qualcomm teams while advancing AI-assisted development, verification, and engineering practices.
The summary above was generated by AI
About Modular
At Modular, a Qualcomm company, we’re on a mission to revolutionize AI infrastructure by systematically rebuilding the AI software stack from the ground up. Our team, made up of industry leaders and experts, is building cutting-edge, modular infrastructure that simplifies AI development and deployment. By rethinking the complexities of AI systems, we’re empowering everyone to unlock AI’s full potential and tackle some of the world’s most pressing challenges.

If you’re passionate about shaping the future of AI and creating tools that make a real difference in people’s lives, we want you on our team. You can read about our culture and careers to understand how we work and what we value.

About the role:

MAX is Modular's inference tech stack, built to run GenAI models fast across NVIDIA, AMD, Qualcomm, Trainium, TPU and more. The MAX Kernel team is the performance foundation of that stack, developing kernels using Mojo, GEMM, attention (MHA/MLA/MSA), MoE routing and grouped matmul, and the portable tile abstractions that let one kernel lower to many targets.

As the Engineering Manager of the MAX Kernel team, you will lead a team of kernel engineers, own the kernel roadmap and the per-model × per-hardware performance optimization to the best vendor and open-source stacks. You will partner closely with the Framework, Serve, Compiler/Runtime and Hardware Enablement teams, and help grow a kernel engineering community that spans Modular and Qualcomm.

LOCATION: Candidates based in the United States are welcome to apply. You can work in our office in Los Altos, CA or remotely from home. Onboarding for new hires is conducted in-person in our Los Altos, CA office.


What you will do:

  •  Lead, hire, coach and grow a team of ~12 kernel engineers across NVIDIA, AMD and ASIC hardwares; own career development, performance management and team health.
  • Own the MAX kernel roadmap and delivery: match or beat the performance of other open-source kernel libraries on each business essential model × hardware target.
  • Support portable kernel architecture (TileTensor / TensorEngine / TileIO, reusable building blocks such as MegaFFN and the attention family) together with tech leads.
  • Partner cross-functionally with Framework, Serve, Compiler/Runtime, Hardware Enablement and Qualcomm kernel teams on priorities, interfaces and escalations.
  • Advance AI-assisted kernel development (kernel agents, fuzz verification, playbooks) and the kernel engineering community across org boundaries.


What you bring to the table:
 
  •  3+ years as an engineering manager leading GPU kernel, compiler, HPC, or ML performance engineering teams. (minimum requirement)
  • Strong cross-functional communication; able to drive technical decisions with tech leads and communicate status, risks and tradeoffs to engineering leadership and customers.
  • Track record of hiring, retaining and growing senior kernel or performance engineers.
  • Working knowledge of writing and optimizing GPU/accelerator kernels (CUDA, HIP/ROCm, Triton, CUTLASS/CuTe or similar).
  • Deep understanding of profiling, benchmarking, and roofline analysis


Helpful, but not required:
 
  • Experience with Mojo or other kernel and tile-level programming models (Triton, TileLang, CuTe)
  • Multi-target experience beyond NVIDIA: AMD, NPUs/ASICs or edge/on-device.
  • Experience with AI-assisted or agentic kernel development.
  • Open-source contributions to kernel libraries or inference engines.


What Modular brings to the table:

  • Amazing Team. We are a progressive and agile team with some of the industry’s best engineering and product leaders.
  • World-class Benefits. In order to attract the best, we need to offer the best. Your benefits package may include comprehensive healthcare coverage, retirement and savings programs, employee stock purchase opportunities, paid time off, wellbeing resources, family support programs, and learning and development opportunities. Please note that specific benefit packages may vary based on your location, you can read more about benefits offered by Qualcomm here.
  • Competitive Compensation. We offer very strong compensation packages, including RSU grants. We want people to be focused on their best work and believe in tailoring compensation plans to meet the needs of our workforce. 
  • Team Building Events. We organize regular team onsites and local meetups in Los Altos, CA as well as different cities. Traveling 2-4 times a year is expected for all roles. 

Working at Modular will enable you to grow quickly as you work alongside incredibly motivated and talented people who have high standards, possess a growth mindset, and a purpose to truly change the world.
 
The estimated base salary range for this role to be performed in the US is $248,000.00 - $372,000.00 USD. 

The salary for the successful applicant will depend on a variety of permissible, non-discriminatory job-related factors, which include but are not limited to education, training, work experience, business needs, or market demands. This range may be modified in the future. The total compensation for a candidate will also include annual target bonus, equity, and benefits, with equity making up a significant portion of your total compensation.
For candidates who fall outside of the listed requirements, we nevertheless encourage you to apply as we may have upcoming openings that are lower/higher level than the ones advertised. 

Similar Jobs

3 Hours Ago
Remote or Hybrid
45K-85K Annually
Junior
45K-85K Annually
Junior
Artificial Intelligence • Fintech • Insurance • Marketing Tech • Software • Analytics
Handles inbound calls and warm leads, consults customers on insurance needs, recommends appropriate Property and Casualty coverage, and converts prospects into policyholders. The role includes paid licensing and training, customer communication, sales closing, and adherence to remote-work requirements. Employees must work four weekdays and one weekend day, maintain a dedicated home workspace, and remain in their resident state for at least one year due to licensing restrictions.
Top Skills: Cable InternetDsl InternetFiber InternetPcWired High-Speed Internet
6 Hours Ago
In-Office or Remote
132K-208K Annually
Senior level
132K-208K Annually
Senior level
Cloud • Information Technology • Productivity • Security • Software • App development • Automation
Provides strategic technical consulting to enterprise customers using Atlassian Cloud, AI, and related solutions. Responsibilities include solving business challenges, developing prescriptive technical guidance, supporting cloud migrations, designing AI proof-of-concept solutions, promoting service expansion, advocating for customer needs, and collaborating across internal teams. The role requires customer-facing expertise, strong knowledge of Atlassian ecosystems and SaaS architectures, and up to 30% domestic and international travel.
Top Skills: Ai AgentsAtlassian CloudCloud MigrationConfluenceFocusGuardHybrid SaasJira Service ManagementJira SoftwareLarge Language ModelsOn-Premises DeploymentsPrompt EngineeringRovoSaas Architectures
6 Hours Ago
In-Office or Remote
124K-195K Annually
Entry level
124K-195K Annually
Entry level
Cloud • Information Technology • Productivity • Security • Software • App development • Automation
Designs and maintains SaaS integrations, middleware, APIs, webhooks, automation scripts, and data pipelines. Optimizes enterprise software usage, supports AI integrations, troubleshoots complex technical issues, manages vendors and SLAs, tracks licensing and budgets, and translates business needs into technical requirements.
Top Skills: Anthropic ApisAPIsAtlassian JsmAutomation WorkflowsCursor ApisData PipelinesMiddlewareOpenai ApisSaaSWebhooks

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account