AI Kernel EngineerHybridSoftware EngineeringFull timeBurlingame, California, United StatesShare this job DescriptionQuadric has created an innovative general purpose neural processing unit (GPNPU) architecture. Quadric's co-optimized software and hardware is targeted to run neural network (NN) inference workloads in a wide variety of edge and endpoint devices, ranging from battery operated smart-sensor systems to high-performance automotive or autonomous vehicle systems. Unlike other NPUs or neural network accelerators in the industry today that can only accelerate a portion of a machine learning graph, the Quadric GPNPU executes both NN graph code and conventional C++ DSP and control code.RoleThe AI Kernel Engineer in Quadric plays the key role to enable a large number of AI kernels/operators to run efficiently on the Quadric platform. The AI Kernel Engineer at Quadric will [1] develop a highly efficient Quadric kernel library for a variety of AI/LLM models; [2] analyze the performance and optimize the kernel for different hardware configurations; This senior technical role demands deep knowledge of hardware architecture, compiler toolchain and optimization techniques.Our preference is for a candidate located in the California Bay Area who can regularly collaborate from our Burlingame office. This role follows a hybrid schedule with at least two in-office days per week expected, but actual schedule may adjust depending on team and business need. We believe strong technical collaboration, rapid iteration, and shared problem-solving are well supported by working together in person. The team and company also gather periodically for onsite meetings and offsite events to connect, collaborate, and align on priorities.ResponsibilitiesDevelop AI/LLM kernels/operators on Quadric platform for efficient inferenceOptimize the kernel performance for different hardware configurations and workloadsProfile and analyze kernel performance in terms of compute, data and parallelism; identify micro-architecture and software bottlenecks and provide optimization solutionsOptimize kernel C/C++ codes, maximize hardware utilizationCollaborate across related areas of the AI inference stack to support team and business prioritiesMake Improvement to Quadric toolchain, compiler and runtimeProvide technical support and documents to customers and developer communityRequirementsBachelor’s or Master’s in Computer Science and/or Electric Engineering5+ years of experience in AI kernel development and optimizationexperience with model and kernel inference performance profilingexperience with at least one of the following compute development: CUDA, DSP, NEON, Triton-langProficiency in C/C++ and Python, experience with assembly language a plusDemonstrate good capability in problem solving, debug and communicationBenefitsAt Quadric, we value Integrity, Humility, and Happiness. What we expect from one another is simple and clear: Initiative, Collaboration, and Completion. We are a collaborative team focused on building something extraordinary in the edge computing space. Competitive salary and meaningful equityMedical, dental, and vision plan options starting on day one401(k) retirement planFlexible paid time off (unlimited, non-accrual) to support work-life balanceWhen working in-office, enjoy company-provided lunches and a stocked kitchenConvenient office location within walking distance of the Caltrain stationSupport for commuting, including monthly parking or Caltrain passesDowntown Burlingame office location, close to shops, cafes, and local amenitiesA politics-free, highly collaborative environment where talented people can do their best work and make an immediate impactThe opportunity to build long-term career relationships in a company that values strong personal connections alongside professional excellenceThe base salary range for this position is $110,000 to $270,000. This range reflects the full span of levels and geographies at which Quadric hires for...
Want jobs like this matched to you?
Swoopd scores fresh postings against your résumé so you only see the matches that matter.