<aside> <img src="/icons/city_gray.svg" alt="/icons/city_gray.svg" width="40px" /> Makora, Inc.

</aside>

<aside> <img src="/icons/row_gray.svg" alt="/icons/row_gray.svg" width="40px" /> Table of Contents

</aside>

<aside> <img src="/icons/flash_gray.svg" alt="/icons/flash_gray.svg" width="40px" /> Apply Now!

</aside>

Summary

Our R&D team is seeking expert level GPU kernel engineers to help build the world’s best LLMs and Agents for GPU kernel generation.

The goal is simple: design an AI agent that writes and optimizes kernels in the same way you do. You will collaborate with the training team to define robust evaluation, validation, and reward models that will be used to train LLMs in the art of GPU kernel engineering. You will also contribute to the AI agent architecture itself, defining the workflows that enable an LLM to discover and implement high performance GPU kernels.

This job is based in either Gdansk or New York City. Remote work will be considered for exceptional candidates.

About Makora

Makora is a venture-backed AI infrastructure company building the optimization software and inference platform that make frontier AI models run faster, cheaper, and more efficiently across any hardware. There are three core components:

  1. MakoraGenerate writes GPU kernels in CUDA, HIP, and Triton using LLMs
  2. MakoraOptimize automatically selects and swaps GPU kernels in combination with tuning inference engine (vLLM, SGlang, etc..) hyperparameters to optimize performance
  3. MakoraInference uses Generate and Optimize to deliver frontier open-source models as ultra-fast inference endpoints.

Responsibilities

Qualifications

Bonus Points

Our Benefits

To Apply

Fill out this form