Kernelize provides software for AI inference across CPUs, GPUs, and NPUs.
Kernelize builds compiler backends for AI accelerators to enable and optimize Triton kernels for broad adoption in ML training and inference. The Runtime Engineer will build a scalable runtime supporting a wide range of AI hardware accelerators, drive ML optimization using Triton, and help shape core runtime technologies.