Blog
Porting, proving and keeping kernels fast on any chip.
More posts
AI writes the code fast. A signed Ed25519 proof pack shows it runs right on the real chip, and you can check it offline.
Read more
Whole model ports keep 99 percent of the task score for FP8 and INT8. Here is how a pass is defined.
Read more
Porting across CUDA, HIP, NKI, Pallas and Triton. The engine finishes what hipify leaves.
Read more
Rechecks on ROCm, CUDA, Neuron and firmware releases, with a CI gate that warns first and blocks only when you choose.
Read more
Short notes on kernels, chips and proof. One email when a post goes up.