samankeon.com

Blog

Notes on how LLMs work under the hood, and a hands-on series on writing GPU kernels.

#benchmarking#cuda-kernel#image-processing#kernels#llm#positional embedding#rope#triton-kernel