writing about language models, CUDA, and systems engineering
Developing a clean C++/CUDA library for token generation, covering page tables, virtual memory metaphors for GPU allocation, and block size tradeoffs.
~ ~ ~ ~ ~ ~ ~
© 2026 Mayank Joshi