-
Contributing to open-source software in 2026
-
Trip to Kyoto
-
Multi Token Prediction in llama.cpp
-
Prefetching Weights in llama.cpp
-
Simple Loop for Auto-Anything using LLMs
-
Every LLM hallucinates that std::vector deletes elements in a LIFO order
-
Optimizing Token Generation in llama.cpp's CUDA Backend
-
A gentle introduction to GEMM using MMA tensor cores
-
Creating a git repo for your life
-
Why "good first issues" are usually not good first issues