If you want to understand GPU for LLMs – why every AI company is spending billions on graphics cards to run large language models – the answer comes down to two things: parallel math and memory. This post explains both, without the handwaving. What is a CPU, and what’s it good at? A CPU – […]
