02 — Blog

Notes from the learning process.

Writings about research, people, places, and things I want to remember.

LLM Optimizers: From SGD and AdamW to SOAP, Muon, and Scalable Training

LLM ↗

Tokenization: From Building BPE from Scratch to Language Models Without a Fixed Tokenizer

LLM ↗

Welcome to my personal website!

Life ↗