pragalva.me
Blog Projects Homelab CV
Blog Projects Homelab CV

Llms

2026-09-22 site

Observing the Probabilities

What I observed while training a tiny language model: the loss stopped falling, rug's embedding never moved, and cat and dog were not as close as I expected.

#technical #llms
2026-08-21 site

Viewing Language Through a Probabilistic Lens

How next-word prediction, distributed representations, and a neural probability model let useful structure emerge from language.

#technical #llms
2026-06-25 site

Working With Text Data

How raw text becomes something an LLM can actually train on: tokenizing, building a vocabulary, byte pair encoding, sliding windows, and turning token IDs into token and positional embeddings.

#technical #llms
2026-06-08 site

Understanding LLMs

A foundational look at large language models, recurrent and convolutional architectures, transformers, and the difference between BERT and GPT.

#technical #llms