Pre-Training vs. Fine-Tuning
5 MIN READ
Two phases. One brain. Here's how a model goes from clueless to expert.
Related Reads
Mixture of experts: How AI models got smarter without getting slower
The secret architecture behind trillion-parameter models that stay fast and affordable.
Steering vectors: the DJ mixer inside every LLM
Prompting talks to the model. Steering vectors operate inside it.
Tokenization: How AI Reads Text
AI doesn't read words. It reads bricks of text called tokens.