Fine-tuning small language models
Small models have no capacity to waste. Noisy or duplicated examples teach the wrong habits, and synthetic instruction data narrows what they can do.
- Instruction and preference sets written or verified by experts
- Domain slices sized for focused models
- Deduplicated against public benchmarks so evals stay honest
Faster convergence and better generalization per parameter — smaller models that punch above their weight.