
model trainingdeep learningscalingAI agentsPost-Training19 min read
Reasoning Models: How LLMs Learned to Think Before They Speak
Explore how reasoning models like o1, o3, and DeepSeek-R1 use inference-time compute scaling and chain-of-thought to solve problems standard LLMs cannot.
RayZAPR 6, 2026

