Microsoft published a 109-page technical report on MAI-Thinking-1. Here’s the abbreviated version of how a modern lab actually trains a frontier reasoning model — from scraping the web to reinforcement learning, judges, and anti-cheating. The post How do you make an LLM, anyway?