Skip to content
The day in AI
Mon 9 Jun 2025
Monday 9 June 2025
←
→
No written reading for this day yet. Here is everything that was logged.
Everything from this day
Research
9 Jun 2025
Microsoft Research / Peking / Tsinghua
Reinforcement Pre-Training (RPT)
Recasts next-token prediction as a reasoning task with verifiable reward (did the token match?), scaling RL over ordinary web text.