Tweet by natolambert

September 28, 2025

RL research is becoming like pretraining/modeling. This is a huge vibe shift. Most research published on RL isn't using enough compute to make many of these decisions matter as much. This is slowly shifting. https://t.co/FwhLhsUUZq

Author
natolambert
Date
September 28, 2025