Tweet by natolambert
September 28, 2025
RL research is becoming like pretraining/modeling. This is a huge vibe shift. Most research published on RL isn't using enough compute to make many of these decisions matter as much. This is slowly shifting. https://t.co/FwhLhsUUZq
- Author
- natolambert
- Date
- September 28, 2025
- Canonical URL
- /tweets/natolambert-1972323625207509321-dedc9d