ModalTools
Reinforcement learning is an infrastructure problem

What we've seen helping teams run Reinforcement Learning at scale on Modal. Plus an open-source library to skip the scaffolding.
Read at Modal ↗Related

ResearchDiverse reasoning traces teach LLMs to make better decisions Amazon Science

VoicesRecent Developments in LLM Architectures: KV Sharing, mHC, and Compressed Attention Ahead of AI

LabsIntroducing AIMIP: The AI weather and climate model intercomparison project Ai2

VoicesMy Workflow for Understanding LLM Architectures Ahead of AI

VoicesOpen-world evaluations for measuring frontier AI capabilities AI Snake Oil
