Beam is a large language model that combines pretraining and reinforcement learning to achieve strong performance in coding, agentic, and reasoning tasks. It was trained on 23.8 trillion high-quality tokens from the web and proprietary licensed datasets using a novel scaling recipe for reinforcement learning.