Reflection Previews Beam AI Coding Model, Optimized for Efficiency


TL;DR

  • Beam Preview: Reflection AI unveiled Beam for coding and tool-using work, offering selected users access before its public weight release.
  • Coding Results: Company benchmarks put Beam ahead of Western rival Inkling, with mixed results against Chinese model GLM-5.2.
  • Reasoning Control: Beam’s lower effort settings favor shorter responses; higher settings allow longer reasoning for demanding tasks.
  • Open Weights: Reflection promises downloadable model parameters and an Apache 2.0 license later in October, enabling developers to run and adapt Beam themselves.

Reflection AI unveiled Beam, its first model for writing code and carrying out tasks with software tools, on October 5, 2026. The startup is targeting businesses and governments that want a Western-built AI system they can run and customize themselves, but that control still awaits its promised release of downloadable model parameters. Selected users receive an early version, and others can request access through a waitlist.

Reflection’s chief executive Misha Laskin told Semafor that Reflection sees demand from organizations unwilling or unable to use Chinese models. Beam’s immediate pitch is a combination of coding capability and lower inference compute, the processing work needed to generate an answer.

Coding Gains and Current Alternatives

Reflection’s scorecard gives Beam 80.1 on Terminal Bench v2.1, a test of tasks performed through a computer’s command-line interface, against 63.8 for Inkling from Thinking Machines Lab. On SWE Bench Pro v1, which tests software-engineering work, Beam scores 65.5 against Inkling’s 54.3.

The comparison with Z.ai’s GLM-5.2 is closer: Beam’s 65.5 exceeds GLM-5.2’s 62.1 on SWE Bench Pro v1, while its 80.1 trails 81.0 on Terminal Bench v2.1. GLM-5.2 already offered public weights and local deployment options in June, giving developers an existing model they could run themselves.

Inkling’s developers used an internal coding harness, the software that supplies tools and runs the model’s task attempts, for Terminal Bench. Reflection’s technical report remains promised for later in October.