Reflection, the company that had a rough public debut with its first model, is back with something that looks more deliberate. The company has announced Beam, a text model built for coding, reasoning, and agentic tool use, and is currently accepting applications for early access. Full downloadable weights are expected later in October, pending final safety evaluations.
That Apache 2.0 license is the headline detail here. It puts Beam in direct competition with Meta’s Llama series and Mistral’s open models, which are the current defaults for developers who want to run inference locally, fine-tune freely, or build products without paying per token. If Beam’s actual performance holds up, the license alone makes it worth evaluating.
Reflection says it trained capability and safety separately, which is a meaningful architectural choice. Many labs bake safety constraints into the same training run as capability, which can create tradeoffs that are hard to isolate later. Keeping them distinct at least gives developers a clearer picture of what they’re working with, and gives Reflection more flexibility to update either side without retraining from scratch. The company plans to release a technical report and model card alongside the weights, which is the minimum bar for serious open model releases in 2025.
The tool-use angle matters more than it might seem. Models that can reliably call external APIs, parse structured outputs, and chain actions together are increasingly what production teams actually need. GPT-4o and Claude 3.5 Sonnet are still the benchmarks for agentic reliability, but both are closed and expensive at scale. Open alternatives like Qwen2.5-Coder and DeepSeek-Coder-V2 have closed the gap on raw coding benchmarks, but tool-use consistency remains uneven. If Beam is competitive there, it fills a real gap.
Still, Reflection has credibility to rebuild. Its first public model release drew criticism over benchmark claims and a messy rollout, and the AI developer community has a long memory. The decision to do early access before the full release is probably smart, it gives the company a chance to catch issues before a wider audience scrutinizes the weights.
- Early access applications are open now
- Weights release targeted for late October 2025
- License: Apache 2.0, allowing commercial use and modification
- Technical report, model card, and developer tools included at release
- Separate capability and safety training pipelines
The actual weight release will be the real test. Numbers on a blog post mean less than what developers find when they run evals themselves.



