OLMo
by Allen Institute for AI · allenai.org ↗
Fully open large language model family from Ai2, released with its training data, code, and weights.
open-source llm research
- Category
- Open Models
- Business model
- Open (open weights/source)
- Availability
- Global
- Launched
- 2024-02
- Record updated
- 2026-07-30
- Model family
- OLMo
- License
- Apache-2.0
- Founded
- 2014
- Headquarters
- US
- Canonical URL
- https://globalaiproductindex.com/products/olmo/
Overview
OLMo is a fully open large language model family from the Allen Institute for AI (Ai2), released with its complete training data, code, evaluation suites, and weights. Designed for transparent, reproducible research, later generations like OLMo 2 and OLMo 3 are competitive with similarly sized open-weight models. Everything is free under Apache 2.0.
Key features
- Fully open: weights, data, code, and checkpoints
- Dolma training corpus openly released
- Reproducible training and evaluation pipelines
- Apache 2.0 licensing for commercial use
Use cases
- Academic research on LLM training dynamics
- Auditing model behavior against known training data
- Building on a transparent, commercially usable base
Pricing
OLMo models, data, and code are entirely free under Apache 2.0; users bear only their own compute costs.
Frequently asked questions
What makes OLMo different from other open models?
OLMo releases the full stack—training data, code, logs, and weights—not just the final model, enabling true reproducibility.
Who develops OLMo?
The Allen Institute for AI (Ai2), a non-profit research institute in Seattle.
Can OLMo be used commercially?
Yes, everything is released under the permissive Apache 2.0 license.
Similar products
- Baichuan — Open-weight bilingual LLM family from Beijing-based Baichuan Intelligence, now focused on medical AI with its Baichuan-M reasoning models.
- Command R — Cohere's open-weight LLM family optimized for retrieval-augmented generation and tool use; succeeded by Command A as Cohere's flagship.
- DBRX — Open-weight mixture-of-experts LLM released by Databricks in 2024; retired from Databricks serving in 2025, though the weights remain available.
- DeepSeek — Open-weight LLM family and chat assistant from Hangzhou-based DeepSeek, known for strong reasoning models. The official DeepSeek-V4-Flash-0731 API brings a major agent-capability upgrade via post-training, with native Responses API support and Codex adaptation.
- Falcon — Family of open-weight large language models developed by the Technology Innovation Institute (TII).
- Gemma — Google's family of lightweight open-weight models derived from the same research as Gemini.
Sources
This record was last reviewed on 2026-07-30.
Machine-readable record: /api/products/olmo.json · Spot an error? Suggest a correction