Our fully open releases give researchers the data, code, checkpoints, and methods they need to inspect claims, reproduce findings, and advance new science. Read more about why that’s so important to us. ⬇️ allenai.org/blog/who-get...
Who gets to understand AI? | Ai2 allenai.org
AI Weekly's analysis
→
- Ai2 argues meaningful AI transparency requires not just open weights but training data, code, methods, checkpoints, evaluations, and documentation.
- Post cites three studies enabled by open Olmo releases, covering clinical demographic bias, benchmark inflation, and how models reason about drug names.
- Without that access, Ai2 warns, technical direction of the field risks becoming concentrated inside a small number of companies.
Read full analysis →
What's next for OlmoEarth Platform: change-detection alerts, agentic tools for nonexperts, & precomputed global embeddings that could one day skip the forward pass entirely. Read more about the engineering in our latest blog: 🌐 buff.ly/wkAWugp
The OlmoEarth Platform: Geospatial inference at planetary scale | Ai2 allenai.org
This is the work that’s shaping what comes next for OlmoEarth. Learn more: allenai.org/olmoearth
OlmoEarth | Ai2 allenai.org
For a deeper look at what SciArena revealed about how scientists evaluate AI-generated answers to Qs about the scientific literature, read our NeurIPS 2025 Spotlight paper—which includes findings beyond the leaderboard. 📄 buff.ly/A9OrLph
SciArena: An Open Evaluation Platform for Non-Verifiable Scientific Literature-Grounded Tasks arxiv.org
o3 finished on top of the SciArena leaderboard – ahead of Claude Opus 4.1, Gemini 3 Pro Preview, & open-weights models like DeepSeek-R1 – with answers researchers called more detailed + to the point. Learn more in our updated blog: buff.ly/o1EyhhG
SciArena: A new platform for evaluating foundation models in scientific literature tasks | Ai2 allenai.org
Shippy is currently in private preview for select partners. For more on how we built it, what’s next, and how we’re ensuring users can trust and fully trace its outputs, read our new engineering blog: buff.ly/xur4hXq
What building Shippy taught us about building agents | Ai2 allenai.org
Try olmOCR 2 in the Ai2 Playground, check out our blog for more info, & download the weights and data from Hugging Face: ▶️ Playground: buff.ly/pJhoWXN 📝 Blog: buff.ly/8uUabID 🤗 Model & data: buff.ly/EUYHIuN
playground.allenai.org
Modular training is gaining momentum as frontier models become costlier to train & deploy. This project shows how open, distributed approaches can make development more practical for national projects, public institutions, & smaller teams. → Learn more: buff.ly/N0b3zI7
Modular LLMs at scale: How the Danish Foundation Models project is using FlexOlmo to pool national expertise without pooling sensitive data | Ai2 allenai.org
The pre-training dataset for OlmoEarth includes Sentinel-2, Sentinel-1, and Landsat satellite imagery, paired with various "maps" such as ESA WorldCover and the USDA Cropland Data Layer. For details, see: - Hugging Face: huggingface.co/datasets/all... - OlmoEarth v1 paper: arx…
allenai/olmoearth_pretrain_dataset · Datasets at Hugging Face huggingface.co
The pre-training dataset for OlmoEarth includes Sentinel-2, Sentinel-1, and Landsat satellite imagery, paired with various "maps" such as ESA WorldCover and the USDA Cropland Data Layer. For details, see: - Hugging Face: huggingface.co/datasets/all... - OlmoEarth v1 paper: arx…
arxiv.org
Score & density show up across many fields. We hope one pretrained model like DiScoFormer can serve them all, at scale. 📝 Blog: https://t.co/bomZWcgcfy 📄 Report: https://t.co/g9M5oaYY2m
arxiv.org
This update came directly from partners asking for cleaner embeddings. OlmoEarth v1.2 comes in Nano, Tiny, Small, & Base—all open source + available now. 🤗 Models: buff.ly/T7IU2ZD 💻 Training & fine-tuning code: buff.ly/kj0fXJh 📄 Tech report: buff.ly/arhawz1
OlmoEarth - a allenai Collection huggingface.co