Ai2

Why they matter

Tracked through public AI activity and peer connections inside the directory.

AI signals
10
past 30d
Sources
4
distinct domains
Discussões
0
past 30d
Latest signal
11d ago
View every signal from Ai2 →
Breakthrough AI to solve the world's biggest problems. › Join us: http://allenai.org/careers › Get our newsletter: https://share.hsforms.com/1uJkWs5aDRHWhiky3aHooIg3ioxm

Articles & links

Our fully open releases give researchers the data, code, checkpoints, and methods they need to inspect claims, reproduce findings, and advance new science. Read more about why that’s so important to us. ⬇️ allenai.org/blog/who-get...

Who gets to understand AI? | Ai2 allenai.org
AI Weekly's analysis
  • Ai2 argues meaningful AI transparency requires not just open weights but training data, code, methods, checkpoints, evaluations, and documentation.
  • Post cites three studies enabled by open Olmo releases, covering clinical demographic bias, benchmark inflation, and how models reason about drug names.
  • Without that access, Ai2 warns, technical direction of the field risks becoming concentrated inside a small number of companies.
Read full analysis →
View on Bluesky · ♥ 4 ↻ 1 ↩ 0 · 2 from the directory shared this · 24d ago

Try olmOCR 2 in the Ai2 Playground, check out our blog for more info, & download the weights and data from Hugging Face: ▶️ Playground: buff.ly/pJhoWXN 📝 Blog: buff.ly/8uUabID 🤗 Model & data: buff.ly/EUYHIuN

playground.allenai.org
View on Bluesky · ♥ 1 ↻ 1 ↩ 0 · 38d ago

The pre-training dataset for OlmoEarth includes Sentinel-2, Sentinel-1, and Landsat satellite imagery, paired with various "maps" such as ESA WorldCover and the USDA Cropland Data Layer. For details, see: - Hugging Face: huggingface.co/datasets/all... - OlmoEarth v1 paper: arx…

arxiv.org
View on Bluesky · ♥ 1 ↻ 0 ↩ 1 · 49d ago

Recent commentary

LLMs are no longer created w/ human data alone. They rely on other models to generate & filter data, evaluate outputs, & guide dev work. So what is a modern LLM built on? Olmo 3 → 89 model + 183 dataset dependencies; Nemotron 3 → 273 + 560 We made ModSleuth to trace this. 🧵

View on Bluesky · ♥ 53 ↻ 11 ↩ 1 · 68d ago

Today we're introducing a preview of TutorMoments, a framework that measures whether AI tutors can make one of the hardest calls in teaching: when to step in and help a student, & when to hold back and let them do the heavy thinking. 🧵

View on Bluesky · ♥ 51 ↻ 6 ↩ 2 · 11d ago

We built SciArena to test how well AI models handle scientific literature questions, as judged by researchers. It's retiring July 15, and the results are in: ~1,700 users cast ~3,900 votes. Here's what they told us. 🧵

View on Bluesky · ♥ 9 ↻ 5 ↩ 2 · 33d ago

AI image generators don't "draw"—they follow a compass: the score function, which points toward more probable images. The same compass drives Bayesian sampling and plasma physics. We built DiScoFormer to estimate the score far better when data gets complex. 🧵

View on Bluesky · ♥ 14 ↻ 1 ↩ 2 · 49d ago

When a model writes, where do its words come from? Are they new, or do they match exactly with language it saw in training? An AI-writing detector can't tell you. @tuhinchakr.bsky.social's group at Stony Brook has been dissecting AI-generated prose with our infini-gram engine. 🧵

View on Bluesky · ♥ 13 ↻ 1 ↩ 1 · 17d ago

Two updates to Asta, our ecosystem of AI agents for science: a one-click handoff from AutoDiscovery to Asta’s data analysis tools, & paper search that evaluates its own results + searches again when they fall short. 🧵

View on Bluesky · ♥ 10 ↻ 2 ↩ 1 · 27d ago

Building an LLM means evaluating it over & over as it changes. Tweak a hyperparameter or scale the model up, & every new checkpoint sends you back through the same benchmarking loop. We're releasing olmo-eval, a workbench built for this kind of iterative model development. 🧵

View on Bluesky · ♥ 8 ↻ 3 ↩ 1 · 67d ago

As a nonprofit research institute dedicated to advancing open science, we're encouraged to see growing support for open models across the AI ecosystem. We believe the evidence behind advanced AI systems shouldn’t be locked up in a few hands.

View on Bluesky · ♥ 9 ↻ 1 ↩ 1 · 24d ago

What does it actually take to build cutting-edge AI systems? On July 30 during #SeattleTechWeek, the researchers behind Ai2's open models sit down to talk through the deep technical work behind them. 🧵

View on Bluesky · ♥ 9 ↻ 0 ↩ 2 · 31d ago

What can you build with a fully open robotics model in a weekend? 🤖 Robotics engineer @0xbinh.bsky.social used MolmoAct 2, our open vision-language-action model, in the voice-controlled robot that won @southparkcommons.bsky.social's AI hackathon. Watch our interview with him ↓ 🎥

View on Bluesky · ♥ 13 ↻ 0 ↩ 0 · 40d ago

In Ai2's orbit

Center = Ai2. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.