Nathan Lambert: Thinking Machines just released with a ~1T param, 41B active, apache-2 model Benchmarks are a clear step up from Nemotron Ultra (55B active), new best American model, and omni input. A bit behind G…
Nathan Lambert: A big day for open model supporters in the U.S. Blog: thinkingmachines.ai/news/introdu... Model card: huggingface.co/thinkingmach...
Ben Recht: Yeah, that part is weird, no? The small model is *much* smaller and yet indistinguishable from the big one on benchmarks? I guess we'll have to wait and see, as it's not available yet.