↻
hardmaru reposted
I am incredibly proud of our Tokyo team for shipping this. By orchestrating the world’s models, we are delivering the resilient blueprint required for AI sovereignty. Read our full vision and results here: sakana.ai/fugu-release 🐡
Sakana AI sakana.ai
View on Bluesky ·
♥ 12
↻ 0
↩ 0
·
6 from the directory shared this ·
99d ago
“World Models in Natural and Artificial Intelligence” brings together pioneers including Douglas Hofstadter, Michael Levin, Josh Tenenbaum, Samuel Gershman, and Melanie Mitchell to ask: What if the next leap in AI requires not just more data, but systems that model themselves?
royalsocietypublishing.org
↻
hardmaru reposted
Sakana AI
@sakanaai.bsky.social
Fugu-Ultra is now live on OpenRouter! ⚡ We share a core vision with the OpenRouter team: the future of AI isn’t a single monolithic model, but the collective intelligence of the world’s best models working together. Try it: openrouter.ai/sakana/fugu-... 🐡
Fugu Ultra - API Pricing & Providers openrouter.ai
AI Weekly's analysis
→
Read full analysis →
View on Bluesky →
↻
hardmaru reposted
By orchestrating the world's models, we are building the resilient infrastructure required for AI sovereignty. Try: sakana.ai/fugu Blog: sakana.ai/fugu-max-rel... 🐡
Sakana Fugu — Multi-agent System as A Model sakana.ai
AI Weekly's analysis
→
- Fugu routes tasks through a dynamic multi-agent pipeline exposed as a single OpenAI-compatible API, removing orchestration setup from users.
- The system draws on two ICLR 2026 papers: TRINITY assigns Thinker/Worker/Verifier roles; Conductor uses reinforcement learning to design coordination strategies.
- Fugu Ultra scored 73.7 on SWE Bench Pro and 93.2 on LiveCodeBench; base Fugu reached 95.5 on GPQA-D, per Sakana's own benchmark reporting.
Read full analysis →
By orchestrating the world's models, we are building the resilient infrastructure required for AI sovereignty. Try: sakana.ai/fugu Blog: sakana.ai/fugu-max-rel... 🐡
Sakana AI sakana.ai
AI Weekly's analysis
→
- Fugu Max prices at $2 per million input tokens and $6 per million output tokens, which Sakana claims runs 40-60% below Sonnet 5, GPT 5.6 Terra, and Kimi K3.
- The system routes each task across a swappable pool of open-weight and specialized models, with NVIDIA's Nemotron folded in via an August 2026 collaboration.
- Fugu Ultra v2 scores 48.3 on Chartography against Opus 5's 27.3, and does so without Fable 5, Fable 5.1, or GPT-6-Astra in its agent pool.
Read full analysis →
We just pushed a big update to Sakana Chat Free to use: chat.sakana.ai A big motivation for this release is getting people in Japan, especially kids, excited about software development. Vibe-coding turns Kanji & Math drills during the summer break into fun games they can actua…
Sakana Chat chat.sakana.ai
AI Weekly's analysis
→
Read full analysis →
View on Bluesky ·
♥ 15
↻ 1
↩ 0
·
3 from the directory shared this ·
47d ago