↻
Miguel Alonso Jr. reposted
↻
Miguel Alonso Jr. reposted
Sung Kim
@sungkim.bsky.social
GLM-5.3 is now open-weight. Their most capable model for agentic coding and cyber defense is now available to download, run, and customize. Weights: huggingface.co/zai-org/GLM-5.3 Tech blog: z.ai/blog/glm-5.3
zai-org/GLM-5.3 · Hugging Face huggingface.co
View on Bluesky →
↻
Miguel Alonso Jr. reposted
↻
Miguel Alonso Jr. reposted
Tim Kellogg
@timkellogg.me
if you're having trouble with Astra, it's probably old skills. Point your agent at these two URLs and use them to upgrade skills: 1. developers.openai.com/blog/rethink... 2. github.com/openai/codex...
codex/codex-rs/skills/src/assets/samples/skill-creator/SKILL.md at main · openai/codex github.com
View on Bluesky →
↻
Miguel Alonso Jr. reposted
Tim Kellogg
@timkellogg.me
I know you're all fed up. Employees at frontier labs have gotten to sign like 3 open letters just this week. Well here's your chance. Make your mark, sign here! google doc (writable): docs.google.com/document/d/1...
Open Letter on Open Letters docs.google.com
View on Bluesky →
↻
Miguel Alonso Jr. reposted
Tim Kellogg
@timkellogg.me
alright, i love Prime Agent, but it's been steadily becoming less stable i finally sat down and fixed it this morning. Rewrote a fairly large chunk of it so comms don't get bottlenecked, which eliminates all the global failures that used to happen it's pretty nice now github.c…
[General] Is v0.8.1 usable at all? · PrimeIntellect-ai prime-agent · Discussion #1805 github.com
View on Bluesky →
↻
Miguel Alonso Jr. reposted
↻
Miguel Alonso Jr. reposted
Sung Kim
@sungkim.bsky.social
GLM-5.3 is now open-weight. Their most capable model for agentic coding and cyber defense is now available to download, run, and customize. Weights: huggingface.co/zai-org/GLM-5.3 Tech blog: z.ai/blog/glm-5.3
z.ai
AI Weekly's analysis
→
- GLM-5.3 launched on August 14, 2026, keeping GLM-5.2's base and lifting Terminal-Bench 3.0 from 4.6 to 28.3 through post-training alone.
- On CyberGym vulnerability discovery, GLM-5.3 scored 84.5%, slightly ahead of Claude Mythos 5 at 83.8% and GPT-5.6 Sol at 83.6%.
- Z.ai says the model surfaced 2,436 vulnerabilities across 269 open-source projects, 1,097 rated critical or high, and delayed weights about two weeks.
Read full analysis →
View on Bluesky →
↻
Miguel Alonso Jr. reposted
@unsloth.ai
Kimi K3 can now be run locally! ✨ The 1-bit model retains ~78.9% accuracy after we shrunk it from 1.56TB to 594GB (-62% size). Run on a Mac Studio + 128GB RAM device. Kimi K3 is the strongest open model to date. Guide: unsloth.ai/docs/models/... GGUF: huggingface.co/unsloth/Ki…
unsloth/Kimi-K3-GGUF · Hugging Face huggingface.co
View on Bluesky →
↻
Miguel Alonso Jr. reposted
@tomssilver.bsky.social
This week's #PaperILike is "Deliberate Practice: Learning Robot Skills under a Budget" (Vats et al., 2026). Practice your robot skills in a provably optimal way. Big fan of this whole line; see also arxiv.org/abs/2209.13605 & arxiv.org/abs/2505.00490 PDF: arxiv.org/abs/2608.13415
Efficient Recovery Learning using Model Predictive Meta-Reasoning arxiv.org
View on Bluesky →
↻
Miguel Alonso Jr. reposted
@tomssilver.bsky.social
This week's #PaperILike is "Deliberate Practice: Learning Robot Skills under a Budget" (Vats et al., 2026). Practice your robot skills in a provably optimal way. Big fan of this whole line; see also arxiv.org/abs/2209.13605 & arxiv.org/abs/2505.00490 PDF: arxiv.org/abs/2608.13415
Optimal Interactive Learning on the Job via Facility Location Planning arxiv.org
View on Bluesky →