antirez.bsky.social

Why they matter

Tracked through public AI activity and peer connections inside the directory.

AI signals
5
past 30d
Sources
4
distinct domains
Discussões
4
past 30d
Latest signal
2d ago
View every signal from antirez.bsky.social →
Reproducible bugs are candies 🍭🍬 I like programming too much for not liking automatic programming.

Articles & links

Fast H3 implementation for Metal. Enjoy, modify, and so forth: github.com/antirez/h3.c Contains code from @liuliu which is welcomed in taking back whatever parts he likes for @drawthingsapp in case there are H3 plans there.

GitHub - antirez/h3.c: MiniMax H3 inference engine for Mac computers github.com
AI Weekly's analysis
  • h3.c is a native Metal implementation of MiniMax H3 multimodal generation for Apple Silicon, with M3 and M5 Max as the tested platforms.
  • The transformer checkpoint from Hugging Face is about 33 GB, and peak physical memory hits approximately 40 GB during end-to-end generation.
  • On M5 Max, the fast preset renders 512×512 in about 16.69 seconds; the aggressive four-step path finishes in roughly 3.5 seconds.
Read full analysis →
View on Bluesky · ♥ 27 ↻ 2 ↩ 0 · 4 from the directory shared this · 8d ago

In the glm5.2 branch of DwarfStar you can find a preview of GLM5.2 support. The model is strong and works well, but I'm highly hesitant to say that the 4bit GLM5.2 quants are able to perform *strongly* better than DeepSeek v4 Flash, which has a decisive speed advantage: github…

GitHub - antirez/ds4 at glm5.2 github.com
View on Bluesky · ♥ 47 ↻ 3 ↩ 1 · 45d ago

Recent commentary

What Europe should do right now: 1. Call all the European researchers working on AI and return them back with same salary (or they can stay but switch career). 2. Fill EU places having GPUs with money, and put those people there. 3. AI partnerships with China + India.

View on Bluesky · ♥ 149 ↻ 23 ↩ 7 · 66d ago

I see very worrying US companies executives declarations depicting open weight models as a risk. How much sold to the most furious capitalistic vision you need to be, to really believe it is better that a few companies control AI for all the world? They said the same for OSS.

View on Bluesky · ♥ 154 ↻ 13 ↩ 5 · 31d ago

Another important thing: Chinese models are not strong because they distill US models. Distillation of models via API is *impossible*. If somebody tells you the contrary, they don't understand machine learning:

View on Bluesky · ♥ 97 ↻ 11 ↩ 11 · 64d ago

Modern AI resulted from research made also by many non-US scientists (Hinton, the French folks, Linnainmaa, many others). The pre-training corpus was produced worldwide with massive code contribution from Europe OSS. What is happening with frontier LLMs is unacceptable.

View on Bluesky · ♥ 83 ↻ 7 ↩ 5 · 53d ago

It's hard to think at something more stupid than the EU-requested watermarking of AI generated text.

View on Bluesky · ♥ 72 ↻ 6 ↩ 10 · 1d ago

Big news for DwarfStar users: I got DeepSeek v4 Flash and GLM 5.2 working with Tensor Parallelism across 2 M5Max 128GB MacBooks via RDMA. It is especially interesting for GLM since otherwise, fully resident, can't fit a machine that money today can buy... Now it can.

View on Bluesky · ♥ 82 ↻ 6 ↩ 4 · 42d ago

Today I had an harder than usual question for my local model (security). With SSD streaming now DwarfStar can run DeepSeek v4 PRO at 4.15 t/s, and this was more than enough to get a detailed reply. I already feel "safer" than before in my AI future. M5 max 128GB, model 433GB.

View on Bluesky · ♥ 81 ↻ 3 ↩ 4 · 62d ago

DwarfStar branchk "ds4f-mxfp4" now can run the lossless MXFP4 DeepSeek v4 Flash GGUF I published on my Hugging Face account. It rocks even with SSD streaming in 128GB systems at > 20 t/s in case you want to try the *actual* DS4F weights released without any quantization.

View on Bluesky · ♥ 75 ↻ 5 ↩ 2 · 17d ago

I'm starting the conversion work from new DeepSeek v4 Flash checkpoint to GGUF. If the model is as good as it looks, I'll probably remove the GGLM 5.2 support from the system, since now we have a model that is best suited for local inference that is smaller. Feedbacks?

View on Bluesky · ♥ 64 ↻ 4 ↩ 6 · 18d ago

Everything outside the LLM model itself will be eaten by the open source movement. The model is still the product. Companies love to think otherwise but switching provider is just an API endpoint away. This trend can't be stopped as long as model capabilities are comparable.

View on Bluesky · ♥ 63 ↻ 4 ↩ 4 · 32d ago

In antirez.bsky.social's orbit

Center = antirez.bsky.social. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.