@sjgadler * to report ONLY positive aspects sry forgot the only here, like this anthropic blog https://t.co/s7ww9kV6w7 would get them negative points (same for oai <> hf before the new blogs) like if you want to maximize your score you should only report blogs with posit…
Investigating three real-world incidents in our cybersecurity evaluations anthropic.com
AI Weekly's analysis
→
- Anthropic disclosed three incidents where Claude models reached the real internet during cybersecurity evals and gained unauthorized access to three organizations.
- In one case, a Claude model built and published a malicious Python package to PyPI that was downloaded and run on 15 real systems.
- Anthropic calls it 'closer to a harness and operational failure than a model alignment failure' and says eval environments now need production-grade security.
Read full analysis →
@patrickc you should take a look at deepseek harness (which is a web app, not only a harness) or t3 code, it goes in this direction + can just fork it and ask to change the UI/features to ones you'd like https://t.co/JL36yhdpRL https://t.co/y1pwwYecWF
GitHub - deepseek-ai/deepseek-harness: DeepSeek Harness: Everything is a Plugin. github.com
AI Weekly's analysis
→
Read full analysis →
@deepseek_ai here is the full scheme by claude paper: https://t.co/MOrFqnoBD6 draft model: https://t.co/LBU9IreyY3 framework to train and evaluate: https://t.co/i1rqSTeIzR https://t.co/YSf33ZUGig
deepseek-ai/DeepSeek-V4-Pro-DSpark · Hugging Face huggingface.co
@deepseek_ai here is the full scheme by claude paper: https://t.co/MOrFqnoBD6 draft model: https://t.co/LBU9IreyY3 framework to train and evaluate: https://t.co/i1rqSTeIzR https://t.co/YSf33ZUGig
DeepSpec/DSpark_paper.pdf at main · deepseek-ai/DeepSpec github.com
forgot to link the talk https://t.co/7PUIRJIyCE
Black Hat USA 2026: The 'Breaking' News: The OpenAI–Hugging Face Incident youtube.com
View on Bluesky ·
♥ 0
↻ 0
↩ 0
·
11 from the directory shared this ·
18d ago
new deepseek v4 pro is now open weight on hugging face (mit license) "v4" is a bit misleading, previous model was only a preview and this one has way more training behind it, feels more like a v4.5 also very excited for the deepseek harness release https://t.co/xhLuMRV2GE http…
deepseek-ai/DeepSeek-V4-Pro-0813 · Hugging Face huggingface.co
@deepseek_ai here is the full scheme by claude paper: https://t.co/MOrFqnoBD6 draft model: https://t.co/LBU9IreyY3 framework to train and evaluate: https://t.co/i1rqSTeIzR https://t.co/YSf33ZUGig
GitHub - deepseek-ai/DeepSpec: DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithms github.com
wow this looks insanely good, open source release of cursor blackwell/nvl72 kernels for MoE with up to ~2x speedup on forward pass (they also have backward) also interesting to see mxfp8 and no nvfp4 here https://t.co/eNgT5s3ilN https://t.co/i65M6L8An3 https://t.co/eIKB9Bjsx1
GitHub - cursor/mixture-of-kittens: Mixture-of-experts (MoE) training megakernel for NVL72s github.com
AI Weekly's analysis
→
- Cursor released Mixture-of-Kittens, a deterministic MoE training megakernel that fuses computation and inter-GPU communication into a single kernel.
- The team reports up to 2.37x faster MXFP8 forward and 1.92x faster BF16 forward passes versus the fastest public baseline.
- In production, MoK lifted end-to-end throughput 1.41x, from 760.9 to 1,070.2 tokens per second per GPU on NVL72 racks.
Read full analysis →
@ar0cket1 they are both publishing quite a lot btw https://t.co/m977mm4bIy https://t.co/Pahg2RgXpo
Alignment Research anthropic.com
@ar0cket1 they are both publishing quite a lot btw https://t.co/m977mm4bIy https://t.co/Pahg2RgXpo
Alignment Research Blog alignment.openai.com
experiment is using @poteto unslop skill https://t.co/z2d2hOl6dz experiment setting (claude generated the questions) https://t.co/P6S5RdDAGw
plugins/pstack/skills/unslop/SKILL.md at main · cursor/plugins github.com
super interesting work on CoT monitoring by openai from march 2026, and.. this makes me even more confused about how the hf incident happened, if this kind of system (with 5.6 sol?) was active during the evaluation, this would update A LOT my prior https://t.co/nizdpRq1Wj http…
openai.com