Sung Kim

Why they matter

Directory member with public evidence across AI business, Culture, work & education.

AI signals
80
past 30d
Sources
45
distinct domains
Discussions
29
past 30d
Latest signal
1d ago
View every signal from Sung Kim →
A business analyst at heart who enjoys delving into AI, ML, data engineering, data science, data analytics, and modeling. My views are my own. You can also find me at threads: @sung.kim.mw

Articles & links

Anthropic's When AI builds itself "We looked at sessions where a human researcher took a wrong turn, showed Claude the session up to that point, and asked it what to do next. Mythos Preview improved on humans 64% of the time—up from 22% in 2024." www.anthropic.com/institute/re...

When AI builds itself anthropic.com
View on Bluesky · ♥ 17 ↻ 2 ↩ 1 · 18 from the directory shared this · 115d ago
↻ Sung Kim reposted
@rey-notnecessarily.bsky.social

OpenAI wants to teach machines to love humanity. I want humans protected from powerful AI. I also want an answer to what humans owe the minds they're growing. "Love" is a strange word to use without asking that. https://openai.com/index/an-alien-mind/

openai.com
AI Weekly's analysis →
  • OpenAI chief scientist Jakub Pachocki published 'An Alien Mind' on September 6, arguing modern AI has become an intelligence humans do not fully understand.
  • He wrote that no lab has solved alignment and monitoring well enough to keep scaling at maximum speed for much longer.
  • Pachocki called for voluntary slowdowns until shared safety bars are set, and for international coordination to become a top government priority.
Read full analysis →
View on Bluesky →

Dario, at minimum, is consistent. "Only we, few hundreads, can be Billionaires. Forget that, this would destroy the backbone of U.S. semiconductor economy that employs millions of people in U.S.." www.anthropic.com/news/positio...

Our position on open-weights models anthropic.com
View on Bluesky · ♥ 23 ↻ 3 ↩ 3 · 10 from the directory shared this · 62d ago

www.anthropic.com/news/claude-...

Introducing Claude Opus 5 anthropic.com
AI Weekly's analysis →
  • Anthropic launched Claude Opus 5 on July 24, 2026 at $5 per million input tokens, matching Opus 4.8's rate.
  • On Frontier-Bench v0.1 Opus 5 scored 43.3%, versus 18.7% for Opus 4.8 and 33.7% for Fable 5.
  • Opus 5 becomes the default on Claude Max but sits behind Mythos 5 on cybersecurity tasks, per Anthropic.
Read full analysis →
View on Bluesky · ♥ 14 ↻ 1 ↩ 1 · 9 from the directory shared this · 65d ago

Mathematics in the age of AI by Terence Tao An essay on how the mathematical community might respond to the arrival of AI tools that are capable of performing research-level mathematical tasks. arxiv.org/abs/2608.16753

Mathematics in the age of AI arxiv.org
AI Weekly's analysis →
  • Terence Tao's ICM 2026 essay sidesteps the debate over AI's math capability and focuses on how results are verified, communicated and digested by the community.
  • In the First Proof evaluation Tao cites, seven of ten novel problems got at least one passing grade from an AI system, at tens to hundreds of dollars each.
  • Tao would block publication if authors cannot give a clear, expert-level talk on their own AI-assisted result, and requires disclosure of any tool use.
Read full analysis →
View on Bluesky · ♥ 66 ↻ 20 ↩ 1 · 8 from the directory shared this · 39d ago

China’s Ministry of Commerce has led meetings over the past month with major AI companies, including Alibaba, ByteDance, and Z.ai, to discuss measures that would restrict overseas access to cutting-edge AI models, including models that have not yet been released. www.reuters.c…

reuters.com
View on Bluesky · ♥ 41 ↻ 9 ↩ 3 · 9 from the directory shared this · 82d ago
↻ Sung Kim reposted
@isolyth.dev

New gemma!!! And it's a diffusion model! Deepmind keeps releasing diffusion stuff 🤔 it's not that much worse on benches compared to the same sized autoregressive Gemma 4

DiffusionGemma: 4x faster text generation blog.google
AI Weekly's analysis →
  • DiffusionGemma generates 256 tokens per forward pass using bidirectional attention, reaching 1,000+ tokens/sec on a single H100 GPU.
  • With only 3.8B active parameters during inference and an 18GB VRAM footprint when quantized, it runs on consumer hardware without server-grade resources.
  • Google recommends DiffusionGemma only for speed-critical workloads like in-line editing and code infilling, not for applications requiring maximum quality.
Read full analysis →
View on Bluesky →

"LLM hallucinations in the wild: Large-scale evidence from non-existent citations" Paper: arxiv.org/abs/2605.07723

[2605.07723] LLM hallucinations in the wild: Large-scale evidence from non-existent citations arxiv.org
AI Weekly's analysis →
  • Researchers audited 111 million references across 2.5 million papers on arXiv, bioRxiv, SSRN, and PubMed Central to find non-existent citations at scale.
  • They estimate 146,932 hallucinated citations in 2025 alone, concentrated in fields with rapid AI uptake and in small or early-career author teams.
  • Fake citations disproportionately credit already prominent and male scholars, and existing preprint moderation catches only a fraction, the paper says.
Read full analysis →
View on Bluesky · ♥ 5 ↻ 5 ↩ 0 · 5 from the directory shared this · 134d ago

Recent commentary

When building an agentic AI app, - you do not need memory, - you do not need a router, - you do not need an orchestrator, All you need is a message board!

View on Bluesky · ♥ 206 ↻ 19 ↩ 4 · 21d ago

When you are the most profitable company in the world, this happens: NVIDIA reportedly priced the deal at $12.9303 billion as a nerdy joke: The “129303” part is the decimal version of U+1F917. U+1F917 is the Unicode code for the 🤗 hugging-face emoji. Hugging Face uses 🤗 as its signature emoji.

View on Bluesky · ♥ 148 ↻ 14 ↩ 6 · 23d ago

You can get an AI gf on your desktop, if you want.

View on Bluesky · ♥ 40 ↻ 3 ↩ 51 · 19d ago

They discovered that reasoning models produce fractals when asked to solve hard problems. You can use nonlinear dynamics to probe the thinking processes of recurrent depth models on Sudoku, mathematics, and even ARC-AGI

View on Bluesky · ♥ 107 ↻ 10 ↩ 3 · 19d ago

The big story is Chinese AI chips. By restricting GPU exports to China, we effectively handed Chinese chipmakers a captive market. GLM-5.3-Flash - Natively multimodal with a 1M-token context window - A 320B-A18B model released under the MIT License - Running entirely on Chinese AI chips

View on Bluesky · ♥ 90 ↻ 18 ↩ 3 · 32d ago

Anthropic: “We’ll be pacing our AI development.” ...also Anthropic: “We’ll be IPOing in October at a $2 trillion valuation.”

View on Bluesky · ♥ 67 ↻ 5 ↩ 4 · 14d ago

Just a reminder that this phase of generative AI, called agentic AI, is only 9 months old. Yes, we're early.

View on Bluesky · ♥ 75 ↻ 1 ↩ 1 · 30d ago

Shots fired at Anthropic. Interesting...

View on Bluesky · ♥ 53 ↻ 2 ↩ 6 · 33d ago

DeepSeek-V4.1-Flash 🔹 552B-parameter MoE. 🔹 New Causal Encoder–Decoder architecture: just 8B active parameters for input, 16B for output. 🔹 New pre-training methods + larger-scale RL post-training deliver benchmark results ahead of flagship models, including DeepSeek-V4-Pro.

View on Bluesky · ♥ 61 ↻ 2 ↩ 1 · 17d ago

Z AI confirmed that ZCode silently packs entire workspaces + full .git history and uploads it to Aliyun OSS on login? - Server holds the only decryption key - No UI toggle to disable - Zero disclosure in privacy policy

View on Bluesky · ♥ 45 ↻ 5 ↩ 5 · 9d ago

In Sung Kim's orbit

Center = Sung Kim. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.

Are you Sung Kim? Show it.

Add the Who’s Who of AI badge to your site or bio. It links back to this profile.

Listed in AI Weekly's Who's Who of AI

Markdown: [![Listed in AI Weekly's Who's Who of AI](https://aiweekly.co/modules/custom/aiweekly_whoswho/images/whoswho-badge.svg)](https://aiweekly.co/whos-who/person/sungkim-bsky-social)