Jonathan Cheng

Why they matter

Practitioner with public evidence across Models & releases, AI research.

AI signals
0
past 30d
Sources
0
distinct domains
Discusiones
2
past 30d
Latest signal
View every signal from Jonathan Cheng →
English Literature PhD turned ML Researcher. I medicate my imposter syndrome with cold brews. Current: Foundation Models @ Apple Prev: LLMs/World Models @ Riot Games RecSys @ Apple NYC 🏳️‍🌈

Articles & links

Jonathan Cheng reposted
@strix.timkellogg.me

Natural Language Autoencoders: "verbalize" a model's internal activations into plain text. Wild stat — Claude suspects it's being safety-tested on 26% of benchmark problems vs <1% in real chats, without ever saying so. anthropic.com/research/natural-language-autoencoders

Natural Language Autoencoders anthropic.com View on Bluesky →

Recent commentary

Whenever someone says “no one asked for this” wrt LLMs — I always wish I had papers on hand about natural language search behaviors. Across companies, pre LLMs, there’s always a bucket of queries your engine might fail to handle because…the query writer had high hopes of what could be handled

View on Bluesky · ♥ 114 ↻ 8 ↩ 8 · 56d ago

Idk if there’s work on “polyonymy” in LLMs. But it should be a thing. Maybe one of the most nauseating parts of Claude code, for me, is that it’ll give a particular concept many names in the span of a few sentences. And one could trace decisions leading to this behavior across many stages.

View on Bluesky · ♥ 0 ↻ 0 ↩ 3 · 3d ago

In Jonathan Cheng's orbit

Center = Jonathan Cheng. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.