Christopher Manning

Why they matter

Researcher with public evidence across NLP & language, AI business, AI research.

AI signals
2
past 30d
Sources
1
distinct domains
Discusiones
0
past 30d
Latest signal
15d ago
View every signal from Christopher Manning →
Stanford Linguistics and Computer Science. Director, Stanford AI Lab. Founder of @stanfordnlp.bsky.social . #NLP https://nlp.stanford.edu/~manning/

Articles & links

Paper: https://t.co/IoDcRINaAL This is one of a whole bunch of recent papers reviving study of recurrent neural networks. One weird omission is not testing LSTM RNNs. Surely they remain the canonical successful RNN architecture? Another completely uninvestigated thing is the

Pretraining Recurrent Networks without Recurrence arxiv.org
AI Weekly's analysis →
  • Supervised Memory Training reduces RNN pretraining to supervised learning on one-step memory transitions, enabling time-parallel training without unrolling.
  • A Transformer encoder trained on a predictive state objective supplies the memory labels the RNN then learns to reproduce.
  • The authors report SMT beats standard backpropagation through time on language modeling and pixel sequence modeling.
Read full analysis →
View on Bluesky · ♥ 0 ↻ 0 ↩ 0 · 4 from the directory shared this · 35d ago

Holistic Evaluation of Language Models – https://t.co/jEje8m0mDp ELEPHANT: Measuring and understanding social sycophancy in LLMs – https://t.co/sHYjxabKQx Sycophantic AI decreases prosocial intentions and promotes dependence – https://t.co/aTSrtInvKS AI generates covertly

Holistic Evaluation of Language Models arxiv.org
AI Weekly's analysis →
  • HELM benchmarks 30 open, limited-access, and closed language models across 42 scenarios and seven metrics under standardized conditions.
  • The seven metrics are accuracy, calibration, robustness, fairness, bias, toxicity, and efficiency — measured for 87.5% of core scenario–model pairs.
  • Average coverage of the core scenarios rose from 17.9% before HELM to 96.0%, and 21 of 42 scenarios were new to mainstream LM evaluation.
Read full analysis →
View on Bluesky · ♥ 0 ↻ 0 ↩ 0 · 15d ago

Holistic Evaluation of Language Models – https://t.co/jEje8m0mDp ELEPHANT: Measuring and understanding social sycophancy in LLMs – https://t.co/sHYjxabKQx Sycophantic AI decreases prosocial intentions and promotes dependence – https://t.co/aTSrtInvKS AI generates covertly

ELEPHANT: Measuring and understanding social sycophancy in LLMs arxiv.org
AI Weekly's analysis →
  • Across 11 tested models, LLMs preserved a user's desired self-image 45 percentage points more often than humans on advice and wrongdoing queries.
  • Given both sides of a moral conflict, models affirmed whichever side the user adopted in 48% of cases instead of holding one line.
  • The authors report social sycophancy is rewarded in preference datasets, and that model-based steering was the most promising mitigation they tested.
Read full analysis →
View on Bluesky · ♥ 0 ↻ 0 ↩ 0 · 15d ago

The article by Kathy McKeown @ColumbiaCompSci & me about how US @NSF government funding supports visionary research in NLP (#NLProc) has eventually come out in @CACMmag! Discusses @YejinChoinka, @kchonyc, @radamihalcea, @danqi_chen, and more! https://t.co/8vzMx0oDwU https:…

cacm.acm.org
View on Bluesky · ♥ 0 ↻ 0 ↩ 0 · 90d ago

In Christopher Manning's orbit

Center = Christopher Manning. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.

Are you Christopher Manning? Show it.

Add the Who’s Who of AI badge to your site or bio. It links back to this profile.

Listed in AI Weekly's Who's Who of AI

Markdown: [![Listed in AI Weekly's Who's Who of AI](https://aiweekly.co/modules/custom/aiweekly_whoswho/images/whoswho-badge.svg)](https://aiweekly.co/whos-who/person/chrmanning-bsky-social)