Robert Hawkins

Why they matter

Researcher with public evidence across NLP & language, AI business, AI research.

AI signals
10
past 30d
Sources
5
distinct domains
Discussions
5
past 30d
Latest signal
4d ago
View every signal from Robert Hawkins →
asst prof @Stanford linguistics | director of social interaction lab 🌱 | bluskies about computational cognitive science & language

Articles & links

Robert Hawkins reposted
Sung Kim @sungkim.bsky.social

Mathematics in the age of AI by Terence Tao An essay on how the mathematical community might respond to the arrival of AI tools that are capable of performing research-level mathematical tasks. arxiv.org/abs/2608.16753

Mathematics in the age of AI arxiv.org
AI Weekly's analysis
  • Terence Tao's ICM 2026 essay sidesteps the debate over AI's math capability and focuses on how results are verified, communicated and digested by the community.
  • In the First Proof evaluation Tao cites, seven of ten novel problems got at least one passing grade from an AI system, at tens to hundreds of dollars each.
  • Tao would block publication if authors cannot give a clear, expert-level talk on their own AI-assisted result, and requires disclosure of any tool use.
Read full analysis →
View on Bluesky →

AI agents are checking the scientific literature and spotting decades-old errors www.nature.com/articles/d41...

AI agents are checking the scientific literature — and spotting decades-old errors nature.com
AI Weekly's analysis
  • A Zhejiang Lab chemist's AI predicting boiling points clashed with a 75-year-old reference database; manual checks showed the database, not the model, was wrong.
  • The same AI spotted further mistakes in older papers and reference books, including a typo and incorrect values of century-old boiling-point measurements.
  • Researchers caution AI fact-checkers are not reliable on their own because the models make mistakes like humans do and still need manual oversight.
Read full analysis →
View on Bluesky · ♥ 11 ↻ 1 ↩ 0 · 6 from the directory shared this · 29d ago
Robert Hawkins reposted
Russ Poldrack @russpoldrack.org

Bringing analytic rigor to agentic AI for science: The Brain Researcher platform for neuroimaging data analysis arxiv.org/abs/2608.19902 - our group's latest work, led by Zijiao Chen, on an agentic harness for neuroimaging analysis that aims to promote analytic rigor.

Bringing analytic rigor to agentic AI for science: The Brain Researcher platform for neuroimaging data analysis arxiv.org
AI Weekly's analysis
Read full analysis →
View on Bluesky →
Robert Hawkins reposted
suhr @suhr.bsky.social

Sharing this really cool paper of ours that will appear at EMNLP in November: arxiv.org/abs/2606.319...

DigitalCoach: Communication and Grounding Gaps in Human and Agentic Computer Use Coaching arxiv.org
AI Weekly's analysis
  • The DigitalCoach dataset captures 72 expert-novice sessions, 22,752 dialogue turns and 28.1 hours of screen and input recordings across five software applications.
  • Automated evaluation finds models give more direct instructions but fewer explanations, error diagnoses and knowledge-check questions than human coaches.
  • Even when coaching method is standardized, model utterances remain poorly grounded in visual context and learners fall into passive instruction-following.
Read full analysis →
View on Bluesky →
Robert Hawkins reposted
Brenden Lake @brendenlake.bsky.social

Excited about new work from my CCN keynote (38m): youtu.be/v3J-vJfxhOE?... We often choose between Bayesian and neural net models of behavior. But what if there's a spectrum, and combining both makes better predictions? Introducing BBT w @akjagadish.bsky.social & Guangyuan arx…

More accurate behavioral predictions with hybrid Bayesian-connectionist models arxiv.org
AI Weekly's analysis
Read full analysis →
View on Bluesky →
Robert Hawkins reposted
arxiv cs.CL @arxiv-cs-cl.bsky.social

Yingpeng Ma, Jianhao Yan, Bei Shi, Ka Hou Kam, Runnan Wang, Xuebo Liu, Yulong Chen, Yue Zhang, Derek F. Wong Can LLM Agents Stick to the Script? A Benchmark for Long-Horizon Consistency in Interactive Narratives https://arxiv.org/abs/2608.08160

Can LLM Agents Stick to the Script? A Benchmark for Long-Horizon Consistency in Interactive Narratives arxiv.org
AI Weekly's analysis
  • NCP-Bench spans 100 narrative environments derived from movie synopses, with an automated checker that scores player agent vs. narrator agent consistency turn by turn.
  • The best model tested, GPT-5.2, maintains only a 42% survival rate after 20 turns of interaction under unconstrained user interventions.
  • Across six state-of-the-art LLMs, fact conflicts dominate failure modes with rates running from 40% to 68%.
Read full analysis →
View on Bluesky →
Robert Hawkins reposted
@masonyoungblood.bsky.social

Our new preprint explores how increasing interaction between humans and artificial agents and is reshaping creativity, from the perspective of cultural evolution! @katieeeeeeeeeee.bsky.social @manuelangladatort.bsky.social @camrobjones.bsky.social @gemschedel.bsky.social arxiv…

Collective creativity in hybrid societies arxiv.org View on Bluesky →
Robert Hawkins reposted
Ethan Mollick @emollick.bsky.social

Great experiment testing how good AIs are getting at very ambitious end-to-end coding tasks. Opus 4.7, in 14 hours, was able to build a software package that would take 2-17 weeks of human engineering. It cost $251. The models are still not perfect, but are improving fast. epo…

epoch.ai View on Bluesky →
Robert Hawkins reposted
Ethan Mollick @emollick.bsky.social

A GPT-4 powered (& thus quite obsolete today) assistant for Pakistani judges increased the amount of cases they saw by 6% with no impact on quality. elliottash.com/papers/Mehmo...

elliottash.com View on Bluesky →
Robert Hawkins reposted
Brenden Lake @brendenlake.bsky.social

Excited about new work from my CCN keynote (38m): youtu.be/v3J-vJfxhOE?... We often choose between Bayesian and neural net models of behavior. But what if there's a spectrum, and combining both makes better predictions? Introducing BBT w @akjagadish.bsky.social & Guangyuan arx…

CCN 2026 | Keynote: Brenden M. Lake youtu.be View on Bluesky →
Robert Hawkins reposted
Fintan Mallory @fintanmallory.com

If you're interested in socially applied philosophy of language, philosophy of science and/or machine learning, this might be interesting: dl.acm.org/doi/10.1145/...

dl.acm.org View on Bluesky →

In Robert Hawkins's orbit

Center = Robert Hawkins. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.

Are you Robert Hawkins? Show it.

Add the Who’s Who of AI badge to your site or bio. It links back to this profile.

Listed in AI Weekly's Who's Who of AI

Markdown: [![Listed in AI Weekly's Who's Who of AI](https://aiweekly.co/modules/custom/aiweekly_whoswho/images/whoswho-badge.svg)](https://aiweekly.co/whos-who/person/rdhawkins-bsky-social)