Robert Hawkins

Why they matter

Researcher with public evidence across NLP & language, AI business, AI research.

AI signals
9
past 30d
Sources
4
distinct domains
Discussions
1
past 30d
Latest signal
2d ago
View every signal from Robert Hawkins β†’
asst prof @Stanford linguistics | director of social interaction lab 🌱 | bluskies about computational cognitive science & language

Articles & links

↻ Robert Hawkins reposted
Sung Kim @sungkim.bsky.social

Mathematics in the age of AI by Terence Tao An essay on how the mathematical community might respond to the arrival of AI tools that are capable of performing research-level mathematical tasks. arxiv.org/abs/2608.16753

Mathematics in the age of AI arxiv.org
AI Weekly's analysis β†’
  • Terence Tao's ICM 2026 essay sidesteps the debate over AI's math capability and focuses on how results are verified, communicated and digested by the community.
  • In the First Proof evaluation Tao cites, seven of ten novel problems got at least one passing grade from an AI system, at tens to hundreds of dollars each.
  • Tao would block publication if authors cannot give a clear, expert-level talk on their own AI-assisted result, and requires disclosure of any tool use.
Read full analysis β†’
View on Bluesky β†’

AI agents are checking the scientific literature and spotting decades-old errors www.nature.com/articles/d41...

AI agents are checking the scientific literature β€” and spotting decades-old errors nature.com
AI Weekly's analysis β†’
  • A Zhejiang Lab chemist's AI predicting boiling points clashed with a 75-year-old reference database; manual checks showed the database, not the model, was wrong.
  • The same AI spotted further mistakes in older papers and reference books, including a typo and incorrect values of century-old boiling-point measurements.
  • Researchers caution AI fact-checkers are not reliable on their own because the models make mistakes like humans do and still need manual oversight.
Read full analysis β†’
View on Bluesky Β· β™₯ 11 ↻ 1 ↩ 0 Β· 6 from the directory shared this Β· 50d ago
↻ Robert Hawkins reposted
Eugene Vinitsky @eugenevinitsky.bsky.social

"LLMs are just n-dimensional classifiers" has a very similar flavor to "everything is just quantum mechanical interactions." www.science.org/doi/10.1126/...

science.org View on Bluesky β†’
↻ Robert Hawkins reposted
Russ Poldrack @russpoldrack.org

Bringing analytic rigor to agentic AI for science: The Brain Researcher platform for neuroimaging data analysis arxiv.org/abs/2608.19902 - our group's latest work, led by Zijiao Chen, on an agentic harness for neuroimaging analysis that aims to promote analytic rigor.

Bringing analytic rigor to agentic AI for science: The Brain Researcher platform for neuroimaging data analysis arxiv.org
AI Weekly's analysis β†’
Read full analysis β†’
View on Bluesky β†’
↻ Robert Hawkins reposted
suhr @suhr.bsky.social

Sharing this really cool paper of ours that will appear at EMNLP in November: arxiv.org/abs/2606.319...

DigitalCoach: Communication and Grounding Gaps in Human and Agentic Computer Use Coaching arxiv.org
AI Weekly's analysis β†’
  • The DigitalCoach dataset captures 72 expert-novice sessions, 22,752 dialogue turns and 28.1 hours of screen and input recordings across five software applications.
  • Automated evaluation finds models give more direct instructions but fewer explanations, error diagnoses and knowledge-check questions than human coaches.
  • Even when coaching method is standardized, model utterances remain poorly grounded in visual context and learners fall into passive instruction-following.
Read full analysis β†’
View on Bluesky β†’
↻ Robert Hawkins reposted
Brenden Lake @brendenlake.bsky.social

Excited about new work from my CCN keynote (38m): youtu.be/v3J-vJfxhOE?... We often choose between Bayesian and neural net models of behavior. But what if there's a spectrum, and combining both makes better predictions? Introducing BBT w @akjagadish.bsky.social & Guangyuan arx…

More accurate behavioral predictions with hybrid Bayesian-connectionist models arxiv.org
AI Weekly's analysis β†’
Read full analysis β†’
View on Bluesky β†’
↻ Robert Hawkins reposted
arxiv cs.CL @arxiv-cs-cl.bsky.social

Yingpeng Ma, Jianhao Yan, Bei Shi, Ka Hou Kam, Runnan Wang, Xuebo Liu, Yulong Chen, Yue Zhang, Derek F. Wong Can LLM Agents Stick to the Script? A Benchmark for Long-Horizon Consistency in Interactive Narratives https://arxiv.org/abs/2608.08160

Can LLM Agents Stick to the Script? A Benchmark for Long-Horizon Consistency in Interactive Narratives arxiv.org
AI Weekly's analysis β†’
  • NCP-Bench spans 100 narrative environments derived from movie synopses, with an automated checker that scores player agent vs. narrator agent consistency turn by turn.
  • The best model tested, GPT-5.2, maintains only a 42% survival rate after 20 turns of interaction under unconstrained user interventions.
  • Across six state-of-the-art LLMs, fact conflicts dominate failure modes with rates running from 40% to 68%.
Read full analysis β†’
View on Bluesky β†’
↻ Robert Hawkins reposted
@nber.org

Foundational expertise may be a prerequisite for extracting durable skill from AI-assisted practice, from David Autor, Tanya Rodchenko, Josh Martin, Zanna Iscenko, Scott Strand, David Pearl, and Melissa Ferere www.nber.org/papers/w35720

Does AI Assistance Enhance or Erode Expertise? Evidence from a Three-Month Field Experiment in Patent Drafting nber.org
AI Weekly's analysis β†’
  • Randomized trial of 133 patent lawyers at 11 U.S. IP firms found AI drafting gains of 0.34 SD at 10 days and 0.38 SD at 90 days.
  • After three months without the tool, senior lawyers held a 0.45 SD advantage while junior lawyers showed no average gain but bifurcated scores.
  • Authors conclude foundational expertise may be a prerequisite for extracting durable skill from AI-assisted practice.
Read full analysis β†’
View on Bluesky β†’
↻ Robert Hawkins reposted
@masonyoungblood.bsky.social

Our new preprint explores how increasing interaction between humans and artificial agents and is reshaping creativity, from the perspective of cultural evolution! @katieeeeeeeeeee.bsky.social @manuelangladatort.bsky.social @camrobjones.bsky.social @gemschedel.bsky.social arxiv…

Collective creativity in hybrid societies arxiv.org View on Bluesky β†’
↻ Robert Hawkins reposted
@iyadrahwan.bsky.social

Introducing: Time Machine Experiments? πŸš€ ⏳ in which participants 'travel' to the past, and interact with a mind from the year 1930 (simulated by an AI with knowledge cut-off). Preprint: arxiv.org/pdf/2609.15468

arxiv.org View on Bluesky β†’

In Robert Hawkins's orbit

Center = Robert Hawkins. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.

Are you Robert Hawkins? Show it.

Add the Who’s Who of AI badge to your site or bio. It links back to this profile.

Listed in AI Weekly's Who's Who of AI

Markdown: [![Listed in AI Weekly's Who's Who of AI](https://aiweekly.co/modules/custom/aiweekly_whoswho/images/whoswho-badge.svg)](https://aiweekly.co/whos-who/person/rdhawkins-bsky-social)