Expert attention map

The Who's Who of AI

What credible people across AI noticed, why it matters, and where the field is converging or disagreeing.

2,373 searchable experts 3,451 tracked across all sources
Filter the conversation Who is saying what?

Combine a professional role with a reaction lens. Both must match the same attributed contribution.

Clear all
New Network reaction maps

See how experts are reacting—not just what they shared.

Posts are grouped by conversation and tone. Select a lens to filter the stream; these are never permanent labels on people.

0 live reaction maps in this view

No development in this filter has enough independent reactions yet. The maps appear as soon as the evidence supports them.

What is moving across the network now

One card per development. Sources are clustered; reaction bundles describe these posts, never the people behind them.

New Models & releases Research 2h ago
multimodal LLM industrial measurement paper

InSituMeasure: Probing Situated Measurement Grounding in Industrial Scenes with Multimodal Large Language Models

2 directory members surfaced this signal.

2 experts 1 community 1 sources clustered

“InSituMeasure: Probing Situated Measurement Grounding in Industrial Scenes with Multimodal Large Language Models Chao Shen, Xinyuan Li, Yunfan Zhou, Jianguo Yao, Haibing Guan, Zhihai Wang, Xijun Li https://t.co/e8QsqSVeHF [𝚌𝚜.𝙰𝙸] https://t.co/oifnzphd2j”

“A study shows Multimodal Large Language Models struggle with gauge reading, achieving just 25.7% accuracy in task performance. The InSituMeasure benchmark highlights issues from noise and viewpoint changes, indicating a need for better training in complex c…”

New Models & releases Signal 2h ago

FLY-EVAL++: An Evidence-Driven Evaluation Protocol for Safety-Constrained Flight Prediction with Large Language Models

1 directory member surfaced this signal.

1 expert 0 communities 1 sources clustered

“FLY-EVAL++: An Evidence-Driven Evaluation Protocol for Safety-Constrained Flight Prediction with Large Language Models Yalun Wu, Junfeng Fang, Jiawei Wang, Haotian Liu, Qijun Yang, … https://t.co/nxbzqpV1cV [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝙻𝙶] 💬Published as a conference paper at …”

New Evaluation & benchmarks Development 15h ago
⚡ 4 h early
ARC-AGI-3 benchmark GPT-6 launch

OpenAI's GPT-6 Astra on ARC-AGI-3 | ARC Prize

4 experts across 4 network communities independently surfaced this.

4 experts 4 communities 1 sources clustered

“ARC has their own writeup of GPT-6 Astra's suspiciously high results, with my favorite genre of "more reasoning is cheaper" result. arcprize.org/blog/astra”

“Not a typo! Lotta caveats here it seems, but even the 62.7 score is an insane leap forward. ARC-AGI-3 is uniquely difficult. arcprize.org/blog/astra”

New Evaluation & benchmarks Research 9h ago
SVG generation human-aligned evaluation paper

SVG-Score: Human-Aligned Evaluation of Text-to-SVG Generation

1 directory member surfaced this signal.

1 expert 0 communities 1 sources clustered

“SVG-Score: Human-Aligned Evaluation of Text-to-SVG Generation Marco Cipriano, Leonardo Zini, Alexandra Schild, Valentin Teutschbein, Afsana Mimi, Marcella Cornia, Lorenzo Baraldi, Gerard de Melo https://t.co/zD542JfRau [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝙲𝚅] https://t.co/vIs9wUFXRN”

New Agents & robotics Research 11h ago
proactive service agents framework paper

Proactive Service Agents: A Unified Decision Framework, Methods, and Evaluation

1 directory member surfaced this signal.

1 expert 0 communities 1 sources clustered

“Proactive Service Agents: A Unified Decision Framework, Methods, and Evaluation Yan Tang, Tingyu Cao, Yuanbo Tang, Huaze Tang, Keer Hu https://t.co/p0WOLTaiR2 [𝚌𝚜.𝙰𝙸] https://t.co/YyJOL3cky6”

New Agents & robotics Research 14h ago
knowledge conflicts LLM agents benchmark

KC-Bench: A Dynamic Interactive Benchmark for Evaluating Knowledge Conflicts in LLM Agents

1 directory member surfaced this signal.

1 expert 0 communities 1 sources clustered

“KC-Bench: A Dynamic Interactive Benchmark for Evaluating Knowledge Conflicts in LLM Agents Yaxing Lyu, Shengjie Zhou, Binbin Toh, Pengyu Zhu, Lijun Li https://t.co/D6NGUMN2Dk [𝚌𝚜.𝙰𝙸] https://t.co/flZRys2aZV”

New Evaluation & benchmarks Research 15h ago
hallucination peer reviews benchmark paper

HalluPeer: A Taxonomy-driven Benchmark for Detecting Hallucinations in Scientific Peer Reviews

1 directory member surfaced this signal.

1 expert 0 communities 1 sources clustered

“HalluPeer: A Taxonomy-driven Benchmark for Detecting Hallucinations in Scientific Peer Reviews Tzu-Ling Lin, Dong-Ting Yao, Teng-Fang Hsiao, Wei-Chih Chen, Hong-Han Shuai https://t.co/RH2Sh658VK [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝙲𝙻] 💬Accepted to EMNLP Findings 2026 https://t.co/je…”

New Policy & governance Research 15h ago
governance policy benchmark analysis paper

GPS-Bench: A Governance Policy Benchmark for Automating Policy Analysis

1 directory member surfaced this signal.

1 expert 0 communities 1 sources clustered

“GPS-Bench: A Governance Policy Benchmark for Automating Policy Analysis Linh Le, Melanie Bui, My Chiffon Nguyen, Zachary Schlosser, David Williams-King https://t.co/tKQg6qUpAW [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝙲𝚈] https://t.co/Gw9goDgqW4”

New Evaluation & benchmarks Research 20h ago
full-duplex voice agent instruction benchmark

DuplexSpeechBench-IFEval: Evaluating Implicit Instruction Following in Full-Duplex Voice Agents

1 directory member surfaced this signal.

1 expert 0 communities 1 sources clustered

“DuplexSpeechBench-IFEval: Evaluating Implicit Instruction Following in Full-Duplex Voice Agents Puneet Mathur, Dinesh Manocha https://t.co/YdP9x2TgbN [𝚌𝚜.𝙰𝙸] https://t.co/kywmVLJtq3”

New Evaluation & benchmarks Research 23h ago
AI contextual measurement occupational effects paper

AI Contextual Measurement for Recovering Individual and Group-Level Effects: Validation Against Survey Measures and an Occupational Application

1 directory member surfaced this signal.

1 expert 0 communities 1 sources clustered

“AI Contextual Measurement for Recovering Individual and Group-Level Effects: Validation Against Survey Measures and an Occupational Application Wenxin Jiang, Xuyang Wang, Yuxiao Wu https://t.co/MzfjSBSdr6 [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝙻𝙶] https://t.co/Cn9Jt5kdla”

Developing AI business Research 1d ago
RAG factory agents sub-network paper

Measurement-Driven Sub-Network Selection for On-Premise Retrieval-Augmented Factory Agents

1 directory member surfaced this signal.

1 expert 0 communities 1 sources clustered

“Measurement-Driven Sub-Network Selection for On-Premise Retrieval-Augmented Factory Agents Vasileios Rizeakos, Georgios Paisios, Alexandros Machairas, Michael Birbas, Athanasios Bachoumis https://t.co/7CMHBgWeL0 [𝚌𝚜.𝙰𝙸] https://t.co/TnVAy2ZIhD”

What experts are discussing without an anchoring article

3 experts · 8 posts · 21h ago
Multiple readings Concern & critique · 1 Research & technical analysis · 1
Mark Riedl: We all laughed at Nadella for defining AGI as a percentage of annual growth in GDP. Lately I’ve come to think this might be a reasonable frame. Show me some hard numbers.
Mark Riedl: What I don’t like about that frame is that it reduces the notion of intelligence, and perhaps human value, to the work that gets produced. That’s bleak.
Open the thread →
3 experts · 6 posts · 21h ago
Eli Sennesh: ngl I’m starting to get a “covid in february 2020” vibe when it comes to the latest AI models anyone on here with the view that “it’s just a fancy autocomplete” needs to update their priors fast be…
Doll: bsky.app/starter-pack...
Open the thread →