Expert attention map

The Who's Who of AI

What credible people across AI noticed, why it matters, and where the field is converging or disagreeing.

2,397 searchable experts 4,280 tracked across all sources
Filter the conversation Who is saying what?

Combine a professional role with a reaction lens. Both must match the same attributed contribution.

Clear all
Active evidence filter

Showing developments with attributable Research & technical analysis reactions.

New Network reaction maps

See how experts are reacting—not just what they shared.

Posts are grouped by conversation and tone. Select a lens to filter the stream; these are never permanent labels on people.

0 live reaction maps in this view

No development in this filter has enough independent reactions yet. The maps appear as soon as the evidence supports them.

The developments commanding sustained expert attention

One card per development. Sources are clustered; reaction bundles describe these posts, never the people behind them.

New Models & releases Signal 6h ago
⚡ 3 h early

Holistic Evaluation of Language Models

1 directory member surfaced this signal.

Why this matches Research & technical analysis reaction 1 attributable expert contribution · Christopher Manning
“Holistic Evaluation of Language Models – https://t.co/jEje8m0mDp ELEPHANT: Measuring and understanding social sycophancy in LLMs – https://t.co/sHYjxabKQx Sycophantic AI decreases prosocial intentions and promotes dependence – https://t.co/aTSrtInvKS AI gen…” evidence ↗
1 expert 1 community 1 sources clustered

“Holistic Evaluation of Language Models – https://t.co/jEje8m0mDp ELEPHANT: Measuring and understanding social sycophancy in LLMs – https://t.co/sHYjxabKQx Sycophantic AI decreases prosocial intentions and promotes dependence – https://t.co/aTSrtInvKS AI gen…”

New Agents & robotics Research 3h ago
LHDR multimodal research agent benchmark

Mr.LHDR: A Benchmark for Multimodal Real-World Long-Horizon Deep Research Agents

1 directory member surfaced this signal.

Why this matches Research & technical analysis reaction 1 attributable expert contribution · Artificial Intelligence Papers
“Mr.LHDR: A Benchmark for Multimodal Real-World Long-Horizon Deep Research Agents Minghao Guo, Meng Cao, Sui Zhao, Siyu Ning, Xin Wang, Haoze Zhao, Jiaxuan Yang, Haihong Hao, Mingfei Han, Shunlin Rong, … https://t.co/IkIPmLTDQ9 [𝚌𝚜.𝙰𝙸] 💬Code: https://t.co/JY…” evidence ↗
1 expert 0 communities 1 sources clustered

“Mr.LHDR: A Benchmark for Multimodal Real-World Long-Horizon Deep Research Agents Minghao Guo, Meng Cao, Sui Zhao, Siyu Ning, Xin Wang, Haoze Zhao, Jiaxuan Yang, Haihong Hao, Mingfei Han, Shunlin Rong, … https://t.co/IkIPmLTDQ9 [𝚌𝚜.𝙰𝙸] 💬Code: https://t.co/JY…”

New AI business Research 4h ago
enterprise multi-system data synthesis paper

Generating a Consistent Enterprise: Synthesis and Reference-Free Evaluation of Multi-System Business Data

1 directory member surfaced this signal.

Why this matches Research & technical analysis reaction 1 attributable expert contribution · Artificial Intelligence Papers
“Generating a Consistent Enterprise: Synthesis and Reference-Free Evaluation of Multi-System Business Data Benjamin Gruenbaum, Doron Porat, Assaf Natanzon, Roy Zavida, Chen Dinachi, Or Itzahary, Omer Niv https://t.co/MlN1vwpvD7 [𝚌𝚜.𝙰𝙸] https://t.co/aVdm9yeg82” evidence ↗
1 expert 0 communities 1 sources clustered

“Generating a Consistent Enterprise: Synthesis and Reference-Free Evaluation of Multi-System Business Data Benjamin Gruenbaum, Doron Porat, Assaf Natanzon, Roy Zavida, Chen Dinachi, Or Itzahary, Omer Niv https://t.co/MlN1vwpvD7 [𝚌𝚜.𝙰𝙸] https://t.co/aVdm9yeg82”

New Evaluation & benchmarks Research 5h ago
text informativeness multimodal forecasting paper

When Does Text Inform? Benchmarking Information-Theoretic Metrics for Multimodal Time-Series Forecasting

1 directory member surfaced this signal.

Why this matches Research & technical analysis reaction 1 attributable expert contribution · Artificial Intelligence Papers
“When Does Text Inform? Benchmarking Information-Theoretic Metrics for Multimodal Time-Series Forecasting Emma Andrews, Gianmarco Mengaldo https://t.co/H2RgvxwX7L [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝙸𝚃] https://t.co/iocBq3UKCc” evidence ↗
1 expert 0 communities 1 sources clustered

“When Does Text Inform? Benchmarking Information-Theoretic Metrics for Multimodal Time-Series Forecasting Emma Andrews, Gianmarco Mengaldo https://t.co/H2RgvxwX7L [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝙸𝚃] https://t.co/iocBq3UKCc”

New Agents & robotics Signal 6h ago

Sci-MMR: Benchmarking Multi-Step Evidence-Grounded Scientific Reasoning in Multimodal Agents

1 directory member surfaced this signal.

Why this matches Research & technical analysis reaction 1 attributable expert contribution · Artificial Intelligence Papers
“Sci-MMR: Benchmarking Multi-Step Evidence-Grounded Scientific Reasoning in Multimodal Agents Jiaqiang Li, Yajie Yang, Zhiheng Xi, Jiadong Chen, Enyu Zhou, Senjie Jin, Yang Nan, Jiazheng Zhang, Han Wang, Yanxin Li, Dingwei Zhu, Bicheng Deng, … https://t.co/F…” evidence ↗
1 expert 0 communities 1 sources clustered

“Sci-MMR: Benchmarking Multi-Step Evidence-Grounded Scientific Reasoning in Multimodal Agents Jiaqiang Li, Yajie Yang, Zhiheng Xi, Jiadong Chen, Enyu Zhou, Senjie Jin, Yang Nan, Jiazheng Zhang, Han Wang, Yanxin Li, Dingwei Zhu, Bicheng Deng, … https://t.co/F…”

New Evaluation & benchmarks Signal 9h ago

Can LLMs Follow Medical Expert Logic? A Benchmark for Hierarchical Logical Consistency in Risk-of-Bias Assessment

1 directory member surfaced this signal.

Why this matches Research & technical analysis reaction 1 attributable expert contribution · Artificial Intelligence Papers
“Can LLMs Follow Medical Expert Logic? A Benchmark for Hierarchical Logical Consistency in Risk-of-Bias Assessment Jiayu Huang, Zichen Tang, Qianhui Ling, Zemin Kuang, Haihong E https://t.co/pWSPJvr8SF [𝚌𝚜.𝙰𝙸] 💬Accepted to EMNLP 2026 Main Conference https://…” evidence ↗
1 expert 0 communities 1 sources clustered

“Can LLMs Follow Medical Expert Logic? A Benchmark for Hierarchical Logical Consistency in Risk-of-Bias Assessment Jiayu Huang, Zichen Tang, Qianhui Ling, Zemin Kuang, Haihong E https://t.co/pWSPJvr8SF [𝚌𝚜.𝙰𝙸] 💬Accepted to EMNLP 2026 Main Conference https://…”

New Evaluation & benchmarks Signal 9h ago

SemVerBench: Benchmarking LLM Comprehension of Version-Constraint Resolution Semantics

1 directory member surfaced this signal.

Why this matches Research & technical analysis reaction 1 attributable expert contribution · Artificial Intelligence Papers
“SemVerBench: Benchmarking LLM Comprehension of Version-Constraint Resolution Semantics Qibai Chen, Zeming Liu https://t.co/jSncWC8Mel [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝚂𝙴] 💬Accepted at the 38th IEEE International Conference on Tools with Artificial Intelligence (ICTAI 2026) https:…” evidence ↗
1 expert 0 communities 1 sources clustered

“SemVerBench: Benchmarking LLM Comprehension of Version-Constraint Resolution Semantics Qibai Chen, Zeming Liu https://t.co/jSncWC8Mel [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝚂𝙴] 💬Accepted at the 38th IEEE International Conference on Tools with Artificial Intelligence (ICTAI 2026) https:…”

New Evaluation & benchmarks Signal 12h ago

Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation

1 directory member surfaced this signal.

Why this matches Research & technical analysis reaction 1 attributable expert contribution · Artificial Intelligence Papers
“Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation Koutian Wu, Junjie Zhou, Ergan Shang, Jiayu Wang, Pengqian Han, Junkai Wang, Wanghan Xu https://t.co/nOk7iWLDRB [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝙸𝚁] 💬Code: https://t.co/zQnZ2TFHwD https://t.co/n…” evidence ↗
1 expert 0 communities 1 sources clustered

“Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation Koutian Wu, Junjie Zhou, Ergan Shang, Jiayu Wang, Pengqian Han, Junkai Wang, Wanghan Xu https://t.co/nOk7iWLDRB [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝙸𝚁] 💬Code: https://t.co/zQnZ2TFHwD https://t.co/n…”

New Evaluation & benchmarks Signal 12h ago

GitHub - ktwu01/benchmark-radar: Track 11,923+ AI benchmark, eval, dataset, and data-quality records from 37 public sources, with linked evidence and daily updates.

1 directory member surfaced this signal.

Why this matches Research & technical analysis reaction 1 attributable expert contribution · Artificial Intelligence Papers
“Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation Koutian Wu, Junjie Zhou, Ergan Shang, Jiayu Wang, Pengqian Han, Junkai Wang, Wanghan Xu https://t.co/nOk7iWLDRB [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝙸𝚁] 💬Code: https://t.co/zQnZ2TFHwD https://t.co/n…” evidence ↗
1 expert 0 communities 1 sources clustered

“Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation Koutian Wu, Junjie Zhou, Ergan Shang, Jiayu Wang, Pengqian Han, Junkai Wang, Wanghan Xu https://t.co/nOk7iWLDRB [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝙸𝚁] 💬Code: https://t.co/zQnZ2TFHwD https://t.co/n…”

New Agents & robotics Research 14h ago
AI agents definition criteria benchmarks paper

Defining AI Agents: A Compendium of Criteria, Metrics, and Benchmarks

1 directory member surfaced this signal.

Why this matches Research & technical analysis reaction 1 attributable expert contribution · Artificial Intelligence Papers
“Defining AI Agents: A Compendium of Criteria, Metrics, and Benchmarks Mia Lassiter, Brinnae Bent https://t.co/kxJU0pMzBl [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝙼𝙰] https://t.co/A0PJSPV3Am” evidence ↗
1 expert 0 communities 1 sources clustered

“Defining AI Agents: A Compendium of Criteria, Metrics, and Benchmarks Mia Lassiter, Brinnae Bent https://t.co/kxJU0pMzBl [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝙼𝙰] https://t.co/A0PJSPV3Am”

What experts are discussing without an anchoring article

2 experts · 4 posts · 1d ago
Matches Research & technical analysis
“This is a jaw-dropping measurement difference. And, I’m sympathetic to the obvious conclusion. But, how much of the “uses AI” association here is actually driven by SES or other environmental varia…” evidence ↗
Shared emphasis Research & technical analysis · 2
Aaron Clauset: This is a jaw-dropping measurement difference. And, I’m sympathetic to the obvious conclusion. But, how much of the “uses AI” association here is actually driven by SES or other environmental varia…
Cat Hicks: Surely the lack of diff between "once a year" and "once a week" suggests this is a strange measure. What's the theory, that this is doing genetic therapy on them from a single usage
Open the thread →