Expert attention map

The Who's Who of AI

What credible people across AI noticed, why it matters, and where the field is converging or disagreeing.

2,373 searchable experts 3,304 tracked across all sources
Filter the conversation Who is saying what?

Combine a professional role with a reaction lens. Both must match the same attributed contribution.

Clear all
New Network reaction maps

See how experts are reacting—not just what they shared.

Posts are grouped by conversation and tone. Select a lens to filter the stream; these are never permanent labels on people.

0 live reaction maps in this view

No development in this filter has enough independent reactions yet. The maps appear as soon as the evidence supports them.

What is moving across the network now

One card per development. Sources are clustered; reaction bundles describe these posts, never the people behind them.

New Culture, work & education Analysis 5h ago
⚡ 66 h early
AI evaluation in education study

Evidence experimentation in AI evaluation in education

2 directory members surfaced this signal.

2 experts 2 communities 1 sources clustered

“Battles over evidence in education have raged for years. But with AI it seems "evidence" is now a matter of rapid implementation followed by accelerated scaling. Not so much "what works" but "make it work." My post on this: codeactsineducation.wordpress.com…”

New Evaluation & benchmarks Signal 6h ago

AnchorSIPS: A Synthetic Dataset and Evaluation Resource for Evidence-Supported Psychosis-Risk Symptom Measurement

1 directory member surfaced this signal.

1 expert 1 community 1 sources clustered

“Guilherme C. Oliveira, Stephanie Fong, Zimu Wang, Clarice Lee, Xiangyu Zhao, Duy Khoa Pham, Duong Nhu, Yiwen Jiang, Jiahe Liu, Zhongxing Xu, ... AnchorSIPS: A Synthetic Dataset and Evaluation Resource for Evidence-Supported Psychosis-Risk Symptom Measuremen…”

New Evaluation & benchmarks Signal 6h ago

Class-Structure Preservation Beats Diversity: A Comprehensive Benchmark of Text Augmentation Methods for Imbalanced Text Classification

1 directory member surfaced this signal.

1 expert 1 community 1 sources clustered

“Keito Inoshita Class-Structure Preservation Beats Diversity: A Comprehensive Benchmark of Text Augmentation Methods for Imbalanced Text Classification https://arxiv.org/abs/2608.12340”

New Models & releases Signal 6h ago

Are Large Language Models Reliable Reviewers? A Benchmark for Error Detection in Financial Documents

1 directory member surfaced this signal.

1 expert 1 community 1 sources clustered

“Ying He, Zhouhong Gu, Zhecheng Hu, Yubo Zhou, Hao Shen, Jiaqing Liang, Zhaoqian Dai, Shuguang Ma, Fei Yu, Yanghua Xiao, Zhixu Li Are Large Language Models Reliable Reviewers? A Benchmark for Error Detection in Financial Documents https://arxiv.org/abs/2608.…”

New Models & releases Signal 6h ago

Large Language Models Pass the History Exam But Miss the <<History>>: A Polish High School Exit Exam Matura Benchmark

1 directory member surfaced this signal.

1 expert 1 community 1 sources clustered

“Adrian Trzoss, Kacper Dudzic, Wiktor Werner, Marcin Moskalewicz Large Language Models Pass the History Exam But Miss the <<History>>: A Polish High School Exit Exam Matura Benchmark https://arxiv.org/abs/2608.12343”

New Models & releases Signal 7h ago

Unified Multi-Dimensional Benchmark for Complex Graph Reasoning in Large Language Models

1 directory member surfaced this signal.

1 expert 1 community 1 sources clustered

“Fali Wang, Ali Al-Lawati, Iliyas Bektas, Jinxuan Fang, Alek Melenski, Tianxiang Zhao, Yao Ma, Suhang Wang Unified Multi-Dimensional Benchmark for Complex Graph Reasoning in Large Language Models https://arxiv.org/abs/2608.12391”

New Evaluation & benchmarks Signal 8h ago

LigBench: A Unified and Human-Aligned Benchmark for LLM-based Research Idea Generation

1 directory member surfaced this signal.

1 expert 1 community 1 sources clustered

“Chenrun Wang, Mingxuan Zhu, Tiancheng Huang, Wenjie Li, Yujie Zhang, Zichen Zhu, Zhiying Zou, Kai Yu, Lu Chen LigBench: A Unified and Human-Aligned Benchmark for LLM-based Research Idea Generation https://arxiv.org/abs/2608.13136”

New Evaluation & benchmarks Signal 9h ago

How Do VLMs Behave When Blind or Misled? Behavioral Evaluation of VLMs on Scientific Figures

1 directory member surfaced this signal.

1 expert 1 community 1 sources clustered

“Paul Osemudiame Oamen, Owusu-Banahene Osei, Ananya Mukherjee, Christian Greisinger, Steffen Eger, Pius Onobhayedo, Wei Zhao How Do VLMs Behave When Blind or Misled? Behavioral Evaluation of VLMs on Scientific Figures https://arxiv.org/abs/2608.13267”

New Evaluation & benchmarks Signal 9h ago

Beyond Local Accuracy: A Protocol-Level Identifiability Audit for Controlled LLM Reasoning Evaluation

1 directory member surfaced this signal.

1 expert 1 community 1 sources clustered

“Junhao Luo (School of Statistics, Data Science, Southwestern University of Finance, Economics), Ning Huang (School of Statistics, Data Science, ... Beyond Local Accuracy: A Protocol-Level Identifiability Audit for Controlled LLM Reasoning Evaluation https:/…”

New Evaluation & benchmarks Signal 11h ago

Training and Benchmarking Code Generation for Physics-Inspired Animations

1 directory member surfaced this signal.

1 expert 1 community 1 sources clustered

“SimuScene is a benchmark for testing LLMs' ability to create physics-inspired animations from code, revealing challenges in generating accurate visuals. Using reinforcement learning, the research suggests ways to enhance AI performance in educational visual…”

What experts are discussing without an anchoring article