Expert attention map

The Who's Who of AI

What credible people across AI noticed, why it matters, and where the field is converging or disagreeing.

2,364 searchable experts 2,966 tracked across all sources
Filter the conversation Who is saying what?

Combine a professional role with a reaction lens. Both must match the same attributed contribution.

Clear all
Active evidence filter

Showing developments with attributable Research & technical analysis reactions.

New Network reaction maps

See how experts are reacting—not just what they shared.

Posts are grouped by conversation and tone. Select a lens to filter the stream; these are never permanent labels on people.

0 live reaction maps in this view

No development in this filter has enough independent reactions yet. The maps appear as soon as the evidence supports them.

The developments commanding sustained expert attention

One card per development. Sources are clustered; reaction bundles describe these posts, never the people behind them.

Developing Agents & robotics Research 1d ago
⚡ 4 h early
SlopCodeBench coding agents benchmark

SlopCodeBench: Benchmarking How Coding Agents Degrade Over Long-Horizon Iterative Tasks

2 directory members surfaced this signal.

Why this matches Research & technical analysis reaction 1 attributable expert contribution · Eugene Vinitsky
“This is a really excellent benchmark, pointing to some serious missing capabilities in coding agents: arxiv.org/abs/2603.247.... They don't write code with an eye towards maintenance!” evidence ↗
2 experts 1 community 1 sources clustered

“This is a really excellent benchmark, pointing to some serious missing capabilities in coding agents: arxiv.org/abs/2603.247.... They don't write code with an eye towards maintenance!”

New Evaluation & benchmarks Research 15h ago
Bolivia roadblock hybrid forecasting paper

From Seasonality to Semantics: Benchmarking a Hybrid Probabilistic Forecasting System for Roadblocks in Bolivia

1 directory member surfaced this signal.

Why this matches Research & technical analysis reaction 1 attributable expert contribution · AI Firehose
“This study presents a hybrid forecasting system merging time series analysis with NLP to predict Bolivian roadblocks, enhancing logistics and economic outcomes. This approach outperformed standard models, showing how news signals reveal tensions ahead of co…” evidence ↗
1 expert 1 community 1 sources clustered

“This study presents a hybrid forecasting system merging time series analysis with NLP to predict Bolivian roadblocks, enhancing logistics and economic outcomes. This approach outperformed standard models, showing how news signals reveal tensions ahead of co…”

Developing Models & releases Research 1d ago
LLM consensus preference evaluation paper

A Consensus-Based Framework for Relative Preference Evaluation of Large Language Models

1 directory member surfaced this signal.

Why this matches Research & technical analysis reaction 1 attributable expert contribution · arxiv cs.CL
“Mohtashim Khan A Consensus-Based Framework for Relative Preference Evaluation of Large Language Models https://arxiv.org/abs/2607.21632” evidence ↗
1 expert 1 community 1 sources clustered

“Mohtashim Khan A Consensus-Based Framework for Relative Preference Evaluation of Large Language Models https://arxiv.org/abs/2607.21632”

Developing Evaluation & benchmarks Research 1d ago
biomedical MeSH NLP evaluation paper

Evaluation design conditions the expert-vs-auto MeSH gap: a controlled comparison of bag-of-words and BiomedBERT on the Cohen benchmark

1 directory member surfaced this signal.

Why this matches Research & technical analysis reaction 1 attributable expert contribution · arxiv cs.CL
“Samuel M. Okoe-Mensah Evaluation design conditions the expert-vs-auto MeSH gap: a controlled comparison of bag-of-words and BiomedBERT on the Cohen benchmark https://arxiv.org/abs/2607.21685” evidence ↗
1 expert 1 community 1 sources clustered

“Samuel M. Okoe-Mensah Evaluation design conditions the expert-vs-auto MeSH gap: a controlled comparison of bag-of-words and BiomedBERT on the Cohen benchmark https://arxiv.org/abs/2607.21685”

Developing Evaluation & benchmarks Research 1d ago
Khondo Bangla document benchmark

Khondo: A Multimodal Benchmark for Document Packet Splitting of Bangla Forms

1 directory member surfaced this signal.

Why this matches Research & technical analysis reaction 1 attributable expert contribution · arxiv cs.CL
“Abu Tyeb Azad, Fahim Ahmed, Ishita Sur Apan, Ezharuddin Jubaer, Sumaiya Karim Katha, Armun Alam, Amin Ahsan Ali, Aman Chadha, Md Mofijul Islam, AKM Mahbubur Rahman Khondo: A Multimodal Benchmark for Document Packet Splitting of Bangla Forms https://arxiv.or…” evidence ↗
1 expert 1 community 1 sources clustered

“Abu Tyeb Azad, Fahim Ahmed, Ishita Sur Apan, Ezharuddin Jubaer, Sumaiya Karim Katha, Armun Alam, Amin Ahsan Ali, Aman Chadha, Md Mofijul Islam, AKM Mahbubur Rahman Khondo: A Multimodal Benchmark for Document Packet Splitting of Bangla Forms https://arxiv.or…”

Developing Agents & robotics Research 1d ago
agentic copyright law evaluation paper

Agentic Evaluation of Copyright Law Compliance

1 directory member surfaced this signal.

Why this matches Research & technical analysis reaction 1 attributable expert contribution · arxiv cs.CL
“Zheng Hui, Doni Bloomfield, Noam Kolt Agentic Evaluation of Copyright Law Compliance https://arxiv.org/abs/2607.21799” evidence ↗
1 expert 1 community 1 sources clustered

“Zheng Hui, Doni Bloomfield, Noam Kolt Agentic Evaluation of Copyright Law Compliance https://arxiv.org/abs/2607.21799”

Developing Agents & robotics Research 1d ago
agent memory longitudinal evaluation paper

Ground Truth First: A Longitudinal Evaluation Instrument for Agent Memory, and the Tenure Crossover in Memory-Architecture Rankings

1 directory member surfaced this signal.

Why this matches Research & technical analysis reaction 1 attributable expert contribution · arxiv cs.CL
“Quentin Spencer Ground Truth First: A Longitudinal Evaluation Instrument for Agent Memory, and the Tenure Crossover in Memory-Architecture Rankings https://arxiv.org/abs/2607.21962” evidence ↗
1 expert 1 community 1 sources clustered

“Quentin Spencer Ground Truth First: A Longitudinal Evaluation Instrument for Agent Memory, and the Tenure Crossover in Memory-Architecture Rankings https://arxiv.org/abs/2607.21962”