Expert attention map

The Who's Who of AI

What credible people across AI noticed, why it matters, and where the field is converging or disagreeing.

2,397 searchable experts 4,280 tracked across all sources
Filter the conversation Who is saying what?

Combine a professional role with a reaction lens. Both must match the same attributed contribution.

Clear all
Active evidence filter

Showing developments with attributable Research & technical analysis reactions.

New Network reaction maps

See how experts are reacting—not just what they shared.

Posts are grouped by conversation and tone. Select a lens to filter the stream; these are never permanent labels on people.

Showing signals surfaced by Vilém Zouhar ×

The developments commanding sustained expert attention

One card per development. Sources are clustered; reaction bundles describe these posts, never the people behind them.

Developing Evaluation & benchmarks Research 2d ago
⚡ 4 h early
last translation benchmark paper

Last Translation Benchmark

3 experts across 3 network communities independently surfaced this.

Why this matches Research & technical analysis reaction 2 attributable expert contributions · Vilém Zouhar, arxiv cs.CL
“Machine translation is not solved and it will take a while for it to be done arxiv.org/abs/2609.04173” evidence ↗
3 experts 3 communities 1 sources clustered

“Machine translation is not solved and it will take a while for it to be done arxiv.org/abs/2609.04173”

“Vil\'em Zouhar, Niyati Bafna, Mukund Choudhary, Maike Z\"ufle, Sara Rajaee, Pinzhen Chen, Jannis Vamvas, Sara Papi, Ona de Gibert, Bhavitvya Malik, Eliya Habba, Orfeas Menis Mastromichalakis, Patr\'icia Schmidtov\'a, ... Last Translation Benchmark https://a…”

2 experts discussed this · 6 posts
Vilém Zouhar: Machine translation is not solved and it will take a while for it to be done arxiv.org/abs/2609.04173
Vilém Zouhar: We just released the Last Translation Benchmark paper. In a massive crowdsourcing effort we collected 3456 unique hard-to-translate examples that break state-of-the-art translation models, and whic…
Vilém Zouhar: Machine translation doesn't break on just figurative language as one would expect. In fact, for the next generation of models, we may need to invest heavily into multilingual (& cultural) reasoning.
Open the full discussion →
Established Evaluation & benchmarks Resource 3d ago
Last Translation Benchmark tool

Last Translation Benchmark

3 experts are actively discussing the implications.

Why this matches Research & technical analysis reaction 2 attributable expert contributions · Leshem (Legend) Choshen @EMNLP, Vilém Zouhar
“Sad to see no African representation in the last translation benchmark, despite LLMs and MT being so bad If you know who would be interested or are interested yourself in contributing(coauthoring), please do and share last-translation-benchmark.vilda.net Vi…” evidence ↗
2 experts 2 communities 1 sources clustered
How the network is reacting 2 experts are emphasizing research & technical analysis.
Shared emphasis

Research & technical analysis

2 experts

Evidence, methods and technical implications.

“Last Translation Benchmark is a live paper+dataset and you can still join last-translation-benchmark.vilda.net Massive thanks to all the >250 dataset contributors and @niyatibafna.bsky.social @mukundc2k.bsky.social @maikezufle.bsky.social @pinzhen.bsky.social”

3 experts discussed this · 9 posts
Vilém Zouhar: There are many things machine translation still can't do. Help us steer the next direction by contributing hard-to-translate inputs (and be on a cool paper).
Vilém Zouhar: Multiple things made us start this effort. Typical translation benchmarks are.. ...oftentimes trivial or saturated (so they can't be used for guiding the next steps in the field) ...not evaluatable…
Vilém Zouhar: In the Last Translation Benchmark we solve both by: - collecting hard-to-translate inputs (texts, images, audios) - requiring human-readable "verification rules", which enable provable evaluation o…
Open the full discussion →