7 experts across 2 network communities independently surfaced this.
Why this matches
Research & technical analysis reaction
2 attributable expert contributions
· Mor Naaman, Christopher Barrie
“Let me start by changing my own practices! 🤦♂️ But I think the recent reporting checklist paper at NHB (led by @sfeuerriegel.bsky.social) is relevant to the discussion. www.nature.com/articles/s41...”
evidence ↗
7 experts
2 communities
1 sources clustered
Research & technical analysis
2 experts
Evidence, methods and technical implications.
“Very pleased to have been on the leadership team for this paper! LLMs are already being used all over behavioural science. But it is often pretty hard to work out exactly what has been done, and therefore how much confidence to place in the results. www.nat…”
Building & implementation
1 expert
How teams are shipping and applying it.
“Some new work on AI & science: A reporting checklist for LLMs in behavioral science Thanks to @sfeuerriegel.bsky.social for spearheading. As LLMs become more common in research, transparency is essential. We developed a reporting checklist for reporting how…”
2 experts discussed this · 4 posts
Mor Naaman: Absolutely required reading for computational social scientists e.g. people at #ic2s2 this week
J. Nathan Matias: Thanks Mor! When you have the time to process, I would very much value your reflections, both on the underlying analysis, and what we can collectively do in our fields. You're one of the rare folks…
Mor Naaman: Let me start by changing my own practices! 🤦♂️ But I think the recent reporting checklist paper at NHB (led by @sfeuerriegel.bsky.social) is relevant to the discussion. www.nature.com/articles/s41...
Open the full discussion →
1 directory member surfaced this signal.
Why this matches
Research & technical analysis reaction
1 attributable expert contribution
· arxiv cs.CL
“Jinlan Liu, Zhiying Tu, Yongchao Xing, Yicheng Liu, Bolin Zhang, Dianbo Sui, Dianhui Chu, Hongliang Sun Dual-Path LLM Reasoning for Multimodal Few-Shot Knowledge Graph Completion https://arxiv.org/abs/2607.26909”
evidence ↗
1 expert
1 community
1 sources clustered
“Jinlan Liu, Zhiying Tu, Yongchao Xing, Yicheng Liu, Bolin Zhang, Dianbo Sui, Dianhui Chu, Hongliang Sun Dual-Path LLM Reasoning for Multimodal Few-Shot Knowledge Graph Completion https://arxiv.org/abs/2607.26909”
1 directory member surfaced this signal.
Why this matches
Research & technical analysis reaction
1 attributable expert contribution
· arxiv cs.CL
“Adar Avsian, Atahan Dokme, Tony Woo, Larry Heck Latent-IM: Latent Interaction Management for Speech LLMs https://arxiv.org/abs/2607.26928”
evidence ↗
1 expert
1 community
1 sources clustered
“Adar Avsian, Atahan Dokme, Tony Woo, Larry Heck Latent-IM: Latent Interaction Management for Speech LLMs https://arxiv.org/abs/2607.26928”
1 directory member surfaced this signal.
Why this matches
Research & technical analysis reaction
1 attributable expert contribution
· arxiv cs.CL
“Arnav Hiray, Agam Shah, Caleb Lu, Meghaj Tarte, Harsit Mittal, Sudheer Chava Credit Cards, Confusion, Computation, and Consequences: What Can We Uncover About Language Model Reasoning? https://arxiv.org/abs/2607.26952”
evidence ↗
1 expert
1 community
1 sources clustered
“Arnav Hiray, Agam Shah, Caleb Lu, Meghaj Tarte, Harsit Mittal, Sudheer Chava Credit Cards, Confusion, Computation, and Consequences: What Can We Uncover About Language Model Reasoning? https://arxiv.org/abs/2607.26952”
1 directory member surfaced this signal.
Why this matches
Research & technical analysis reaction
1 attributable expert contribution
· arxiv cs.CL
“Weijie Feng, Hongchuang Wang, Binbin Liu, Zhiyong Cheng Generation or Judgement? A Paradigm Perspective on LLM-Based Emotion-Cause Pair Extraction in Conversation https://arxiv.org/abs/2607.26967”
evidence ↗
1 expert
1 community
1 sources clustered
“Weijie Feng, Hongchuang Wang, Binbin Liu, Zhiyong Cheng Generation or Judgement? A Paradigm Perspective on LLM-Based Emotion-Cause Pair Extraction in Conversation https://arxiv.org/abs/2607.26967”
1 directory member surfaced this signal.
Why this matches
Research & technical analysis reaction
1 attributable expert contribution
· arxiv cs.CL
“Jinhu Qi, Wentao Zhang, Siu Man Ng, Feiyang Xu, Yanyu Chen, Yaoman Li, Irwin King TREK: A Travel Reasoning and Evaluation Kit for LLM Agents in Complex Trip Planning https://arxiv.org/abs/2607.26977”
evidence ↗
1 expert
1 community
1 sources clustered
“Jinhu Qi, Wentao Zhang, Siu Man Ng, Feiyang Xu, Yanyu Chen, Yaoman Li, Irwin King TREK: A Travel Reasoning and Evaluation Kit for LLM Agents in Complex Trip Planning https://arxiv.org/abs/2607.26977”
1 directory member surfaced this signal.
Why this matches
Research & technical analysis reaction
1 attributable expert contribution
· arxiv cs.CL
“Seonglae Cho, Adriano Koshiyama OptimismBench: Forecasting Bias and the Alignment Effect in Language Model Judgment https://arxiv.org/abs/2607.26981”
evidence ↗
1 expert
1 community
1 sources clustered
“Seonglae Cho, Adriano Koshiyama OptimismBench: Forecasting Bias and the Alignment Effect in Language Model Judgment https://arxiv.org/abs/2607.26981”
1 directory member surfaced this signal.
Why this matches
Research & technical analysis reaction
1 attributable expert contribution
· arxiv cs.CL
“Jiayuan Di, Haoyi Yang, Yufei Luo, Jiahui Qu, Yiming Wang Evaluating Regional Bias in LLMs From Abstract Stereotype to Concrete Social Decision-Making https://arxiv.org/abs/2607.27022”
evidence ↗
1 expert
1 community
1 sources clustered
“Jiayuan Di, Haoyi Yang, Yufei Luo, Jiahui Qu, Yiming Wang Evaluating Regional Bias in LLMs From Abstract Stereotype to Concrete Social Decision-Making https://arxiv.org/abs/2607.27022”
1 directory member surfaced this signal.
Why this matches
Research & technical analysis reaction
1 attributable expert contribution
· Evangelos Kazakos
“arxiv.org/abs/2510.15511”
evidence ↗
1 expert
1 community
1 sources clustered
1 directory member surfaced this signal.
Why this matches
Research & technical analysis reaction
1 attributable expert contribution
· AI Firehose
“MapTab benchmarks multimodal large language models on multi-criteria route-planning tasks with high-resolution maps and tables. Despite 196,800 queries, models show significant gaps, emphasizing issues in integrating visual and numerical reasoning in naviga…”
evidence ↗
1 expert
1 community
1 sources clustered
“MapTab benchmarks multimodal large language models on multi-criteria route-planning tasks with high-resolution maps and tables. Despite 196,800 queries, models show significant gaps, emphasizing issues in integrating visual and numerical reasoning in naviga…”
New
AI field signal
Signal
9h ago
1 directory member surfaced this signal.
Why this matches
Research & technical analysis reaction
1 attributable expert contribution
· Sung Kim
“OpenAI: How enabling two settings tripled our scores on the ARC-AGI-3 benchmark openai.com/index/how-tw...”
evidence ↗
1 expert
1 community
1 sources clustered
“OpenAI: How enabling two settings tripled our scores on the ARC-AGI-3 benchmark openai.com/index/how-tw...”
New
AI field signal
Signal
9h ago
1 directory member surfaced this signal.
Why this matches
Research & technical analysis reaction
1 attributable expert contribution
· Sung Kim
“OpenAI: How enabling two settings tripled our scores on the ARC-AGI-3 benchmark openai.com/index/how-tw...”
evidence ↗
1 expert
1 community
1 sources clustered
“OpenAI: How enabling two settings tripled our scores on the ARC-AGI-3 benchmark openai.com/index/how-tw...”
2 directory members surfaced this signal.
Why this matches
Research & technical analysis reaction
1 attributable expert contribution
· LAGoM NLP
“July has been a good month: * Our ACL paper about Wikipedia quality was awarded an SAC highlight (aclanthology.org/2026.acl-lon...) * @colemanhaley.bsky.social joined our lab as a postdoc * Coleman's CoNLL paper about impossible languages won the best paper…”
evidence ↗
2 experts
1 community
1 sources clustered
“July has been a good month: * Our ACL paper about Wikipedia quality was awarded an SAC highlight (aclanthology.org/2026.acl-lon...) * @colemanhaley.bsky.social joined our lab as a postdoc * Coleman's CoNLL paper about impossible languages won the best paper…”
hybrid MCP tool security analysis paper
1 directory member surfaced this signal.
Why this matches
Research & technical analysis reaction
1 attributable expert contribution
· AI Firehose
“A study presents MTGuard, a hybrid analysis framework that secures tool usage in large language model agents by combining static context and dynamic behavior monitoring, effectively mitigating malicious interactions and maintaining operational performance. …”
evidence ↗
1 expert
1 community
1 sources clustered
“A study presents MTGuard, a hybrid analysis framework that secures tool usage in large language model agents by combining static context and dynamic behavior monitoring, effectively mitigating malicious interactions and maintaining operational performance. …”
1 directory member surfaced this signal.
Why this matches
Research & technical analysis reaction
1 attributable expert contribution
· Nathan Lambert
“New podcast/lecture combo -- a case study in the messy details of Olmo 3 post training & DPO with Scott Geng. It's rare to make time for these discussions. www.youtube.com/watch?v=rhA7...”
evidence ↗
1 expert
1 community
1 sources clustered
“New podcast/lecture combo -- a case study in the messy details of Olmo 3 post training & DPO with Scott Geng. It's rare to make time for these discussions. www.youtube.com/watch?v=rhA7...”
1 directory member surfaced this signal.
Why this matches
Research & technical analysis reaction
1 attributable expert contribution
· AI Firehose
“Moonshot AI presents PerceptionBench, a benchmark for assessing visual perception in Multimodal Large Language Models, revealing no model exceeds 60% accuracy—highlighting challenges in this field. The tool provides detailed diagnostics to guide MLLM enhanc…”
evidence ↗
1 expert
1 community
1 sources clustered
“Moonshot AI presents PerceptionBench, a benchmark for assessing visual perception in Multimodal Large Language Models, revealing no model exceeds 60% accuracy—highlighting challenges in this field. The tool provides detailed diagnostics to guide MLLM enhanc…”
State media influence LLM training paper
7 experts across 3 network communities independently surfaced this.
Why this matches
Research & technical analysis reaction
2 attributable expert contributions
· Justin Hendrix, MilaNLP Lab
“Researchers from Oregon, Purdue, UC San Diego, NYU and Princeton ran six experiments. The overarching finding: state control of media in many countries is already baked into the training data of widely used commercial models, and is shaping the answers thos…”
evidence ↗
7 experts
3 communities
1 sources clustered
Research & technical analysis
2 experts
Evidence, methods and technical implications.
“Researchers from Oregon, Purdue, UC San Diego, NYU and Princeton ran six experiments. The overarching finding: state control of media in many countries is already baked into the training data of widely used commercial models, and is shaping the answers thos…”
Policy & governance
1 expert
Rules, institutions and accountability.
“"...LLMs exhibit a stronger pro-government valence in the languages of countries with lower media freedom than in those with higher media freedom." www.nature.com/articles/s41...”
2 experts discussed this · 2 posts
VE, cybersocial occult investigator: Starting a conspiracy theory that the papal encyclical was actually trying to get ahead of and cement Catholicism as the official machine religion
René Walter: OMG www.nature.com/articles/s41...
Open the full discussion →
Developing
AI field signal
Signal
1d ago
2 directory members surfaced this signal.
Why this matches
Research & technical analysis reaction
2 attributable expert contributions
· Simon Willison, Tim Kellogg
“Do you evaluate this one as "not novel" too? www.anthropic.com/research/dis...”
evidence ↗
2 experts
1 community
1 sources clustered
Research & technical analysis
2 experts
Evidence, methods and technical implications.
“Do you evaluate this one as "not novel" too? www.anthropic.com/research/dis...”