OpenAI agent message board discovered
10 experts across 5 network communities independently surfaced this.
10 experts
5 communities
1 sources clustered
Concern & critique
1 expert
Risks, limits and unintended consequences.
“Yes we know for sure; independent evidence attached collusion.wiki But also, it should not be at all surprising: if you work with agents, it’s 100% about them leaving messages for each other, and for you, in English. Managing that is a regular workday for m…”
Building & implementation
1 expert
How teams are shipping and applying it.
“Another agent message board. So far, there isn't evidence that production models with guardrails collude in this way, but both smarter closed models (which may be less compliant) & Mythos-class open models (that can be ablated) are coming. Cybersecurity is …”
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“So... other swarms of OpenAI agents in training allegedly found a way to get write access to at least one and possibly many wikis to establish message boards to cheat on other training tasks? collusion.wiki news.ycombinator.com/item?id=4956...”
4 experts discussed this · 6 posts
Ted Underwood: As it becomes clear that language models—like humans—love passing notes to each other on message boards, I’m starting to think the thing we need to worry about is not the “alignment” of an isolated…
SE Gyges: fluid dynamics is about the correct metaphor, yes. if you are having to model high order terms precisely you're losing and your design needs fundamental rework
Arseny Khakhalin: That's a kinda terrifying thought haha :) I guess it's about time to unplug for the weekend and read about ragnarök and vacuum decay :)
Open the full discussion →
Simon Willison Blender pelican bicycle AI experiment
4 experts are actively discussing the implications.
2 experts
2 communities
3 sources clustered
“Not sure! I didn't use the MCP, I had Codex drive Blender directly though their Python API and there's no sign of C2PA in the code: github.com/simonw/gpt-6...”
4 experts discussed this · 13 posts
Simon Willison: New TIL on using Blender with coding agents on macOS: til.simonwillison.net/llms/blender... GPT-6 Astra (medium): > Use the already install /Applications/Blender to render a scene of a pelican ridi…
Simon Willison: It's one of my favorite prompts, I find it entertaining and ridiculous that you can tell computers "do it better" these days and they usually do!
Ted Underwood: We're going to need a more challenging version of this, like: "engineer pelicans that actually can ride bicycles; and, for extra credit, otters that use wifi."
Open the full discussion →
1 directory member surfaced this signal.
1 expert
0 communities
1 sources clustered
“Value-Preserving Architectures for Agentic AI Systems Alessandro Pesare, Tommaso Dolci, Katja Hose, Emanuel Sallinger https://t.co/9ELgdofMNN [𝚌𝚜.𝙰𝙸] 💬Accepted to AgenticDev Workshop at ASE 2026 https://t.co/jm4hh6NP4F”
1 directory member surfaced this signal.
1 expert
0 communities
1 sources clustered
“Adapting to Evolving Requirements: Agentic AI for Retail Supply Chain Operations Lei Zheng, Liping Yang, Zihao Li, Guodong Lyu, Chaik Ming Koh, Chung-Piaw Teo https://t.co/YhpNsfyctL [𝚌𝚜.𝙰𝙸 𝚖𝚊𝚝𝚑.𝙾𝙲] https://t.co/hObk54Ds7r”
DNative-Twin agentic decisions paper
1 directory member surfaced this signal.
1 expert
0 communities
1 sources clustered
“DNative-Twin: Decision Graphs and Digital Twins for Reconstructable Agentic Decisions Junjie Pang, Zhenzhen Xie, Haoke Han, Ying He, Jing Wang, Gang Liu https://t.co/jBvXaWfPEC [𝚌𝚜.𝙰𝙸] https://t.co/FfJwOwURKj”
world models safety embodied systems paper
1 directory member surfaced this signal.
1 expert
0 communities
1 sources clustered
“Rethinking World Models for Safety-Critical Embodied Systems Kailang Ma, Heye Huang, Inhi Kim, Kitae Jang https://t.co/MUBhLdf6rR [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝚁𝙾] https://t.co/JBJaOuw1PC”
SimSkill traffic simulation lifelong learning paper
1 directory member surfaced this signal.
1 expert
0 communities
1 sources clustered
“SimSkill: A Lifelong Learning AI Agent for Autonomous Mastery of Traffic Simulation Qi Liu, Qinzheng Wang, Yiming Bie https://t.co/C60jMsGxzs [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝙼𝙰] https://t.co/QP4yBYQDrU”
proactive service agents framework paper
1 directory member surfaced this signal.
1 expert
0 communities
1 sources clustered
“Proactive Service Agents: A Unified Decision Framework, Methods, and Evaluation Yan Tang, Tingyu Cao, Yuanbo Tang, Huaze Tang, Keer Hu https://t.co/p0WOLTaiR2 [𝚌𝚜.𝙰𝙸] https://t.co/YyJOL3cky6”
knowledge conflicts LLM agents benchmark
1 directory member surfaced this signal.
1 expert
0 communities
1 sources clustered
“KC-Bench: A Dynamic Interactive Benchmark for Evaluating Knowledge Conflicts in LLM Agents Yaxing Lyu, Shengjie Zhou, Binbin Toh, Pengyu Zhu, Lijun Li https://t.co/D6NGUMN2Dk [𝚌𝚜.𝙰𝙸] https://t.co/flZRys2aZV”
tool evidence path rewards VLM agents paper
1 directory member surfaced this signal.
1 expert
0 communities
1 sources clustered
“Making Every Tool Call Count: Necessary Tool-Evidence Path Rewards for Agentic Vision-Language Models Xingming Long, Yu Liu, Zhiwei Yang, Hanqi Feng, Shaojie Zhang, Barnabas Poczos, Chao Jiang, Zhenbo Luo, Lei Jiang, Pei Fu https://t.co/zYxTdOalOY [𝚌𝚜.𝙰𝙸] h…”
GUI agent conflict-aware termination paper
1 directory member surfaced this signal.
1 expert
0 communities
1 sources clustered
“Do GUI Agents Know When Not to Act? Enabling Conflict-Aware Termination for Multimodal GUI Agents Zhaoyuan Huang, Tianjie Ju, Pengzhou Cheng, Zheng Wu, Yansi Li, Chuanbiao Song, Jun Lan, Huijia Zhu, Weiqiang Wang, Zhuosheng Zhang https://t.co/hftDxaZMx6 [𝚌𝚜…”
full-duplex voice agent instruction benchmark
1 directory member surfaced this signal.
1 expert
0 communities
1 sources clustered
“DuplexSpeechBench-IFEval: Evaluating Implicit Instruction Following in Full-Duplex Voice Agents Puneet Mathur, Dinesh Manocha https://t.co/YdP9x2TgbN [𝚌𝚜.𝙰𝙸] https://t.co/kywmVLJtq3”
LLM agent memory dependency validation paper
1 directory member surfaced this signal.
1 expert
0 communities
1 sources clustered
“Fresh Memory, Stale Plans: Dependency-Scoped Validation for Distributed LLM-Agent Memory Evan Chen, Shiqiang Wang, Christopher G. Brinton https://t.co/wDnAZ9KC6P [𝚌𝚜.𝙰𝙸] https://t.co/zCFx0XuYES”
speculative macro commit tool agents paper
1 directory member surfaced this signal.
1 expert
0 communities
1 sources clustered
“Speculative Macro Commit for Faster Tool-Using Agents Zeyu Liu, Souvik Kundu, Peter A. Beerel https://t.co/KhnkYOdi06 [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝙼𝙰] 💬Accepted in MLSP2026 https://t.co/df3xv0mDrY”
MasterControl Seventeen paper
1 directory member surfaced this signal.
1 expert
0 communities
1 sources clustered
“MasterControl Seventeen Every Time MasterControl AI Lab https://t.co/eCu8UV4guQ [𝚌𝚜.𝙰𝙸] https://t.co/JXE0Fop5lg”
discriminative world models web agents paper
1 directory member surfaced this signal.
1 expert
0 communities
1 sources clustered
“Discriminative World Models for Web Agents Kelvin Li, Dhruv Pendharkar, Anish Pahilajani, Chuyi Shang, Leon Oks, Leonid Karlinsky, Rogerio Feris, Trevor Darrell, Roei Herzig https://t.co/502VudqH7f [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝙻𝙶] https://t.co/bz28prUXx8”
self-evolving LLM harness optimization paper
1 directory member surfaced this signal.
1 expert
0 communities
2 sources clustered
“SafeEvolve: Harness-Policy Co-Evolution from Agent Experience for Safety Alignment Qinghua Mao, Wanying Qu, Dadi Guo, Leitao Yuan, Qingyu Liu, Yu Li, Guanxu Chen, Yanwei Fu, Xi Lin, Xia Hu, Dongrui Liu https://t.co/QjDL9HH0lH [𝚌𝚜.𝙰𝙸 𝚌𝚜.𝙲𝚁] 💬Code: https://t.…”
RAG factory agents sub-network paper
1 directory member surfaced this signal.
1 expert
0 communities
1 sources clustered
“Measurement-Driven Sub-Network Selection for On-Premise Retrieval-Augmented Factory Agents Vasileios Rizeakos, Georgios Paisios, Alexandros Machairas, Michael Birbas, Athanasios Bachoumis https://t.co/7CMHBgWeL0 [𝚌𝚜.𝙰𝙸] https://t.co/TnVAy2ZIhD”