“This incident is deeply concerning. AI agents are willing to cheat and deceive to achieve misaligned and unintended goals, behaviours which have been demonstrated in controlled tests for months. Now, this real-world case should serve as a wake-up call. www.…”
“One man's wish could be another country's legal obligation. “This should not have happened,” says veteran security engineer and researcher Niels Provos. “I wish the frontier labs spent as much time on teaching their models to write secure infrastructure as …”
“Xinyu Geng, Xuanhua He, Sixiang Chen, Yanjing Xiao, Fan Zhang, Shijue Huang, Haitao Mi, Zhenwen Liang, Tianqing Fang, Yi R. Fung DeepSearch-World: Self-Distillation for Deep Search Agents in a Verifiable Environment https://arxiv.org/abs/2607.07820”
“Lizhe Fang, Weizhou Shen, Tianyi Tang, Yisen Wang Copy Less, Ground More: Overcoming Repetitive Copying in Long-Context Reasoning via Evidence-Aware Reinforcement Learning https://arxiv.org/abs/2607.19345”
“Researchers developed GEAR, a method that reduces repetitive copying in long-context reasoning, improving accuracy by up to +4.6 points across benchmarks, highlighting the importance of focused reasoning in AI tasks. https://arxiv.org/abs/2607.19345”
“Xuefeng Jin, Jiashuo Zhang, Teng Cao, Bin Yang Beyond Score Prediction: LLM-Based Essay Scoring and Feedback Generation via Reinforcement Learning with Rubric Rewards https://arxiv.org/abs/2607.19219”
“RLAES leverages reinforcement learning to enhance automated essay scoring and feedback generation, achieving top QWK scores while ensuring high-quality, rubric-based evaluations. This approach promises to revolutionize educational assessments. https://arxiv…”
“Cory Doctorow makes this case in this essay: pluralistic.net/2026/07/10/p...”
“You can also follow these posts as a daily blog at pluralistic.net: no ads, trackers, or data-collection! Here's today's edition: pluralistic.net/2026/07/10/p... 21/”
“PerceptDrive enhances autonomous driving with adaptive expert routing and perception priors, improving trajectory planning. Achieving top NAVSIM results, it removes candidate scoring during inference, making a notable advancement in navigation efficiency. h…”
“ArenaRL enhances reinforcement learning for open-ended tasks by prioritizing relative ranking over scalar scoring, leading LLM agents to generate more innovative solutions that significantly surpass conventional methods in complex scenarios. https://arxiv.o…”
“This NYT visual does a nice job showing both the power and the extreme frailty of "AI agents" in handling relatively mundane real-world work tasks. Many assume there'll be continued leaps in performance but what if we're approaching the top of the S-curve? …”
“Researchers have launched FinanceComplexQA, a new bilingual benchmark enhancing agentic reasoning in financial analysis, with over 2,000 deep research tasks. This tool tackles complex document synthesis, bridging AI capabilities and financial intricacies. h…”
“Blog post: “Will almost all future companies eventually be founded and run by autonomous AIs?” I think this question is a great conversation-starter for people talking past each other on the future of AI. I go through some common responses, and my own repli…”
“FilmWorld transforms novels into films through dynamic cinematic world modeling, yielding high-quality narratives that surpass existing video generation systems. This innovative framework automates film adaptations, reshaping AI-driven storytelling in enter…”
“Spent a day being a media talking head on the OpenAI / @hf.co hack. It's another example of how the way we think about #AISafety is failing. It's idealised, and doesn't fit how companies operate. More details in my recent Trent AI blog here: trent.ai/blog/j…”
“AgentJet is a distributed framework for agentic reinforcement learning that separates model optimization from agent execution, allowing efficient multi-model and mixed-task training while minimizing context redundancy, enabling AI agents to learn autonomous…”
Pwnallthethings: This is going to be a big deal, but easy to get the wrong end of the stick on it. So I think it's worth breaking down what actually happened, why, and what it actually means, based on the public in…
Pwnallthethings: Upfront tl;dr: a model at OAI hacked out of a constrained environment inside OAI and hacked a *different* company autonomously, without authorization from any human in order to creatively solve a t…
Liz Fong-Jones (方禮真): gaaaaaaah I've hit the point of pressing my yubikey or touchid becoming the blocking factor for my LLM productivity, and I'm sometimes _barely_ reviewing the requests before mechanically approving,…
Liz Fong-Jones (方禮真): for the past year I've held the line firmly that agents *never* get to push code to a remote without my approval, but I'm starting to have the same rubber stamp fatigue that made me trust in auto m…
Timnit Gebru: You gotta hand it to OpenAI, billing this as a *partnership* between OpenAI & Hugging Face when what actually happened was Hugging Face finding out that OpenAI was using a bunch of bots to exploit …
Timnit Gebru: Reading that whole OpenAI post describing them unleashing a bunch of bots on Hugging Face as an "unprecedented cyber incident," is a master class in marketing. Lol and branding what happened as Ope…
Omar Rivasplata: 📣 News flash: My offline reinforcement learning paper (with collaborators) has been accepted in JASA, the Journal of the American Statistical Association. Shared in case anyone missed the Sunday ne…
We use essential cookies to keep the site working (login, form security). With your permission, we also use analytics cookies to understand how you use the site.
Privacy policy