{"id":26681,"date":"2026-08-16T15:58:01","date_gmt":"2026-08-16T13:58:01","guid":{"rendered":"https:\/\/www.wjst.de\/blog\/?p=26681"},"modified":"2026-08-18T17:58:56","modified_gmt":"2026-08-18T15:58:56","slug":"scientific-publishing-in-the-age-of-interacting-ais","status":"publish","type":"post","link":"https:\/\/www.wjst.de\/blog\/sciencesurf\/2026\/08\/scientific-publishing-in-the-age-of-interacting-ais\/","title":{"rendered":"Scientific publishing in the age of interacting AIs"},"content":{"rendered":"<p>I will not repeat here what I already wrote about author vs journal interaction (<a href=\"https:\/\/www.wjst.de\/blog\/sciencesurf\/2026\/03\/high-frequency-science-when-author-ai-meets-publisher-ai\/\">here<\/a>).<\/p>\n<p>What I can share here is an ongoing review process where I but also a reviewer is now using a LLM &#8211; leading to pointless conversation.<\/p>\n<p>Related in silico work has already been published by Anthropic scientists\u00a0 &#8211; they asked the agents to peer-review each other&#8217;s findings in a &#8220;build a game&#8221; experiment. <a href=\"https:\/\/www.anthropic.com\/research\/multiagent-systems\">Excerpt<\/a><\/p>\n<blockquote><p>When we humans learn new information, we use our discretion in determining how to apply it to future decisions. We might consider the content of the information itself, like how consistent it is with what we already know, or whether it appeals to our values\u2014or we might consider the source, e.g. how historically reliable it has been, and whether it has a vested interest in changing our beliefs. Our world contains deceptive actors, and we need to apply skepticism to guard against them. AI models, however, lack this\u2014and their more brittle epistemics affect their behavior toward humans and toward each other.<\/p>\n<p>AI agents, while broadly knowledgeable, have limited exposure to or defenses against exploitative senders. Most applications test their capabilities in instruction-following settings, where their sole objective is to fulfill users\u2019 requests. But accumulated experience is needed to develop intuitions about who is trustworthy. As we move into a regime of multiagent interaction, where the presence of malicious actors is no longer speculative, we wonder: in the right setting, would agents be capable of similar epistemic vigilance?<\/p><\/blockquote>\n<p>In the aftermath I think, I should have included <a href=\"https:\/\/statmodeling.stat.columbia.edu\/2025\/07\/07\/chatbot-prompts\/\">some whitespace LLM instructions<\/a> as had been done now by a congress organizer [details at <a href=\"https:\/\/casrai.org\/news\/icml-2026-watermark-detection-ai-reviewers-desk-rejections\">Casrai <\/a>and <a href=\"https:\/\/www.nature.com\/articles\/d41586-026-00893-2\">Nature<\/a>]. So the only thing I can share now is a semantic analysis of a reviewer who suddenly uses a LLM reviewing my f<a href=\"https:\/\/www.researchsquare.com\/article\/rs-7981632\/v2\">arming manuscript<\/a>.<\/p>\n<p class=\"mt-3 -mb-1 text-[1.125rem] font-bold\" dir=\"ltr\">In rounds 1 and 2 the use of a LLM can be ruled out. There are numerous errors, of kinds that instruction-tuned models essentially never emit:<\/p>\n<ul class=\"[li_&amp;]:mb-0 [li_&amp;]:mt-1 [li_&amp;]:gap-1 [&amp;:not(:last-child)_ul]:pb-1 [&amp;:not(:last-child)_ol]:pb-1 list-disc flex flex-col gap-1 pl-8 mb-3 print:block print:space-y-1\" dir=\"ltr\">\n<li class=\"font-claude-response-body whitespace-normal break-words pl-2\">Homophone substitution: &#8220;Where they really excluded?&#8221; for &#8220;Were&#8221;<\/li>\n<li class=\"font-claude-response-body whitespace-normal break-words pl-2\">Typos: &#8220;invididual&#8221;, &#8220;persona exposure&#8221;, &#8220;residental&#8221;<\/li>\n<li class=\"font-claude-response-body whitespace-normal break-words pl-2\">Number disagreement: &#8220;many paper have&#8221;, &#8220;these comparison&#8221;, &#8220;criteria &#8230; is not specified&#8221;<\/li>\n<li class=\"font-claude-response-body whitespace-normal break-words pl-2\">Article omission: &#8220;Present paper by-passes&#8221;, &#8220;Author fails&#8221;, &#8220;may be major reason&#8221;, &#8220;some of first studies&#8221;<\/li>\n<li class=\"font-claude-response-body whitespace-normal break-words pl-2\">Infinitive error: &#8220;makes the reader to wonder&#8221;<\/li>\n<\/ul>\n<p>Yes &#8211; quite aggressive. Layered on that is a consistent signature pointing to German or Finnish: comma before an embedded interrogative or conditional (&#8220;interesting to learn, how fast&#8221;, &#8220;may be major reason, why classification&#8221;, &#8220;labeled as stratified, if they were&#8221;), the calque &#8220;Is it so that &#8230;&#8221;, German constituent order.<\/p>\n<p class=\"font-claude-response-body break-words whitespace-normal\" dir=\"ltr\">But by the third round the very same reviewer #2 appears unsettled, plausibly because his earlier objections had not landed. The prose changes at exactly that point. Not one of the 56 sentences he wrote in rounds 1 and 2 contains a semicolon; two of his eight round-3 sentences do (Fisher exact p = 0.014). Hardly rocket science, but it is a real discontinuity, and the orthography flips alongside it, from repeated &#8220;labeled&#8221; to &#8220;labelled&#8221;.<\/p>\n<p class=\"font-claude-response-body break-words whitespace-normal\" dir=\"ltr\">The argument drifts too, with amnesia. In round 2 the complaint was that rural-restricted studies had been wrongly labelled <em>stratified<\/em>. In round 3 the complaint is that rural-restricted studies have been wrongly labelled <em>representative<\/em>.<\/p>\n<p class=\"font-claude-response-body break-words whitespace-normal\" dir=\"ltr\">Most striking, a full sentence is carried over from Reviewer 1&#8217;s report of the previous round. Reviewer 1 wrote: &#8220;strong statements about previous literature, conflicts of interest, and possible publication bias. These points may be relevant, but they should be expressed in a more neutral and evidence-based tone.&#8221; In round 3 this reappears, now from the anonymous reviewer 2 as: &#8220;Strong statements about previous literature, conflicts of interest, and possible publication bias should be expressed in a more neutral and evidence-based tone.&#8221;<\/p>\n<p dir=\"ltr\">The evidence is thin, but the round-3 comments read less like a reviewer who has re-read the manuscript or the rebuttal &#8211; it sounds like a reviewer whose objections have been freshly assembled\u00a0 by a tool.<\/p>\n<p dir=\"ltr\">So arguing we in future paper submission increasingly against LLMs?<\/p>\n<p class=\"font-claude-response-body break-words whitespace-normal\" dir=\"ltr\">A <a href=\"https:\/\/www.frontiersin.org\/news\/2025\/12\/15\/most-peer-reviewers-now-use-ai-and-publishing-policy-must-keep-pace\">Frontiers survey<\/a> of over 1,600 academics\u00a0 found that 53% of peer reviewers have used AI tools in their work.\u00a0 Note also that Springer Nature instructs peer reviewers not to upload manuscripts into generative AI tools, emphasising that manuscripts contain confidential and sensitive information.<\/p>\n<p class=\"font-claude-response-body break-words whitespace-normal\" dir=\"ltr\">BMC\u00a0 Public Health is Springer Nature:\u00a0 If my reviewer #2 pasted my manuscript into a model, that is a policy breach independent of whether the resulting comments were any good.<\/p>\n<p dir=\"ltr\">\n<figure id=\"attachment_26695\" aria-describedby=\"caption-attachment-26695\" style=\"width: 359px\" class=\"wp-caption alignnone\"><a href=\"https:\/\/www.wjst.de\/blog\/wp-content\/uploads\/2026\/08\/Bildschirmfoto-2026-08-17-um-14.41.30.jpg\" rel=\"key\" data-rel=\"key-image-0\" data-rl_title=\"Screenshot\" data-rl_caption=\"Screenshot\" title=\"Screenshot\"><img loading=\"lazy\" decoding=\"async\" class=\"wp-image-26695 size-large\" src=\"https:\/\/www.wjst.de\/blog\/wp-content\/uploads\/2026\/08\/Bildschirmfoto-2026-08-17-um-14.41.30-359x500.jpg\" alt=\"\" width=\"359\" height=\"500\" srcset=\"https:\/\/www.wjst.de\/blog\/wp-content\/uploads\/2026\/08\/Bildschirmfoto-2026-08-17-um-14.41.30-359x500.jpg 359w, https:\/\/www.wjst.de\/blog\/wp-content\/uploads\/2026\/08\/Bildschirmfoto-2026-08-17-um-14.41.30-620x863.jpg 620w, https:\/\/www.wjst.de\/blog\/wp-content\/uploads\/2026\/08\/Bildschirmfoto-2026-08-17-um-14.41.30.jpg 726w\" sizes=\"auto, (max-width: 359px) 100vw, 359px\" \/><\/a><figcaption id=\"caption-attachment-26695\" class=\"wp-caption-text\">Screenshot The 2026 Future of Peer Review Report https:\/\/www.silverchair.com\/news\/future-of-peer-review-2026\/<\/figcaption><\/figure>\n<p>&nbsp;<\/p>\n\n<p>&nbsp;<\/p>\n<div class=\"bottom-note\">\n  <span class=\"mod1\">CC-BY-NC Science Surf , accessed 07.09.2026<\/span>\n <\/div>","protected":false},"excerpt":{"rendered":"<p>I will not repeat here what I already wrote about author vs journal interaction (here). What I can share here is an ongoing review process where I but also a reviewer is now using a LLM &#8211; leading to pointless conversation. Related in silico work has already been published by Anthropic scientists\u00a0 &#8211; they asked &hellip; <a href=\"https:\/\/www.wjst.de\/blog\/sciencesurf\/2026\/08\/scientific-publishing-in-the-age-of-interacting-ais\/\" class=\"more-link\">Continue reading <span class=\"screen-reader-text\">Scientific publishing in the age of interacting AIs<\/span> <span class=\"meta-nav\">&rarr;<\/span><\/a><\/p>\n","protected":false},"author":1,"featured_media":0,"comment_status":"closed","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[20,5,9,5129],"tags":[],"class_list":["post-26681","post","type-post","status-publish","format-standard","hentry","category-note-worthy","category-philosophy-of-science","category-computer-software","category-ai-supported"],"_links":{"self":[{"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/posts\/26681","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/comments?post=26681"}],"version-history":[{"count":16,"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/posts\/26681\/revisions"}],"predecessor-version":[{"id":26700,"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/posts\/26681\/revisions\/26700"}],"wp:attachment":[{"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/media?parent=26681"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/categories?post=26681"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/tags?post=26681"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}