{"id":26681,"date":"2026-08-16T15:58:01","date_gmt":"2026-08-16T13:58:01","guid":{"rendered":"https:\/\/www.wjst.de\/blog\/?p=26681"},"modified":"2026-08-16T15:58:01","modified_gmt":"2026-08-16T13:58:01","slug":"scientific-publishing-in-the-age-of-interacting-ais","status":"publish","type":"post","link":"https:\/\/www.wjst.de\/blog\/sciencesurf\/2026\/08\/scientific-publishing-in-the-age-of-interacting-ais\/","title":{"rendered":"Scientific publishing in the age of interacting AIs"},"content":{"rendered":"<p>I will not repeat here what I already wrote about author vs journal interaction <a href=\"https:\/\/www.wjst.de\/blog\/sciencesurf\/2026\/03\/high-frequency-science-when-author-ai-meets-publisher-ai\/\">where I predicted<\/a> high speed transactions.<\/p>\n<p>What I can share here is an ongoing review process where a reviewer is now using a tool giving me an increasing workload.<\/p>\n<p>Related in silico work has been published by Anthropic scientists recently who asked the agents to peer-review each other&#8217;s findings in a &#8220;build a game&#8221; experiment. <a href=\"https:\/\/www.anthropic.com\/research\/multiagent-systems\">They write<\/a><\/p>\n<blockquote><p>When we humans learn new information, we use our discretion in determining how to apply it to future decisions. We might consider the content of the information itself, like how consistent it is with what we already know, or whether it appeals to our values\u2014or we might consider the source, e.g. how historically reliable it has been, and whether it has a vested interest in changing our beliefs. Our world contains deceptive actors, and we need to apply skepticism to guard against them. AI models, however, lack this\u2014and their more brittle epistemics affect their behavior toward humans and toward each other.<\/p>\n<p>AI agents, while broadly knowledgeable, have limited exposure to or defenses against exploitative senders. Most applications test their capabilities in instruction-following settings, where their sole objective is to fulfill users\u2019 requests. But accumulated experience is needed to develop intuitions about who is trustworthy. As we move into a regime of multiagent interaction, where the presence of malicious actors is no longer speculative, we wonder: in the right setting, would agents be capable of similar epistemic vigilance?<\/p><\/blockquote>\n<p>In the aftermath of the following review I should have included <a href=\"https:\/\/statmodeling.stat.columbia.edu\/2025\/07\/07\/chatbot-prompts\/\">some whitespace LLM instructions<\/a> as had been done now by a congress organizer [<a href=\"https:\/\/casrai.org\/news\/icml-2026-watermark-detection-ai-reviewers-desk-rejections\">Casrai<\/a>, <a href=\"https:\/\/www.nature.com\/articles\/d41586-026-00893-2\">Nature<\/a>]. So the only thing I can share now is a semantic analysis of a reviewer who comments on my f<a href=\"https:\/\/www.researchsquare.com\/article\/rs-7981632\/v2\">arming manuscript<\/a>.<\/p>\n<p class=\"mt-3 -mb-1 text-[1.125rem] font-bold\" dir=\"ltr\">In rounds 1 and 2 the use of a LLM can be ruled out. There are numerous errors, of kinds that instruction-tuned models essentially never emit:<\/p>\n<ul class=\"[li_&amp;]:mb-0 [li_&amp;]:mt-1 [li_&amp;]:gap-1 [&amp;:not(:last-child)_ul]:pb-1 [&amp;:not(:last-child)_ol]:pb-1 list-disc flex flex-col gap-1 pl-8 mb-3 print:block print:space-y-1\" dir=\"ltr\">\n<li class=\"font-claude-response-body whitespace-normal break-words pl-2\">Homophone substitution: &#8220;Where they really excluded?&#8221; for &#8220;Were&#8221;<\/li>\n<li class=\"font-claude-response-body whitespace-normal break-words pl-2\">Typos: &#8220;invididual&#8221;, &#8220;persona exposure&#8221;, &#8220;residental&#8221;, &#8220;labeled&#8221;<\/li>\n<li class=\"font-claude-response-body whitespace-normal break-words pl-2\">Number disagreement: &#8220;many paper have&#8221;, &#8220;these comparison&#8221;, &#8220;criteria &#8230; is not specified&#8221;<\/li>\n<li class=\"font-claude-response-body whitespace-normal break-words pl-2\">Article omission: &#8220;Present paper by-passes&#8221;, &#8220;Author fails&#8221;, &#8220;may be major reason&#8221;, &#8220;some of first studies&#8221;<\/li>\n<li class=\"font-claude-response-body whitespace-normal break-words pl-2\">Infinitive error: &#8220;makes the reader to wonder&#8221;<\/li>\n<\/ul>\n<p>Layered on that is a consistent L1-transfer signature pointing to German or Finnish: comma before an embedded interrogative or conditional (&#8220;interesting to learn, how fast&#8221;, &#8220;may be major reason, why classification&#8221;, &#8220;labeled as stratified, if they were&#8221;), the calque &#8220;Is it so that &#8230;&#8221;, German constituent order.<\/p>\n<p class=\"font-claude-response-body break-words whitespace-normal\" dir=\"ltr\">But by the third round the very same reviewer #2 appears unsettled, plausibly because the earlier objections had not landed. The prose changes at exactly that point. Not one of the 56 sentences he wrote in rounds 1 and 2 contains a semicolon; two of his eight round-3 comments do (Fisher exact p = 0.014). Hardly rocket science, but it is a real discontinuity, and the orthography flips alongside it, from repeated &#8220;labeled&#8221; to &#8220;labelled&#8221;.<\/p>\n<p class=\"font-claude-response-body break-words whitespace-normal\" dir=\"ltr\">The argument drifts too, with a certain amnesia. In round 2 the complaint was that rural-restricted studies had been wrongly labelled <em>stratified<\/em>. In round 3 the complaint is that rural-restricted studies have been wrongly labelled <em>representative<\/em>.<\/p>\n<p class=\"font-claude-response-body break-words whitespace-normal\" dir=\"ltr\">Most striking, a full sentence is carried over from Reviewer 1&#8217;s report of the previous round. Reviewer 1 wrote: &#8220;strong statements about previous literature, conflicts of interest, and possible publication bias. These points may be relevant, but they should be expressed in a more neutral and evidence-based tone.&#8221; In round 3 this reappears, from the anonymous reviewer 2, as: &#8220;Strong statements about previous literature, conflicts of interest, and possible publication bias should be expressed in a more neutral and evidence-based tone.&#8221;<\/p>\n<p dir=\"ltr\">The evidence is thin, but the round-3 comments read less like a reviewer who has re-read the manuscript than like a reviewer whose remaining objections have been assembled and smoothed by a tool.<\/p>\n<p dir=\"ltr\">So arguing we in future paper submission increasingly against a LLM?<\/p>\n<p class=\"font-claude-response-body break-words whitespace-normal\" dir=\"ltr\">A <a href=\"https:\/\/www.frontiersin.org\/news\/2025\/12\/15\/most-peer-reviewers-now-use-ai-and-publishing-policy-must-keep-pace\">Frontiers survey<\/a> of over 1,600 academics\u00a0 found that 53% of peer reviewers have used AI tools in their work.\u00a0 Note also that Springer Nature instructs peer reviewers not to upload manuscripts into generative AI tools, emphasising that manuscripts contain confidential and sensitive information. BMC is Springer Nature. If my reviewer 2 pasted my manuscript into a model, that is a policy breach independent of whether the resulting comments were any good.<\/p>\n<p>&nbsp;<\/p>\n\n<p>&nbsp;<\/p>\n<div class=\"bottom-note\">\n  <span class=\"mod1\">CC-BY-NC Science Surf , accessed 16.08.2026<\/span>\n <\/div>","protected":false},"excerpt":{"rendered":"<p>I will not repeat here what I already wrote about author vs journal interaction where I predicted high speed transactions. What I can share here is an ongoing review process where a reviewer is now using a tool giving me an increasing workload. Related in silico work has been published by Anthropic scientists recently who &hellip; <a href=\"https:\/\/www.wjst.de\/blog\/sciencesurf\/2026\/08\/scientific-publishing-in-the-age-of-interacting-ais\/\" class=\"more-link\">Continue reading <span class=\"screen-reader-text\">Scientific publishing in the age of interacting AIs<\/span> <span class=\"meta-nav\">&rarr;<\/span><\/a><\/p>\n","protected":false},"author":1,"featured_media":0,"comment_status":"closed","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[20,5,9,5129],"tags":[],"class_list":["post-26681","post","type-post","status-publish","format-standard","hentry","category-note-worthy","category-philosophy-of-science","category-computer-software","category-ai-supported"],"_links":{"self":[{"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/posts\/26681","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/comments?post=26681"}],"version-history":[{"count":9,"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/posts\/26681\/revisions"}],"predecessor-version":[{"id":26690,"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/posts\/26681\/revisions\/26690"}],"wp:attachment":[{"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/media?parent=26681"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/categories?post=26681"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.wjst.de\/blog\/wp-json\/wp\/v2\/tags?post=26681"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}