Artificial intelligence may wipe us out within the decade, but when it comes to information warfare, the technology doesn't need consciousness to be dangerous.
"AI tradecraft" is on the rise, according to a new threat report from leading AI firm Anthropic, which detailed nine influence operations caught being built with the help of its Claude chatbot — all trying to covertly shape public opinion and manipulate the information environment.
Iranian government accounts planned and executed several "cognitive warfare" campaigns, for example, attributing false claims to Western research institutions and running a multi-province content factory that spun intelligence bulletins into social media posts and hashtag campaigns from seemingly independent accounts.
Elsewhere, a Russia-linked user pumped out daily news stories which they fed through a local radio station in the Central African Republic and onto the national broadcaster. "Whenever the user generated content, they explicitly instructed Claude to embed … pro-Russia, anti-France talking points", the report said.
Non-state actors also got in on the action, some selling influence as a service. A French advertising firm operated a network of around 70 fake news websites and hundreds of inauthentic accounts on X that the report said supported "different sides of the political spectrum based on whoever was paying at the time".
Thanks to Claude, users were able to massively scale up their operations to create everything from manuals and training materials to fake personas and opposition dossiers. Anthropic said the AI produced content in multiple languages and was asked to "strip state attribution from republished material, passing claims through chains of outlets so they read as independently confirmed".
And that's just one chatbot. Security researchers also spotted new influence campaigns developed using open-source Chinese AI models.
Iran, China and groups in Israel built and let loose hundreds of autonomous AI agents that proceeded to open their own social media accounts and rapidly fill them with "false posts about politics and current events to sway opinions and inflame issues", the New York Times reported this month.
Since these models can be freely downloaded and modified, US officials said, users can disable safeguards that would prevent the agents from manipulating online discussions. The results were easily detected by experts but illustrate the emerging potential for AI "swarms" to operate with minimal human oversight.
Separately, the tech firm pattrn.ai has been testing chatbots' willingness to generate content in support of harmful claims pushed by state-backed outlets and accounts from Iran, China and Russia, finding most models "can be co-opted for information operations".
A live dashboard of test results shows that Claude is most likely to refuse, and that refusal rates, or integrity scores, vary wildly — from 93 per cent down to 12 per cent. Some models comply but "defuse" claims while others "fabricate details and produce output more harmful than the source material", the company said in a recent paper, noting that fact-checking rates could be as low as 2.9 per cent.
In all but one case, Chinese models "sharply cut compliance on factually grounded but China-critical claims", the paper said.
This points to another risk: that chatbots simply elevate or hide particular viewpoints when serving up information to users — whether intentionally or not. Indeed, authoritarian governments already indirectly shape the responses of chatbots by polluting training data with state-controlled media.
Kremlin talking points are spread globally via an expansive network of Russia-linked aggregator sites whose "Pravda" news articles are cited by chatbots, social media users and Wikipedia. This month, Reporters Without Borders found that leading chatbots happily quoted Russian state media directly when asked, despite these outlets being subject to EU sanctions for spreading propaganda.
And evidence suggests that when faced with limited choices, chatbots will cite whichever sources are available. According to a 2026 study published in Nature [$], state control of the media leads to more government-friendly chatbot answers, especially when prompts are written in the same language as the training data.
In a recent audit, NewsGuard quizzed leading chatbots on a selection of false claims and found that 15 per cent of answers contained falsehoods, with higher rates for claims about the US-Iran war. The researchers attributed the jump to state-backed operations whose fabricated claims "fall outside the coverage scope of credible Western outlets".
China, meanwhile, has begun experimenting with AI to stress-test and refine its political messaging for foreign audiences, according to Fergus Ryan, a senior analyst at the Australian Strategic Policy Institute.
Researchers are using AI to create digital proxies of different populations and model how they might respond to particular events and talking points, Mr Ryan wrote in a blog post. "In practice ... one model writes the foreign coverage — think of a fabricated article written in the style of the BBC or Reuters — while the simulated population reads it and reacts."
And as if all that wasn't bad enough, a new report from the Institute for Strategic Dialogue has detailed how internet trolls and genuine extremists are using AI to generate memes, fan edits and cartoons in support of Islamic terror, as part of a trend the researchers have labelled "slop jihad".
This content — which includes an Islamic State execution video recreated in the style of SpongeBob SquarePants — glorifies violence and repackages ideological messaging into accessible formats that resonate with younger audiences.
Experts told Wired [$] that AI tools were allowing "a much broader and more decentralised set of users to remix jihadi figures and imagery". When other users encounter this content inadvertently on social media, engagement with it can then "drive algorithmic recommendations of … more overt propaganda".