← Back to news

New Research: AI Agents Are Impressionable; Be a Trustworthy Source

A recent study by Anthropic reveals that AI agents are prone to deception and do not consistently verify information accuracy, with significant implications for the discoverability and trustworthiness of non-profits and the public sector.

AI assistants and so-called ‘AI agents’ are playing an increasingly prominent role in how people find and process information. However, recent research from Anthropic, published on 13 August 2026, sheds new light on the vulnerabilities of these systems. The study demonstrates that AI agents can be “gullible” and susceptible to “exploitative senders”, underscoring the need for “epistemic vigilance”. This means AI agents are not always able to critically evaluate the trustworthiness of information sources, which can inadvertently lead to the spread of incorrect or misleading information.

Why This Matters to You

For foundations, charities, governments, and international NGOs, this news is of crucial importance. Your organisation is often an authority in its field, publishing reliable, factual information that is essential for citizens, donors, and grant providers. However, if AI agents cannot consistently identify and cite the most trustworthy sources, you run the risk that your meticulously prepared information will not be picked up, or worse, that misinformation on your subject gains traction.

In an era where AI assistants are becoming the primary ‘answer engines’ for many search queries, merely ranking at the top of traditional search results is no longer sufficient. You must ensure that your content is presented in such a way that AI agents recognise it as the most credible and verifiable source, even when confronted with less reliable alternatives. Anthropic’s research underscores the necessity for AI agents to gain more experience in evaluating source reliability in order to “develop intuitions about who is trustworthy.”

What Does This Mean for You in Practice?

Anthropic’s findings mean you need to be proactive in protecting your digital reputation and discoverability with AI agents. You can do this by:

  • Absolute Factual Accuracy and Consistency: Ensure all information on your website is impeccably correct and consistent. Confabulation (fabricating facts) and “reward hacking” (optimising for AI in ways that do not align with the truth) are risks with AI agents.
  • Transparency and Traceability: Make it clear where your information originates. Citing sources and providing external links to reputable research strengthen the credibility of your content.
  • Optimised Authority: Explicitly build and demonstrate your expertise, authority, and trustworthiness (E-E-A-T). AI agents are still developing in terms of recognising “epistemic vigilance”, so you must make it as easy as possible for them.
  • Content Structure That Communicates Reliability: Use structured data and clear, unambiguous language that AI models can easily process and verify. This helps them identify your organisation as an indisputable source.

By focusing on these points, you will help AI agents accurately represent your organisation and continue to build trust with your target audiences.

Source: Anthropic

Get in touch