# Oxford publishes InfoOps AI safety benchmark for synthetic content

> The University of Oxford has released InfoOps, which it describes as the first live-updated safety benchmark for information operations. Rather than asking whether a model is factually accurate, InfoOps probes something more adversarial: will a system generate fake personas, propaganda, and disinformation designed to power influence campaigns, or does it refuse? The timing is pointed. InfoOps launched on July 31, just days before the EU began enforcing its transparency rules on synthetic content on August 2. The benchmark is designed to update continuously, so it can keep testing models as they evolve rather than freezing a snapshot in time, a direct response to the way static evaluations go stale almost as soon as they are published. It fills a real gap. Most frontier safety testing has focused on things like biohazard or cyber capabilities, leaving the mass production of persuasive fake content comparatively under-measured. By turning that risk into a moving, public benchmark, Oxford gives regulators and labs a shared yardstick for one of the most politically charged failure modes AI has, just as the law starts demanding accountability for it.

_Section: [Daily AI Updates](https://www.wortins.com/daily-ai) · Source: TechTimes · Published Tuesday, August 4, 2026_

## Wortins' read

The University of Oxford has released InfoOps, which it describes as the first live-updated safety benchmark for information operations. Rather than asking whether a model is factually accurate, InfoOps probes something more adversarial: will a system generate fake personas, propaganda, and disinformation designed to power influence campaigns, or does it refuse? The timing is pointed. InfoOps launched on July 31, just days before the EU began enforcing its transparency rules on synthetic content on August 2. The benchmark is designed to update continuously, so it can keep testing models as they evolve rather than freezing a snapshot in time, a direct response to the way static evaluations go stale almost as soon as they are published. It fills a real gap. Most frontier safety testing has focused on things like biohazard or cyber capabilities, leaving the mass production of persuasive fake content comparatively under-measured. By turning that risk into a moving, public benchmark, Oxford gives regulators and labs a shared yardstick for one of the most politically charged failure modes AI has, just as the law starts demanding accountability for it.

## Source

[Read the full story at TechTimes](https://www.techtimes.com/articles/322562/20260731/oxford-publishes-first-live-ai-safety-benchmark-information-operations-risk.htm)

## Related coverage

- [Nvidia and Palantir team up to run supply chains with AI, starting with Nvidia's own million-part operation](https://www.wortins.com/story/nvidia-and-palantir-team-up-to-run-supply-chains-with-ai-sta-4206093d) — [The Decoder](https://the-decoder.com/nvidia-and-palantir-team-up-to-run-supply-chains-with-ai-starting-with-nvidias-own-million-part-operation/)
- [Trump officials say AI will help save rural health care. Some leaders in the field don’t believe it](https://www.wortins.com/story/trump-officials-say-ai-will-help-save-rural-health-care-some-bf0e7d83) — [STAT](https://www.statnews.com/2026/09/10/rural-health-care-ai-adoption-challenges-part-4-unraveled-series/?utm_campaign=rss)
- [Two years ago, Meta killed CrowdTangle. Can a new AI tool fill the void?](https://www.wortins.com/story/two-years-ago-meta-killed-crowdtangle-can-a-new-ai-tool-fill-5e67f291) — [Nieman Lab](https://www.niemanlab.org/2026/09/two-years-ago-meta-killed-crowdtangle-can-a-new-ai-tool-fill-the-void/)
- [OpenAI launches ChatGPT Images 2.5 with faster editing](https://www.wortins.com/story/openai-launches-chatgpt-images-2-5-with-faster-editing-07e86379) — [TestingCatalog](https://www.testingcatalog.com/openai-launches-chatgpt-images-2-5-with-faster-editing/)
- [Paris-based Arlequin AI, which is developing proprietary models based on topological neural networks, raised a €28M Series A co-led by redalpine and OTB (Tamara Djurickovic/Tech.eu)](https://www.wortins.com/story/paris-based-arlequin-ai-which-is-developing-proprietary-mode-48952d84) — [Techmeme](https://www.techmeme.com/260910/p15#a260910p15)
- [Letter: the Senate disaster management subcommittee, led by Sen. Josh Hawley, is probing OpenAI's handling of the Hugging Face breach, calling it "reckless" (Axios)](https://www.wortins.com/story/letter-the-senate-disaster-management-subcommittee-led-by-se-da54156c) — [Techmeme](https://www.techmeme.com/260910/p11#a260910p11)

---

_Curated and written by [Wortins](https://www.wortins.com) — The daily AI briefing. Every story links to its original source; the "Wortins read" on each is our own original analysis. [About Wortins & our editorial approach](https://www.wortins.com/about)._
