Jesse Wynants’ Post

Question for PRs: Do you want AI to train on your press releases? Right now, our newsrooms are crawlable by AI agents seeking to answer human queries – so, when someone asks ChatGPT "what's the best mic for studio recording", it's possible for the AI to then go and find your story as a source. Prezly does not block those. What we DO block are requests from models looking to train on your data. (That said, there's only so much we can do here – our Cloudflare data suggests AI agents don't always respect the robots.txt "Do Not Enter" sign. If you're curious, I can dig into those numbers and share them in a follow-up post.) The reason we block training AI requests is largely due to consent. We host content for thousands of brands. We can't assume everyone is morally or legally comfortable with tech giants hoovering up their IP to train proprietary models. There's also the issue of bandwidth. Unregulated scraping bots aggressively hammer servers, consuming massive bandwidth and computing power, which can slow down newsroom load times for actual humans. BUT... blocking training data isn't as simple as "good v bad". There are potential benefits to letting AI train on newsroom data: 1. Better brand representation. If an AI model is trained on your client's official, factual press releases, it’s far less likely to hallucinate false information 2. Long-term discoverability. As search engines evolve into conversational AI tools, being part of the foundational training data ensures your client stays in the AI's memory cache, rather than relying on a live web search 3. Industry authority. Press releases are structured, high-quality, verified factual data, which is the exact type of clean data that AI companies need. Having your content in those training sets can position your client as a category authority So, PR community, I want to hear from you. Do you view AI training as unauthorized IP theft, or as the ultimate form of media distribution? If Prezly gave you a simple toggle switch in your settings to allow or block AI training scrapers on your newsrooms, would you turn it on? Let me know in the comments.

  • Text on a gradient background reads: "Question for PRs: Do you want AI models to train on your press releases?"

I think clients want that, because they want to appear in AI-based searches - so yes.

To view or add a comment, sign in

Explore content categories