Now you can block OpenAI’s web crawler

Image: OpenAI

OpenAI now lets you block its web crawler from scraping your site to help train GPT models.

In a blog post, OpenAI said website operators can specifically disallow its GPTBot crawler on their site’s Robots.txt file or block its IP address. “Web pages crawled with the GPTBot user agent may potentially be used to improve future models and are filtered to remove sources that require paywall access, are known to gather personally identifiable information (PII), or have text that violates our policies,” OpenAI said in the blog post. For sources that don’t fit the excluded criteria, “allowing GPTBot to access your site can help AI models become more accurate and improve their general capabilities and safety.”

Blocking the GPTBot may be the…

Now you can block OpenAI’s web crawler

By

By

Related Post

Why am I internet-stalking the pope?

Strategic Importance of Tesla Dojo Chips and 4680 Batteries

Apple has a new ‘Viral’ playlist on Apple Music and Shazam

You missed

Why am I internet-stalking the pope?

Strategic Importance of Tesla Dojo Chips and 4680 Batteries

Apple has a new ‘Viral’ playlist on Apple Music and Shazam

Celsius founder Alex Mashinsky sentenced to 12 years in prison

ModernAftertime