Jump to content
  • Sign Up
×
×
  • Create New...

Recommended Posts

  • Diamond Member

This is the hidden content, please

Cloudflare is taking a stand against AI website scrapers

Cloudflare has released a new free tool that prevents AI companies’ bots from scraping its clients’ websites for content to train large language models. The cloud service provider is making this tool available to its entire customer base, including those on free plans. “This feature will automatically be updated over time as we see new fingerprints of offending bots we identify as widely scraping the web for model training,” the company said.

In announcing this update, Cloudflare’s team also shared some data about how its clients are responding to the ***** of bots that scrape content to train generative AI models. According to the company’s internal data, 85.2 percent of customers have chosen to block even the AI bots that properly identify themselves from accessing their sites.

Cloudflare also identified the most active bots from the past year. The Bytedance-owned Bytespider **** attempted to access 40 percent of websites under Cloudflare’s purview, and tried on 35 percent. They were half of the top four AI **** crawlers by number of requests on Cloudflare’s network, along with Amazonbot and ClaudeBot.

It’s proving very difficult to fully and consistently block AI bots from accessing content. The arms race to build models faster has led to instances of companies skirting or outright breaking the existing rules around blocking scrapers. of scraping websites without the required permissions. But having a backend company at the scale of Cloudflare getting serious about trying to put the kibosh on this behavior could lead to some results.

“We ***** that some AI companies intent on circumventing rules to access content will persistently adapt to evade **** detection,” the company said. “We will continue to keep watch and add more **** blocks to our AI Scrapers and Crawlers rule and evolve our machine learning models to help keep the Internet a place where content creators can thrive and keep full control over which models their content is used to train or run inference on.”



This is the hidden content, please

news, Cloudflare, large language model, tomorrow, AI
#Cloudflare #stand #website #scrapers

This is the hidden content, please

This is the hidden content, please

Join the conversation

You can post now and register later. If you have an account, sign in now to post with your account.

Guest
Unfortunately, your content contains terms that we do not allow. Please edit your content to remove the highlighted words below.
Reply to this topic...

×   Pasted as rich text.   Paste as plain text instead

  Only 75 emoji are allowed.

×   Your link has been automatically embedded.   Display as a link instead

×   Your previous content has been restored.   Clear editor

×   You cannot paste images directly. Upload or insert images from URL.

  • Vote for the server

    To vote for this server you must login.

    Jim Carrey Flirting GIF

  • Recently Browsing   0 members

    • No registered users viewing this page.

Important Information

Privacy Notice: We utilize cookies to optimize your browsing experience and analyze website traffic. By consenting, you acknowledge and agree to our Cookie Policy, ensuring your privacy preferences are respected.