this post was submitted on 22 Jul 2026
110 points (100.0% liked)

Technology

43316 readers
349 users here now

A nice place to discuss rumors, happenings, innovations, and challenges in the technology sphere. We also welcome discussions on the intersections of technology and society. If it’s technological news or discussion of technology, it probably belongs here.

Remember the overriding ethos on Beehaw: Be(e) Nice. Each user you encounter here is a person, and should be treated with kindness (even if they’re wrong, or use a Linux distro you don’t like). Personal attacks will not be tolerated.

Subcommunities on Beehaw:


This community's icon was made by Aaron Schneider, under the CC-BY-NC-SA 4.0 license.

founded 4 years ago
MODERATORS
 

For decades, publishers have done everything in their power, from the legal to the not-explicitly illegal, to rank as highly in Google Search as possible. For many websites, traffic from the search engine was their single greatest source of audience and, as a result, revenue.

Now though, a handful of influential players in the digital media ecosystem have begun moving in the opposite direction, laying the groundwork for what was once unthinkable: removing themselves from Google Search.

Beginning Sept. 15, all new websites signing up for Cloudflare, as well as all the customers on its free tier, will have the default settings in their bot management protocol set to block “multi-purpose crawlers” on any webpage that has ads. This means that any crawler that scrapes for both search indexing and AI training will be turned away at the door, unless the site owner decides otherwise.

“We’ve been clear about what we want,” said Cloudflare chief strategy officer Stephanie Cohen. “We want a technical solution that allows you to be discoverable without having to give your content away for free.”

top 17 comments
sorted by: hot top controversial new old
[–] XLE@piefed.social 25 points 1 week ago

Google: Muahaha, we train our AI with the data from the publishers’ sites while indexing them. Then we summarize their site so the user never leaves our site and all the ad revenue is ours, brilliant!

Publishers: People aren’t coming to our site from Google search because it just summarizes our site for them. So we lose no traffic by completely blocking Google from crawling us.

Google: Surprised Pikachu face

Source: @leadore@lemmy.world (https://lemmy.world/comment/24710324)

[–] FaceDeer@fedia.io 7 points 1 week ago (1 children)

“We’ve been clear about what we want,” said Cloudflare chief strategy officer Stephanie Cohen. “We want a technical solution that allows you to be discoverable without having to give your content away for free.”

Sounds like they want DRM,

[–] XLE@piefed.social 18 points 1 week ago* (last edited 1 week ago) (1 children)

Sounds like people don't want AI scrapers like Google taking from their websites without their consent, which allows Google to further centralize and corporatize the web.

[–] FaceDeer@fedia.io 4 points 1 week ago (1 children)

And they may also want a magical unicorn for a pet. Simply wanting a thing doesn't conjure it into existence.

They want the same thing that people who want DRM want - a way to show their content to a viewer without that viewer being able to copy what they're seeing. It's just as impossible either way. As with DRM there may be tricks one can come up with to hinder it temporarily under some circumstances, but fundamentally data is data. If someone can see it then someone can copy it.

[–] XLE@piefed.social 9 points 1 week ago (1 children)

The companies doing the unethical scraping are the megatech companies: Google, OpenAI, Anthropic. They can absolutely be reigned in.

Megacorps invented DRM to try to keep their profits from the public. Now megacorps are trying to scrape up everybody else's data and close off the public. That's literally the goal of Google's zero-click AI.

[–] FaceDeer@fedia.io -3 points 1 week ago (2 children)

Also Moonshot, Alibaba, Baidu. Can they also be reined in?

It doesn't matter who invented DRM. My point is that DRM fundamentally doesn't work. If you pubish a web page in a manner that allows humans to read it then it's also readable by AIs. A pinkie-promise not to is the best you can get.

[–] XLE@piefed.social 10 points 1 week ago (3 children)

This post is about Google, so that's why I'm focusing on Google. This is a positive step forward against one massive, unethical, evil AI company.

Defeatism just benefits them. We don't need it.

[–] minfapper@piefed.social 1 points 1 week ago

Eh, it's less about defeatism and more that I'm watching a fight between two assholes so I'm finding it hard to empathize with either side.

Like... Publishers regularly screwed over consumers, content creators, and tried to break the web multiple times. They can bleed for all I care

[–] The_Decryptor@aussie.zone 1 points 1 week ago

Defeatism just benefits them. We don’t need it.

It's not defeatism btw, it's opposition to the idea that people should have a say if their content is used to train an LLM.

[–] FaceDeer@fedia.io -2 points 1 week ago (1 children)

Guess we'll see, then. I suspect the result is going to be a bunch of websites going "huh? Where did all our traffic suddenly disappear to?"

[–] XLE@piefed.social 4 points 1 week ago (1 children)

Google's zero-click AI search is already harming small publishers! I'm glad we have some common ground on opposing that.

Pushing for regulation against providing AI results would be a great start, which would make the step in this article unnecessary.

[–] FaceDeer@fedia.io 1 points 1 week ago (1 children)

Trying to legislate the world to go back to the way it was in 2023 probably isn't going to work any better than DRM does.

[–] XLE@piefed.social 4 points 1 week ago (2 children)

You can, in fact, legislate to enact a better status quo. America has a history of civil rights laws that demonstrate as much.

You don't have to "legislate the world" to make some positive change.

[–] TehPers@beehaw.org 1 points 1 week ago

While I agree with your points and sentiment, the US civil rights laws (including the Civil Rights Act) are being continuously torn apart by billionaires and religious lunatics, so maybe not the best example.

Regardless, there are plenty of examples of laws succeeding in bettering the status quo going even as far back as Babylon.

[–] SocialistVibes01@lemmy.ml 0 points 1 week ago* (last edited 1 week ago)

Ah, one of those believe in the system guys. The US history on civil rights is one of horror and oppression.

[–] InevitableList@beehaw.org 2 points 1 week ago

That's like saying copyright doesn't work. If a human can see it a human can copy it

[–] Zaleramancer@beehaw.org 3 points 1 week ago

Another contributing factor to the loss of records that future internet historians will have to deal with.