▲ 526 ▼ At least 3 major outlets — The New York Times, The Guardian, and Reddit — have blocked the Internet Archive’s Wayback Machine from accessing their content (www.mediapost.com) submitted 5 months ago by Innerworld@lemmy.world to c/news@lemmy.world 64 comments fedilink hide all child comments
[–] Saryn@lemmy.world 74 points 5 months ago (3 children) Content scraping is harming the information business in ways that could not have been foreseen. What an absolute ridiculous thing to say. permalink fedilink source hideshow 6 child comments replies: [+] ameancow@lemmy.world 18 points 5 months ago* (last edited 5 months ago) [deleted] permalink fedilink source parent [–] Cantaloupe@lemmy.fedioasis.cc 13 points 5 months ago Gee, I wonder what else could be scraping content from all these websites? permalink fedilink source parent [–] REDACTED@infosec.pub 12 points 5 months ago (1 child) To be fair, the archive indeed got heavily abused into simply reading without paywalls. I know this is a controversial opinion, but seeing comments on other threads like "Remember to support news media", then "use archive to bypass paywalls" then anger towards said companies for caring about getting paid or growing, makes one question where exactly does Lemmy draw the line between pirating and paid content. Or are we simply altogether against sites like 404Media just because of paywalls? permalink fedilink source parent hideshow 2 child comments replies: [–] Saryn@lemmy.world 13 points 5 months ago (1 child) That's not the point. The point is content scraping (and crawling) is the cornerstone of the contemporary information environment. It's how we got to this technological paradigm in the first place. This whole "people are bypassing paywalls" is a badly evidenced non-issue, and all too convenient. What these companies are really saying is "Content scraping is bad when others do it. Only I and other big fish get to do it and profit billions out of it. Fuck ordinary citizens. Fuck everyone and everything but me and my dreams of endless wealth and power." To be fair. permalink fedilink source parent hideshow 2 child comments replies: [–] gagcar@lemmus.org 2 points 5 months ago You say bypassing paywalls is a non-issue, but it is basically the only thing I have heard people say to use it for on social media. You can have your problems about data harvesting, but don’t pretend like getting around paywalls was not what the average individual user was using it for. permalink fedilink source parent
[+] ameancow@lemmy.world 18 points 5 months ago* (last edited 5 months ago) [deleted] permalink fedilink source parent
[–] Cantaloupe@lemmy.fedioasis.cc 13 points 5 months ago Gee, I wonder what else could be scraping content from all these websites? permalink fedilink source parent
[–] REDACTED@infosec.pub 12 points 5 months ago (1 child) To be fair, the archive indeed got heavily abused into simply reading without paywalls. I know this is a controversial opinion, but seeing comments on other threads like "Remember to support news media", then "use archive to bypass paywalls" then anger towards said companies for caring about getting paid or growing, makes one question where exactly does Lemmy draw the line between pirating and paid content. Or are we simply altogether against sites like 404Media just because of paywalls? permalink fedilink source parent hideshow 2 child comments replies: [–] Saryn@lemmy.world 13 points 5 months ago (1 child) That's not the point. The point is content scraping (and crawling) is the cornerstone of the contemporary information environment. It's how we got to this technological paradigm in the first place. This whole "people are bypassing paywalls" is a badly evidenced non-issue, and all too convenient. What these companies are really saying is "Content scraping is bad when others do it. Only I and other big fish get to do it and profit billions out of it. Fuck ordinary citizens. Fuck everyone and everything but me and my dreams of endless wealth and power." To be fair. permalink fedilink source parent hideshow 2 child comments replies: [–] gagcar@lemmus.org 2 points 5 months ago You say bypassing paywalls is a non-issue, but it is basically the only thing I have heard people say to use it for on social media. You can have your problems about data harvesting, but don’t pretend like getting around paywalls was not what the average individual user was using it for. permalink fedilink source parent
[–] Saryn@lemmy.world 13 points 5 months ago (1 child) That's not the point. The point is content scraping (and crawling) is the cornerstone of the contemporary information environment. It's how we got to this technological paradigm in the first place. This whole "people are bypassing paywalls" is a badly evidenced non-issue, and all too convenient. What these companies are really saying is "Content scraping is bad when others do it. Only I and other big fish get to do it and profit billions out of it. Fuck ordinary citizens. Fuck everyone and everything but me and my dreams of endless wealth and power." To be fair. permalink fedilink source parent hideshow 2 child comments replies: [–] gagcar@lemmus.org 2 points 5 months ago You say bypassing paywalls is a non-issue, but it is basically the only thing I have heard people say to use it for on social media. You can have your problems about data harvesting, but don’t pretend like getting around paywalls was not what the average individual user was using it for. permalink fedilink source parent
[–] gagcar@lemmus.org 2 points 5 months ago You say bypassing paywalls is a non-issue, but it is basically the only thing I have heard people say to use it for on social media. You can have your problems about data harvesting, but don’t pretend like getting around paywalls was not what the average individual user was using it for. permalink fedilink source parent