this post was submitted on 15 Mar 2024
491 points (95.4% liked)

Technology

59343 readers
5584 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related content.
  3. Be excellent to each another!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, to ask if your bot can be added please contact us.
  9. Check for duplicates before posting, duplicates may be removed

Approved Bots


founded 1 year ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
[–] A_Very_Big_Fan 3 points 8 months ago* (last edited 8 months ago) (1 children)

That article doesn't even claim it's distributing copyrighted material.

If that qualifies as distributing stolen copyrighted material, then this is stealing and distributing the "you shall not pass" LoTR scene. Which, again, ChatGPT won't even do

[–] [email protected] -1 points 8 months ago (1 children)

Sorry, I know reading the whole article is hard:

The complaint cites several examples when a chatbot provided users with near-verbatim excerpts from Times articles that would otherwise require a paid subscription to view.

[–] A_Very_Big_Fan 2 points 8 months ago* (last edited 8 months ago)

Yeah lmao after like 20 paragraphs of nothing, it wasn't hard to believe you didn't know what you were talking about. But I looked at the complaint itself out of curiosity, and it's flimsy and misleading.

The first issue is 100% of the allegedly paywalled text from all 4 articles mentioned in the complaint can be read by non-paying customers for free outside of the paywall. You can't read the whole article, but you can get far enough to read all 4 quotes mentioned in the complaint yourself. The links to each article are in the complaint if you don't believe me. They have nothing to show they bypassed a paywall or that it was trained on unlicensed content.

The second issue is the third exhibit claims it will bypass paywalls when asked. This is demonstrably false because for one, the article they asked it for isn't paywalled, and for two, using their exact prompts word for word doesn't work if you try it yourself.

Two of the four exhibits don't even have screenshots, so there's no evidence it happened in the first place, but more importantly they don't (and apparently won't when asked) disclose what lengths they had to go to in order to get that output. For all we know they gave it 90% of the words and told it to fill in the gaps.