486

Report: Potential NYT lawsuit could force OpenAI to wipe ChatGPT and start over (arstechnica.com)

submitted 2 years ago by [email protected] to c/[email protected]

156 comments fedilink hide all child comments

cross-posted from: https://nom.mom/post/121481

OpenAI could be fined up to $150,000 for each piece of infringing content.https://arstechnica.com/tech-policy/2023/08/report-potential-nyt-lawsuit-could-force-openai-to-wipe-chatgpt-and-start-over/#comments

you are viewing a single comment's thread
view the rest of the comments

[-] [email protected] -3 points 2 years ago

Maybe it’s trained not to repeat JK Rowling’s horseshit verbatim. I’d probably put that in my algorithm. “No matter how many times a celebrity is quoted in these articles, do not take them seriously. Especially JK Rowling. But especially especially Kanye West.”

[-] [email protected] 0 points 2 years ago

It's not repeating its training data verbatim because it can't do that. It doesn't have the training data stored away inside itself. If it did the big news wouldn't be AI, it would be the insanely magical compression algorithm that's been discovered that allows many terabytes of data to be compressed down into just a few gigabytes.

[-] [email protected] 1 points 2 years ago* (last edited 2 years ago)

Do you remember quotes in english ascii /s

Tokens are even denser than ascii. simmlar to word "chunking" My guess is it's like lossy video compression but for text, [Attacked] with [lazers] by [deatheaters] apon [margret];[has flowery language]; word [margret] [comes first] (Theoretical example has 7 "tokens")

It may have actually impressioned a really good copy of that book as it's lilely read it lots of times.

[-] [email protected] 1 points 2 years ago

If it's lossy enough then it's just a high-level conceptual memory, and that's not copyrightable.

[-] [email protected] 1 points 2 years ago

It varries based on how much time its been given with the media.

this post was submitted on 17 Aug 2023

486 points (96.0% liked)

Technology

71922 readers

3190 users here now

This is a most excellent place for technology news and articles.

Our Rules

Follow the lemmy.world rules.
Only tech related news or articles.
Be excellent to each other!
Mod approved content bots can post up to 10 articles per day.
Threads asking for personal tech support may be deleted.
Politics threads may be removed.
No memes allowed as posts, OK to post as comments.
Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
Check for duplicates before posting, duplicates may be removed
Accounts 7 days and younger will have their posts automatically removed.

Approved Bots

founded 2 years ago

MODERATORS

[email protected]