OpenAI Furious DeepSeek Might Have Stolen All the Data OpenAI Stole From Us

ForgottenFlux@lemmy.world · 2 days ago

OpenAI Furious DeepSeek Might Have Stolen All the Data OpenAI Stole From Us

TipRing@lemmy.world · 2 days ago

No honor among thieves.

AwesomeLowlander@sh.itjust.works · 1 day ago

There’s plenty of honor in Deepseek releasing open source.

Doomsider@lemmy.world · 8 hours ago

The new innovate and the old litigate.

sunzu2@thebrainbin.org · 2 days ago

Yas 🐸

Big mad

fallowseed@lemmy.world · 1 day ago

everyone concerned about their privacy going to china-- look at how easy it is to get it from the hands of our overlord spymasters who’ve already snatched it from us.

Sgt_choke_n_stroke@lemmy.world · 2 days ago

Sho@lemmy.world · 2 days ago

The battle of the plagiarism machines has begun

dogslayeggs@lemmy.world · 2 days ago

Regardless of how OpenAI procured their data, I’m absolutely shocked that a company from China would obtain data unauthorized from a company in another country.

chingadera@lemmy.world · 2 days ago

chingadera@lemmy.world · 1 day ago

It just gets better and better y’all.

https://www.theregister.com/2025/01/30/deepseek_database_left_open/

Quokka@mastodon.au · 6 hours ago

@whostosay I know they’re being touted as having done very much with very little, but this kind of thing should have been part of the little.

chingadera@lemmy.world · 6 hours ago

I’m not understanding your reply, do you mind rephrasing?

riot@slrpnk.net · 1 day ago

Security? We don’t need no security!

chingadera@lemmy.world · 1 day ago

You get a free database, and you get free database, and you get a free database! EVERYBODY GETS A FREE DATABASE

Oprahbees.gif

SinningStromgald@lemmy.world · 2 days ago

Fuck you! Pay me for my data asshole!

sunzu2@thebrainbin.org · 2 days ago

Shit posters and linux forums are the back bone of these “AI” after you account for all the commons that parasite took and try to lock up.

dnzm@feddit.nl · 2 days ago

Rooty@lemmy.world · 1 day ago

I love how die hard free market defenders turn into fuming protectionists the second their hegemony is threatened.

CitizenKong@lemmy.world · 1 day ago

Tale as old as capitalism.

owenfromcanada@lemmy.world · 2 days ago

just_another_person@lemmy.world · 2 days ago

👏👏👏👏👏

x00z@lemmy.world · 2 days ago

Tamaleeeeeeeeesssssss

hot hot hot hot tamaleeeeeeeees

MysticKetchup@lemmy.world · 2 days ago

I feel like I didn’t appreciate this movie enough when I first watched it but it only gets better as I get older

just_another_person@lemmy.world · 2 days ago

It’s a true comedy that still holds up. I honestly thought for years that Mel Brooks had something to do with it, but he didn’t. It’s so well crafted that there are many layers to it that you can’t even grasp when watching as a child. Seeing it as an adult just open your eyes to how amazingly well done it was.

I could do without the whole Billy Crystalizing of large portions of it though.

owenfromcanada@lemmy.world · 2 days ago

I always thought Rob Reiner had a similar sense of humor to Mel Brooks. And I liked Billy Crystal in it, it kept that section of the movie from feeling too heavy, though I get it’s not everyone’s thing.

For anyone who hasn’t read it, the book is fantastic as well, and helped me appreciate the movie even more (it’s probably one of the best film adaptations of a book ever, IMO). The humor and wit of William Goldman was captured expertly in the movie.

atrielienz@lemmy.world · 1 day ago

I didn’t realize it was a book. Guess I’ll be searching that out.

obviouspornalt@lemmynsfw.com · 2 days ago

Rob Reiner’s dad Carl was best friends with Mel Brooks for almost all of Carl’s adult life.

https://www.vanityfair.com/hollywood/2020/06/carl-reiner-mel-brooks-friendship

ouRKaoS@lemmy.today · 2 days ago

“Now” is always a good time to rewatch it & get more out of it!

A_A@lemmy.world · 2 days ago

ZILtoid1991@lemmy.world · 1 day ago

Intellectual property theft for me but not for thee!

Critical_Thinker@lemm.ee · 1 day ago

It’s a shame that you can’t copyright the output of AI, isn’t it?

ZILtoid1991@lemmy.world · 1 day ago

Trump executive order on the copyrightability of AI output in 3…

vrighter@discuss.tchncs.de · 1 day ago

so? it won’t have any effect on china, because last i checked, us laws apply only in the us

mechoman444@lemmy.world · 1 day ago

I can’t believe we’re still on this nonsense about AI stealing data for training.

I’ve had this argument so many times before y’all need to figure out which data you want free and which data do you want to pay for because you can’t have it both ways.

Either the data is free or it’s paid for. For everyone including individuals and corporations.

You can’t have data be free for some people and be paid for for others it doesn’t work that way we don’t have the infrastructure to support this kind of thing.

For example Wikipedia can’t make its data available for AI training for a price and free for everyone else. You can just go to wikipedia.com and read all the data that you want. It’s available for free there’s no paywall there’s no subscriptions no account to make no password to put in no username to think of.

Either all data is free or it’s all paid for.

Lifter@discuss.tchncs.de · 1 day ago

Many licences have different rules for redistribution, which I think is fair. The site is free to use but it’s not fair to copy all the data and make a competitive site.

Of course wikipedia could make such a license. I don’t think they have though.

How is the lack of infrastructure an argument for allowing something morally incorrect? We can take that argument to absurdum by saying there are more people with guns than there are cops - therefore killing must be morally correct.

mechoman444@lemmy.world · 22 hours ago

The core infrastructure issue is distinguishing between queries made by individuals and those made by programs scraping the internet for AI training data. The answer is that you can’t. The way data is presented online makes such differentiation impossible.

Either all data must be placed behind a paywall, or none of it should be. Selective restriction is impractical. Copyright is not the central issue, as AI models do not claim ownership of the data they train on.

If information is freely accessible to everyone, then by definition, it is free to be viewed, queried, and utilized by any application. The copyrighted material used in AI training is not being stored verbatim—it is being learned.

In the same way, an artist drawing inspiration from Michelangelo or Raphael does not need to compensate their estates. They are not copying the work but rather learning from it and creating something new.

Omega_Jimes@lemmy.ca · 1 day ago

I mean, sure, but the issue is that the rules aren’t being applied on the same level. The data in question isn’t free for you, it’s not free for me, but it’s free for OpenAI. They don’t face any legal consequences, whereas humans in the USA are prosecuted including an average fine per human of $266,000 and an average prison sentence of 25 months.

OpenAI has pirated, violated copyright, and distributed more copyright than an i divided human is reasonably capable of, and faces no consequences.

https://www.splaw.us/blog/2021/02/looking-into-statistics-on-copyright-violations/

https://www.patronus.ai/blog/introducing-copyright-catcher

My use of the term “human” is awkward, but US law considers corporations people, so i tried to differentiate.

I’m in favour of free and open data, but I’m also of the opinion that the rules should apply to everyone.

LengAwaits@lemmy.world · edit-2 1 day ago

I tend to think that information should be free, generally, so I would probably be fine with “OpenAI the non-profit” taking copyrighted data under fair-use, but I don’t extend that thinking to “OpenAI the for-profit company”.