Sam Clemente

u/countablenewt@allthingstech.social
0 posts · 36 comments

Recent posts

No posts.

Recent comments

@jon@vivaldi.net While I do understand the sentiment, I think the problem with search has been more the ads than anything else

Things like the AI overviews are bad, but search has been more or less powered by AI for a while, it's really just been this new bout of generative AI where they're really saying it out loud

But even without that GenAI, search results have been getting significantly worse for a bit now

@BackOnMyBS It's not entirely straightforward, but I find that stating "I'm confused because you said 'x' which I interpreted as 'y' and this thing you're saying now I interpret as being 'z' doesn't seem consistent, what am I not understanding" (doesn't have to be those exact words)

Or if it's something I said I apologize for the confusion, state my intentions behind what I said, and try to say it in a different way based off how they interpreted the last thing I said

@thestereobus @Cock_Inspecting_Asexual AirPods Pro 2 really work for me personally

If you wanted to stick with the AirPods route the Maxes will do a better job of noise isolation due to size, but they’re bulky and super expensive

Though there are rumors going around that Apple will be updating the AirPods Max today at their release event so definitely keep a lookout for those if you’re interested

@zbyte64 where am I wrong? The process is effectively the same: you get a set of training data (a textbook) and a set of validation data (a test) and voila, I’m trained

To learn how to draw an image of a thing, you look at the thing a lot (training data) and try sketching it out (validation data) until it’s right

How the data is acquired is irrelevant, I can pirate the textbook or trespass to find a particular flower, that doesn’t mean I’m learning differently than someone who paid for it

@zbyte64 data quality, again, was out of the scope of what I was talking about originally

Which, again, was that legal precedent would suggest that the *how* is largely irrelevant in copyright cases, they’re mostly focused on *why* and the *scale of the operation*

I’m not getting sued for copyright infringement by the NYT because I used inspect element to delete content to read behind their paywall, OpenAI is

@zbyte64 from what I understand, you’re referring to the process at scale—the amount of information the AI can take in is inhuman—which I’m not disagreeing with

None of which is relevant to my original point: the scale of their operations, which has already been used countless times in copyright law

The scale at which they operate and their intention to profit is the basis for their infringement, how they’re doing it would be largely irrelevant in a copyright case, is my point

@Subverb that is, quite impressively, the opposite of what I said

Is a person infringing on copyright by producing content? No. It’s about intent and scale. Humans don’t just sit on this knowledge, they do something with it

There is nothing illegal about WHAT it’s doing, there is everything illegal about HOW and WHY

I very clearly stated that OpenAI’s intent and their scale at which they operate are blatant copyright infringement and that it has been backed up with decades of precedents

@Pika @flop_leash_973 This is largely my thoughts on the whole thing, the process of actually training the AI is no different from a human learning

The thing about that, is that there's likely enough precedent in copyright law to actually handle that, with most copyright law it's all about intent and scale and I think that's likely where this will all go

Here the intent is to replace and the scale is astronomical, whereas an individual's intent is to add and the scale is minimal