Sounds like they don’t want consent… Sounds like a rapists mentality.
That’s why it MUST be opt-in.
If your product provided any real value to the consumer than people would be lining up to opt in.
and they could easily do that by paying content creators. Say, matching the ad revenue and sponsorship value of each video. But they won’t.
If you come up with a thing you KNOW FOR A FACT nobody sane would opt in to, then don’t fucking do it at all. Cunts.
But $…
Nobody Wants This But Me: Things AI Devs and Rapists Have In Common
So they’re relying on streamers not realizing the option is on, or perhaps not caring that it’s on, to get their training data. Stupid bastards.
On top of that, people have been having to go back to look at the setting to make sure it saves which likely just means they can “oops” it back on for anyone, anytime.
I clicked it, left, went back, and it hadn’t saved. So I left that screen up for awhile and that time seemed to work.
Going to have to check like once a week now though.
Motherfuckers. I hope a compilation of the loudest Cibidoki’s burps gets embedded irreversibly in the dataset.
Currently, if you can access it without an account, it’s never been opt anything. You can freely scrape anything that’s publicly available.
I guess Twitch is just laying down the ground work expecting copyright laws to get stronger (not a good thing unless you are Twitch and YouTube).
I think this misses the point that you used to have to actually try a little to scrape content.
It is now a one to two click operation to send an AI bot to steal everyone’s shit so you can sell it back to them. That’s unacceptable.
There’s a difference between scraping the net for say, internet archive, or downloading a video you’re going to cut into a new work you make yourself, and then using AI trained specifically to steal peoples work to then sell back to them.
Nuance is really important right now.
It was easier to scrape before than it is now, the difficulty shouldn’t impact the legality in any case.
I think the main dissonance comes from people thinking if they broaden copyright laws, somehow that will mean either less AI or artists getting paid.
These companies are data brokers, they own the data when it comes to this, not the creators. It’s why Reddit made millions but not one cent went to any users.
It was perfectly legal to scrape a picture, cut it into pieces and glue it in a different order before AI. I don’t see how taking that picture and making a tool to make pixel level collages is different. Doesn’t really matter who it’s being “sold” to (I use local for photos and video, never bought a thing).
The current court cases are putting up a walled Garden, not protecting the little people. You fight for copyright juggernauts and big AI. Nuance is a thing but so is pragmatism, anyone with half a brain can see open source is the only thing in the crosshair and this whole mess is leading to anti-consumer laws.
If you really think LLMs have not made scraping the internet easier, then I have a bridge to sell you.
We literally cannot have this conversation if you don’t exist in reality.
the difficulty shouldn’t impact the legality in any case
I mean, the anti bot measures got stronger because of them as well. In any case, it doesn’t matter like I said.
It must be fun ignoring all my points and fixating on the first half phrase. We can’t have this conversation because you want to live in a fantasy land where defending big AI, and their transparent manipulation as to build themselves a monopoly, is a good thing.
All hail our copyright overlords and the mega corps that partner with them.
The fantasy land is the one where you think scraping got harder and that twitch is just “laying down copyright groundwork”.
Have you even been paying attention to what these corps have been doing? Groundwork my ass they’re stealing people’s work for their own benefits.
I do not give a single solitary fuck about anymore about rules, polices, laws, etc that are explicitly designed to fuck over the average person and benefits corporate interests, which is basically the entirety of the US judicial, copyright, trademark, and monetary system.
If you think twitch is just laying groundwork, again I have a second bridge to sell you.
It’s part of my work (not to train AI though). Sure scraping is easier now for people that don’t know how to code. That doesn’t make up for the fact that even a Lemmy instance now has bot protection. The ones that used to have simple bot protection back then now have complex bullshit that changes all the time. It’s a lot more headache and maintenance, even with the LLM doing half the work.
The laws that are coming are designed to fuck over the average person by kneecapping open source. You stance is encouraging that.
If you think idioms bring anything to the conversation resembling an argument or a point, then I have a bridge to sell you.
Negative.
If your platform is predicated on user generated content as twitch is, forcing them to opt out is screwing them over already. How can you not see this? The company is literally leveraging their position as an employer to “legally” steal creators work.
What exactly are you even defending here?
I expect there’s some kind of TOS rule about scraping and recording of streams. Making it opt-out allows Twitch to sell those streams to AI companies similar to how Reddit did after going public.
Currently, the ToS doesn’t really matter, it just let’s them ban you which doesn’t really impact most competent scrapers.









