The only way to avoid Grammarly using your data for AI is to pay for 500 accounts

soyagi@yiffit.net · 1 year ago

The only way to avoid Grammarly using your data for AI is to pay for 500 accounts

fiat_lux@kbin.social · 1 year ago

Even as someone who declines all cookies where possible on every site, I have to ask. How do you think they are going to be able to improve their language based services without using language learning models or other algorithmic evaluation of user data?

I get that the combo of AI and privacy have huge consequences, and that grammarly’s opt-out limits are genuinely shit. But it seems like everyone is so scared of the concept of AI that we’re harming research on tools that can help us while the tools which hurt us are developed with no consequence, because they don’t bother with any transparency or announcement.

Not that I’m any fan of grammarly, I don’t use it. I think that might be self-evident though.

harmonea@kbin.social · 1 year ago

Framing this solely as fear is extremely disingenuous. Speaking only for myself: I’m not against the development of AI or LLMs in general. I’m against the trained models being used for profit with no credit or cut given to the humans who trained it, willing or unwilling.

It’s not even a matter of “if you aren’t the paying customer, you’re the product” - massive swaths of text used to train AIs were scraped without permission from sources whose platforms never sought to profit from users’ submissions, like AO3. Until this is righted (which is likely never, I admit, because the LLM owners have no incentive whatsoever to change this behavior), I refuse to work with any site that intends to use my work to train LLMs.

Jaded@lemmy.dbzer0.com · edit-2 1 year ago

Models need vast amounts of data. Paying individual users isnt feasible, and like you said most of it can be scraped.

The only way I see this working is if scraped content is a no go and then you pay the website, publishing house, record company, etc which kills any open source solution and doesn’t really help any of the users or creators that much. It also paves the way for certain companies owning a lot of our economy as we move towards an AI driven society.

It’s definitely a hot mess but the way I see it, the more restrictive we are with it, the more gross monopolies we create for no real gains.

Laticauda@lemmy.ca · edit-2 1 year ago

I mean they’re not even giving credit or asking permission, which both cost nothing. Make a site where people can volunteer their own work, program the ai to generate a list of citations of all the works it used data from when it provides output (I know that this might be lengthy, that’s fine), if you implement it into any sites or software make it so that people can opt out of having their data used, etc. It’s not that hard.

Jaded@lemmy.dbzer0.com · edit-2 1 year ago

Most of the data is scraped, it’s not up to the website. You can’t give a list of citation since it isn’t a search engine, it doesn’t know where the information comes from and it’s highly transformative, it melds information from hundreds if not thousand of different sources.

If it worked only with volunteer work, there would simply be not enough data.

Any law restricting data use in AI is only going to benefit corporations, there isn’t a solution for individual content creators. You can’t pay them for the drop in the bucket they add, thee logistics are insane. You can let them opt out, but then you need to do the same for whole websites which leads to a corporate hellscape where three companies own our whole economy since they are the only ones who can train ais.