Lemmy is CEO-proof. After Digg, Reddit and Twitter, that term should be a thing

Wander@yiffit.net · 2 years ago

Lemmy is CEO-proof. After Digg, Reddit and Twitter, that term should be a thing

incognito_mode@lemmy.world · 2 years ago

This is a great point. The user data needs to be enshrined in such a way that it can be easily moved in a bulk migration without requiring a direct opt-in from every user. While at the same time making it clear how it’s being used/kept/sold/not sold/etc.

I’m not against LLMs using the data generated on sites like this to inform useful answers when I ask ChatGPT a question. It genuinely makes AI a better tool, but I feel like the contributors of such content should know how their answers are being used.

lightrush · edit-2 2 years ago

LLMs are likely going to scrape no matter the license. I doubt OpenAI got a copyright license from Reddit to ingest it. In fact I’m not even sure they need one if ingestion can be make similar enough to “reading the web site”. And so making content CC probably won’t affect LLM use of public posts.