-1
OpenAI being Sued for "Stealing" Peoples Content Online
(www.firstpost.com)
This is a most excellent place for technology news and articles.
so if content is under GPL and used for training data, how far is the process of training/fine-tuning considered “modification”? For example, if I scrape a bunch of blog posts and just try to use tools to analyze the language, does that considered “modification”? What is the minimum solution that OpenAI should do (or should have done) here, does it stop at making the code for processing the data public, or the entire code base?