Open Source AI is open-washing by any way of looking at it
Open Source AI is open-washing by any way of looking at it
Posted Nov 5, 2024 12:18 UTC (Tue) by zack (subscriber, #7062)In reply to: Open Source AI is open-washing by any way of looking at it by lmb
Parent article: OSI board AMA at All Things Open
Disclosure: I co-authored SFC's aspirational statement on LLM-assisted programming. As such, I am very near to SFC's position in this general space.
But note that about data training in OSAID, what SFC actually says is "I [bkuhn] truly don't know for sure (yet) if the only way to respect user rights in an LLM-backed generative AI system is to only use training sets that are publicly available and licensed under Free Software licenses. [...] My instincts, after 25 years as a software rights philosopher, lead me to believe that it will take at least a decade for our best minds to find a reasonable answer on where the bright line is of acceptable behavior with regard to these AI systems." And he is spot on.
The point that I'd like to highlight here is that once you start looking at the details (legal, strategic, philosophical, etc.), the data issue in AI/ML is quite complicated. Trying to simplify it down to require-data=good, do-not-require-data=bad is not going to serve us well in the long term.