this post was submitted on 24 Jul 2026
30 points (87.5% liked)
Linux
14720 readers
313 users here now
A community for everything relating to the GNU/Linux operating system (except the memes!)
Also, check out:
Original icon base courtesy of lewing@isc.tamu.edu and The GIMP
founded 3 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
As a member of the community, I won't have a problem if the interactions with LLMs are only with open source ones, and the conversation is available as open data to all, without restrictions. The same which CommonVoice does for Voice.
Data matters so much that the companies keep it to themselves. And I absolutely don't want the data to be used for training models which enslave the developers and the people of this planet. The least is that data should be open to all, so the FOSS community would have a chance to use the data to contribute back to FOSS.
Note that it doesn't matter if the model is open source. For Machine Learning the data matters as much as the code does, if not more. And in this case, data matters much much more. The code is widely available and can also get researched independently.