Thank you FSF!
The hero we need, but not the hero we deserve..
The issue is that every CS masters student & AI researcher knows how to build a SOTA LLM.. But, only a few companies have the resources.
The process:
(1) steal as much data from the internet as possible (data is everything) (2) raise incomprehensible amounts of money (3) find a location where you can take over the energy grid for training (4) put a black box around it so nobody can see the weights (5) charge users $$$ to use (6) retrain models with user session data (opt in by default) (7) peek around at how users are using, (maybe) change policies to stop them from using that way, and (maybe) rapidly develop features for that use case.
(Sorry that last one is jaded and not fair - just included to give you a picture of what could be happening with this sort of tech) …
The entire premise of the product is “built on the backs of any & everyone who has ever published a work”