Discussion about this post

User's avatar
Richard Self's avatar

"If you want to know how something works, you look at the code."

The problem with open source LLMs is that the code is the easy part to verify. However, the behaviour of the LLM is fundamentally based on the trained weights, which cannot be verified at all, however many people gaze at the values of billions of trillions of weights.

This has been the problem with all the pattern finding analytics algorithms since about 2002 and Big Data. It is almost trivial to check the code but the behaviour depends on the learned values and weights which cannot be validated at all.

Jason David's avatar

They can't expose their training sets without exposing themselves to fair use lawsuits. Speaking of which, how stupid is the US court system that judges think fair use includes taking every single bit of a company's data and using it to create a product that directly competes with that data for customers?

44 more comments...

No posts

Ready for more?