

I know what an LLM is and how it works. The model for them you currently use to understand them is really bad, I’m sorry to say. It just cannot explain how in context learning is possible, prediction of linebreaks and the model recalling what happened 200 tokens back (where is that information written on the “the” die?), etc. You almost certainly have a deeper understanding of how LLMs work that you have simplified away, if not watch this and then the thousand other more recent videos on how they actually work. You just need to switch from the equivalent model of “gravity makes things fall to the ground” to the equivalent of newtons gravitational laws. Otherwise you will be dumbfounded by completely reasonable things, and forced to reject them in favour of the flawed model you are using.


It’s truly horrible that dangerous bacteria were found in sewage water. Can we not trust any source of previously safe water these days?