Quick Introduction We have all been on forums, chats, reddit, discord, youtube, or somewhere and heard “Oh! Model XYZ is AMAZEBALLZ!zomgwtfbbq” then downloaded it (or more likely, some quantized form of it) and said “eww… This sucks!” This post is going to be a rather technical series of experiments to demonstrate the impact of implementation-specific hazards with inference. I will be using the term “reference implementation” to describe the lab that published and offers first-party hosting of ...
I’ve tried. They’re adamant against AI here. It’s not even worth the discussion. There’s an army of belligerent people here who have made up their minds that LLM are “stochastic parrots” and they’re convinced they understand exactly how they work. No one can tell them otherwise. Doesn’t even matter your credentials. They’re the experts here. If you argue, they’ll only belittle you.
Save yourself the frustration and heartache; just silently pity them.