Quick Introduction We have all been on forums, chats, reddit, discord, youtube, or somewhere and heard “Oh! Model XYZ is AMAZEBALLZ!zomgwtfbbq” then downloaded it (or more likely, some quantized form of it) and said “eww… This sucks!” This post is going to be a rather technical series of experiments to demonstrate the impact of implementation-specific hazards with inference. I will be using the term “reference implementation” to describe the lab that published and offers first-party hosting of ...
Them
You
Those are not the same concepts. The model of LLM remain unchanged. Agents running for hours just modify the prompt they input into an LLM model.
I do not care about the rest of discussion, just swoop in to clarify that distinction.
Fair enough. When I say that AI might be sentient, I include systems like “LLM + agent harness”, treating the whole system as AI.