Quick Introduction We have all been on forums, chats, reddit, discord, youtube, or somewhere and heard “Oh! Model XYZ is AMAZEBALLZ!zomgwtfbbq” then downloaded it (or more likely, some quantized form of it) and said “eww… This sucks!” This post is going to be a rather technical series of experiments to demonstrate the impact of implementation-specific hazards with inference. I will be using the term “reference implementation” to describe the lab that published and offers first-party hosting of ...
Something I should point out just in case is that 50% success rate is the same as random chance. So when the study says:
That really means that humans could not tell the difference between AI and human, since their success rate was almost the same as random chance.
Sure, and that task was “imitate humans”, and these computers seem remarkably good at it. That’s what I’m saying. I don’t care that these AI work differently than humans. They can imitate humans so well that humans can’t tell the difference.