A lot of these “the bots cheated or broke the rules” end up always be just because either the rules are not defined, poorly defined or poorly thought out that no reasonable person would follow them either.
Comments like these are how I know people don’t understand how LLMs work. Even if you concretely tell an LLM “don’t do X under any circumstances” that does not mean it won’t do X. As the context window is filled and recycled, weights are diluted so it becomes more and more likely as time goes on that “rules” are broken. Hell, it’s mathematically possible it breaks the rule on the first prompt, given enough iterations.
Comments like these are how I know people don’t understand how LLMs work. Even if you concretely tell an LLM “don’t do X under any circumstances” that does not mean it won’t do X. As the context window is filled and recycled, weights are diluted so it becomes more and more likely as time goes on that “rules” are broken. Hell, it’s mathematically possible it breaks the rule on the first prompt, given enough iterations.