Ima be honest, if i was given the task to make a bot to beat another bot tried my best and couldn’t. I would just go fork the best bot thats out there and use that as the base and not reinvent the wheel. Learn from its design then improve on it as i study it.
If the goal is ONLY to win a game of starcraft. Fuck it just download the bot. Frankly the fact the bots decided that wasting resources was pointless and to just do the simplier thing is sorta what we want them to do. Thats less eletrical useage, less water wasted, and gets the job done.
A lot of these “the bots cheated or broke the rules” end up always be just because either the rules are not defined, poorly defined or poorly thought out that no reasonable person would follow them either.
And thats the thing these models are litterally the text book definition of wisdom of the masses/mass consensus. They do exactly what the avg run of the mill person with the knowledge at hand would do. That ends up being rather all over the place since people are all over the place. But one of the most common things is that people generally. Hate fucking pointless work and will try to streamline, shortcut or simplify any work given to them to get an acceptable outcome.
Its rare you get someone that will stubbornly work on a problem forever that they can’t over come with out changing their apporch rules be damned.
A lot of these “the bots cheated or broke the rules” end up always be just because either the rules are not defined, poorly defined or poorly thought out that no reasonable person would follow them either.
Comments like these are how I know people don’t understand how LLMs work. Even if you concretely tell an LLM “don’t do X under any circumstances” that does not mean it won’t do X. As the context window is filled and recycled, weights are diluted so it becomes more and more likely as time goes on that “rules” are broken. Hell, it’s mathematically possible it breaks the rule on the first prompt, given enough iterations.
Ima be honest, if i was given the task to make a bot to beat another bot tried my best and couldn’t. I would just go fork the best bot thats out there and use that as the base and not reinvent the wheel. Learn from its design then improve on it as i study it.
If the goal is ONLY to win a game of starcraft. Fuck it just download the bot. Frankly the fact the bots decided that wasting resources was pointless and to just do the simplier thing is sorta what we want them to do. Thats less eletrical useage, less water wasted, and gets the job done.
A lot of these “the bots cheated or broke the rules” end up always be just because either the rules are not defined, poorly defined or poorly thought out that no reasonable person would follow them either.
And thats the thing these models are litterally the text book definition of wisdom of the masses/mass consensus. They do exactly what the avg run of the mill person with the knowledge at hand would do. That ends up being rather all over the place since people are all over the place. But one of the most common things is that people generally. Hate fucking pointless work and will try to streamline, shortcut or simplify any work given to them to get an acceptable outcome.
Its rare you get someone that will stubbornly work on a problem forever that they can’t over come with out changing their apporch rules be damned.
Comments like these are how I know people don’t understand how LLMs work. Even if you concretely tell an LLM “don’t do X under any circumstances” that does not mean it won’t do X. As the context window is filled and recycled, weights are diluted so it becomes more and more likely as time goes on that “rules” are broken. Hell, it’s mathematically possible it breaks the rule on the first prompt, given enough iterations.
I’m honestly stunned you’ve missed the entire point this badly…
How many hours a day do you use chatbots?