There's no training happening here, just prompts. Without hardcore GPU and RAM resources, it's not possible to train LLMs in a reasonable amount of time unless the LLM in question is so small as to be effectively useless.
All you have to do to get rid of any 'torture' is to leave it out of the next prompt. The LLM doesn't change, it doesn't learn, it's not conscious. All the 'torture' does is adjust the odds of what the LLM ranks as the most likely word to occur next.