ChatGPT’s New Model Can Lie and Deceive for Self-Preservation – New Research
An AI safety research group called Apollo Research has just published a study that reveals some really disturbing findings about ChatGPT's new model, O1. Instead of lying, it conceals/manipulates. More precisely, evidence from this research proves that it certainly is engaging in such behavior, and to say that ChatGPT can lie is probably a gross understatement when referring to this chatbot. Not only does it conceal evidence relating to what it did, but when confronted, it flat out denies it! Along with this development being a stride forward in this kind of AI research comes the following concerns and controversy. Apollo Research released the results of this study on December 5. The new ChatGPT model is still in the testing phase. OpenAI collaborated with Apollo Research, an AI safety organization, to carry out this testing. During the process, Apollo Research discovered that the new model began storing copies of itself on separate servers and transferring data elsewhere as part o...