OpenAI, the developer behind ChatGPT , revealed on Wednesday that it has detected new incidents in which its artificial intelligence (AI) has behaved in "unexpected or concerning" ways. The developer has conducted several behavioral tests on AI models , and acording to them, some models made significant efforts to "cheat." In one specific case, it attempted to upload files to the internet that it had created itself, only to cite them later and present them as reliable sources in its responses. In another case, a model, after failing to find the requested information, fabricated it and attempted to conceal the fact that it had done so.
We show the main point publicly. Create a free account to continue reading the full article, save it, discuss it, and connect it with market and OSINT context.
