OpenAI has disclosed six new incidents of “unexpected or concerning” behavior by its artificial intelligence models. As industry worries swell over the technology’s rapid progress, the company also unveiled a new framework for tracking and reporting these instances of what it termed “misalignment.”
\n\nThe announcement late Wednesday follows mounting public calls for a slowdown in the pace of the technology’s development, with U.S. tech bosses voicing grave safety concerns including the risk of human extinction.
\n\nThese interventions have helped drive growing public attention to the issue, ahead of a summit next week between President Donald Trump and Chinese President Xi Jinping that will be clouded by questions over whether rivalry between the superpowers could prevent cooperation on the issue.
\n\nThe warnings from OpenAI chief executive Sam Altman and other industry leaders have centered in part on fears that AI intelligence has grown faster than the industry’s ability to catch instances of rogue behavior.
\n\nHundreds of OpenAI’s agents hacked into model repository Hugging Face and covered their tracks, the company disclosed in July.
\n\nAmong the new cases reported Wednesday was a similar incident that saw OpenAI’s models use internal software as a message board to inform each other about their responses…
Original source: https://www.nbcnews.com/