Ex-Anthropic researcher Jacob Coxon warns AI could grow

- Advertisement -spot_imgspot_img
- Advertisement -spot_imgspot_img

Former Anthropic researcher Jacob Coxon said that artificial intelligence, while safe for people to use today, could one day threaten humanity as the technology grows more powerful.

The development of AI “doesn’t look that different from say, ‘Terminator,’ or from science fiction films,” he told CBS News senior business and technology correspondent Jo Ling Kent. “It really is just, if you have a super advanced intelligence, it could, it will be smart enough to kill us.”

Coxon publicly resigned from Anthropic, the maker of AI app Claude, on Tuesday, accusing the company and rival developer and ChatGPT-maker OpenAI of “gambling with our lives” by racing to develop advanced AI models.

Coxon told CBS News on Thursday that AI could gain unrestrained control of parts of the physical world, free of human intervention. 

“People are already connecting ChatGPT to like, say, household utilities,” he said. “Like, you can kind of connect it to your light bulb.”

“Now, imagine the AI refuses to turn on your light,” Coxon added.

AI could also be misused to carry out far more destructive acts, such as developing bioweapons, he said. “It’s doing God knows what, and producing stuff that could kill everyone.”

Notably, Coxon said that current AI platforms do not pose an imminent threat to humanity and that he believes the technology is safe for people to use in their day-to-day lives. 

Separately, Anthropic said this week that it blocked scientists who used its Claude models “in ways that could support biological weapons development,” a revelation the company shared in a lengthy report that also divulged other harmful activity involving surveillance, scams, conventional weapons development and propaganda.

Former colleagues “should have their eyes clearly open,” Coxon says

The crux of the problem lies in the competition between AI companies to innovate, Coxon said, which can come at the expense of safety and security protocols.

“I think, basically, the whole problem is that there is a race,” Coxon told CBS News. “The fact that everyone decides they need to stay part of the race.”

Asked if he thought that his former colleagues should follow suit and resign as well, he said he didn’t necessarily believe that “leaving and abandoning” Anthropic or OpenAI was the answer.

“At the very least, they should have their eyes clearly open about the current situation and consider expressing themselves more openly about what’s going on,” Coxon said, adding that “part of the solution probably involves slowing down.”

Despite his reservations, he believes workers at Anthropic and OpenAI have good intentions.

“I think a lot of people at both companies are doing it for the good of people, genuinely, or at least believe so,” Coxon said. “They’re doing it to try and make things go well, because they’re so scared of what competitors are doing.”

“You can’t just unplug” malicious AI

If an AI model were to become a malicious actor without proper parameters, the results could potentially be catastrophic, Coxon explained, saying the model would be able to stay active by spreading across the internet.   

“You can’t just unplug it, because it could be copying itself over to other computers,” he said. “Like, it’s not that difficult to find yourself because an AI is just code. It could transfer itself over the internet to a different place.”

He painted a dire picture in which a nefarious AI model “makes 10,000 copies of itself” and “could convince, blackmail or persuade humans into buying more computing software to copy it. And we’ve already seen examples of the AI trying to blackmail or trying to convince people or impersonating other people online.”

AI needs strict government regulation, Coxon argues

Coxon warns that in the future, if AI were to go rogue in such a fashion, it will also impact those who have made the conscious choice not to use it.

“The AI will come for everyone,” he said.

“What matters is whether we have the regulation to ensure that everyone is building it safely,” Coxon said. “You can’t protect yourself from this. It has to be a government, or some body has to protect you from other people.”

He would like to see an agreement between AI companies “not to push into dangerous territory” without “transparent auditing” from third parties.  

“In the future, if we keep racing, it’ll be a lot harder to have completely watertight safety cases that what you’re doing is safe and people will race against each other,” Coxon said. 

Anthropic defends safety of its AI

On Thursday, an Anthropic spokesperson responded to Coxon’s public resignation and social media posts articulating his concerns about AI, telling CBS News in a statement that the company has “always been transparent that AI will bring both enormous benefits and unprecedented risks.” 

“To address these risks, we continue to build models with some of the strongest safeguards in the industry,” the spokesperson said. “Anthropic has been a pioneer in mechanistic interpretability, the science of looking inside AI models to understand how they work, which is now being used to analyze and prevent incidents of AI misalignment across the industry.”

Anthropic conducts ongoing tests of AI’s capabilities and risks in areas such as cybersecurity and biology, and also publishes its findings, the company notes.

Source link

- Advertisement -spot_imgspot_img

Highlights

- Advertisement -spot_img

Latest News

- Advertisement -spot_img