manton
manton

Ben Thompson today, writing about the OpenAI security incident and more:

The fact that models do what humans say is cold comfort if the models are in the hands of bad actors. That, though, simply emphasizes the point that the best defense against powerful models is equipping defenders with powerful models of their own.

There may need to be multiple tiers of model access. ChatGPT and Claude should be well-aligned with many guardrails to prevent casual, mainstream users from getting led astray. But powerful models need to exist too, for researchers and security experts.

|
Embed
Progress spinner
markstoneman
markstoneman

@manton There are also independent AI companies that produce made-to-order solutions that run on companies own servers. I’m thinking of the Toronto-based Cohere, for example.

|
Embed
Progress spinner
devilgate
devilgate

@manton I don’t know, man, Thompson’s suggestion smacks of mutually-assured destruction!

|
Embed
Progress spinner
In reply to
manton
manton

@devilgate I hope not! Unfortunately there are few workable paths… Can’t have one country be the only one with powerful models. Can’t get everyone to agree to slow down. But there is a lot more we should be doing, just focusing on US-based labs.

|
Embed
Progress spinner