How do we prevent AI agents from going rogue? It starts with a new kind of measurement | Bruce Schneier and Barath Raghavan ⊕
found
a story from The Guardian ⚠️ › International
You are not logged in so some information on this page has been withheld. To see more, please log in or sign up.
since
auto-detected in 11 stories
3 days ago
found
a story from The Guardian ⚠️ › International
Like genies of folklore, AI agents take their instructions literally – to potentially disastrous effect. We must track their ability to do what we actually mean
In July, Hugging Face, a company that h…
7 days ago
Sometimes, what matters isn’t your output but what you put into the process. Think of it like work v the gym
I teach public policy at the Harvard Kennedy School and the Munk School at the University o…
25 days ago
These systems will soon be able to track our public and private lives. But we can make the policy choices to reject it
In the near future, AI-powered surveillance systems will be able to track everyth…
32 days ago
Modern AI systems are, in effect, a universal adviser to help people do harmful things. We’ll need to harness AI for defense, too
Earlier this week, national security agencies from the Five Eyes – tha…
37 days ago
found
a story from The Guardian ⚠️ › International
A court in Germany found that Google was responsible for what its chatbots say in search summaries. This is the accountability we need
Earlier this month, a German court ruled that Google is liable fo…
45 days ago
We have opened the AI Pandora’s box. Now we have to make the best of it
On 9 June, Anthropic released its Fable generative AI model. Three days later, the US government classified it as a dangerous m…
52 days ago
found
a story from The Guardian ⚠️ › International
Also: Anthropic advocates for a ‘pause’ on AI advancement – days after filing to go public on the US stock market
Hello, and welcome to TechScape. I’m your host, Blake Montgomery, the US tech editor a…
84 days ago
found
a story from The Guardian ⚠️ › International
The system’s power is comparable to others – but it still has frightening implications for the future of hacking
Last month, Anthropic made a remarkable announcement about its new model, Claude Mythos…
3 months ago
Large language models aren’t trained on real-life conversations. As we encounter their language, it could affect our own
Because of the way they are trained, large language models capture only a slice…
4 months ago
The lesson here isn’t that one AI company is more ethical than another. It’s that we must renovate our democratic structures
OpenAI is in and Anthropic is out as a supplier of AI technology for the US…