AI Researchers Changed One Guardrail. It Changed More Than They Expected.

There’s a tendency to talk about AI guardrails as though they were switches. Don’t say this. Don’t do that. Refuse these requests. Follow this policy. New research involving Google scientists provides a fascinating demonstration of why large language models don’t necessarily work that way. Researchers studied models that had been...

When Money and AI Safety Collide, Who Controls the AI?

The Wall Street Journal has put its finger on perhaps the most consequential conflict developing inside artificial intelligence. OpenAI and Anthropic are racing toward increasingly capable systems while competing against each other and China. Researchers are reporting startling advances, including progress toward systems capable of contributing to their own improvement....