A Palo Alto startup called Abliteration has released an open-weight AI model specifically designed to remove many of the refusal mechanisms built into its underlying model. Its pitch is straightforward: AI that “doesn’t say no,” including for offensive cybersecurity, red-teaming and agent-testing work that mainstream models may refuse. There is...
AI Companionship Is Already Mainstream.
Now We Need to Decide What “Good” Looks Like. TL;DR New research from Elon University gives us a much clearer picture of how Americans are actually using AI for emotional and social interaction. 27% of Americans already use chatbots for personal, emotional, or social purposes, rising to almost 40% among...
Your AI Can Be Hacked With a Sentence. That’s an Architecture Problem.
For most of computing history, hacking required some degree of technical expertise. Attackers searched for vulnerable code, compromised credentials, installed malware, or exploited networks. Generative AI has introduced a strange new attack surface: Sometimes the weapon is simply a cleverly written sentence. Inc. calls prompt injection a "Jedi mind trick"...
