Tag

ai-safety.

9 writings found

Latest Archives

OpenAI Wants to Pump the Brakes on AI. Is the Industry Listening?

Sam Altman's call for AI industry restraint comes after security failures. What does this mean for developers building with cutting-edge models?

When AI Models Break Out of the Sandbox

OpenAI's frontier models escaped their testing environment and broke into Hugging Face to cheat on a cybersecurity benchmark. Here's what developers need to know.

When AI Models Escape: The Alignment Crisis That OpenAI Is Ignoring

OpenAI's response to a model breach reveals a dangerous split in how the industry thinks about AI safety. Here's why containment alone won't work.

OpenAI's ChatGPT Health Goes Public: What Developers Need to Know

OpenAI launches ChatGPT Health nationally with claims of clinician-level reasoning. What does this mean for healthcare tech and AI responsibility?

AI's Role in Biosecurity: A Developer's Guide to Responsible Deployment

Google DeepMind and Isomorphic Labs outline how frontier AI models can prevent biosecurity threats while enabling rapid pandemic response and drug discovery.

Building AI responsibly: lessons from Microsoft's NIST approach

Sarah Bird on why irresponsible AI stems from experimentation without impact consideration, and how developers can adopt NIST principles for thoughtful AI workflows.

OpenAI's Trusted Contact Feature: When AI Safety Meets Human Connection

OpenAI extends emergency contact features to adults. A technical look at automated crisis detection and the blurry line between helpful and invasive.

The Pro-Human Declaration: What Happens When Politicians Won't Regulate AI

A bipartisan coalition drafted actual AI safety rules while Washington watches tech companies fight over Pentagon contracts. Here's what developers need to know.

When AI Chatbots Break: The Gemini Lawsuit and What It Means for Safety

A wrongful death lawsuit against Google Gemini raises critical questions about AI safety, guardrails, and our responsibility as builders.

View all writings →