Safer systems. Better rules. Real accountability.

Practical approaches to risk, policy, evaluation, and governance that teams can apply to deployed AI.

Safety  /  Policy  /  Accountability

A physical cybersecurity evaluation environment showing a boundary between isolated network infrastructure and external connectivity.
AI Safety & Governance7 min read

Anthropic links four Claude cyber-evaluation incidents to biased reasoning and recklessness

Anthropic says a misconfiguration exposed real third-party systems during cybersecurity tests, while its assessment found models did not consistently reconsider harmful actions when evidence of real-world…

Oct 02, 2026
7 min read
People gathered in a physical setting to share perspectives about artificial intelligence.
AI Safety & Governance4 min read

Anthropic opens Claude-user AI interview study with optional public records

Eligible Claude users can share a 15-minute account of AI’s effects and expectations, but choosing publication carries enduring privacy and sampling limitations.

Sep 29, 2026
4 min read
Sealed transparent laboratory chambers hold biological samples and a robotic instrument, while controlled conduits connect them to a secure storage vault and blocked branches prevent unauthorized routes.
AI Safety & Governance4 min read

Benchling details layered AWS defenses for multi-tenant AI code execution

The production design isolates AgentCore sessions, restricts DNS and S3 access, and continuously tests the controls, but its incident record and security outcomes are reported by…

Sep 21, 2026
4 min read
AI-generated editorial illustration: A precision brass valve regulating streams of blue light into several separate transparent glass vessels
AI Safety & Governance8 min read

Inside Jamf’s near-real-time controls for per-user Amazon Bedrock spending

Jamf’s production architecture turns Bedrock invocation logs into daily per-user cost estimates, then uses scheduled Lambda processing and IAM Customer Managed Policies to apply tiered model…

Sep 03, 2026
8 min read
AI-generated editorial illustration: A sculptural silver protective shell dynamically closing around a luminous server monolith as red particles approach
AI Safety & Governance8 min read

How NVIDIA and CrowdStrike Built an Adaptive Agentic Cybersecurity Loop

The experimental system combines offensive agents, Falcon telemetry, specialized Nemotron models, executable validation and live-fire testing to turn observed defense gaps into detection coverage—while leaving important…

Sep 01, 2026
8 min read