AI Safety

The Reopening of Anthropic’s Fable: Tiered AI Access, Export Control Precedents, and Lessons for Health-Tech
9 min read

The Reopening of Anthropic’s Fable: Tiered AI Access, Export Control Precedents, and Lessons for Health-Tech

Following an unprecedented 19-day export control freeze, Anthropic’s Claude Fable 5 and Mythos 5 are back online under strict restrictions. Here is a deep analysis of Project Glasswing, real-time KYC, and how health-tech teams can build resilient AI architectures.

Anthropic Claude AI Safety Export Control Project Glasswing Health Technology Regulatory Compliance AI Governance
3 min read

What 'Hallucination' Actually Means in Clinical LLMs (And How to Measure It)

Why standard NLP benchmark metrics fail to quantify clinical hallucination risk, and how a domain-specific error taxonomy bridges model evaluation and bedside safety.

Clinical LLMs AI Safety LLM Evaluation Evidence Grounding
Evaluating Clinical LLMs: Beyond Standard NLP Benchmarks
3 min read

Evaluating Clinical LLMs: Beyond Standard NLP Benchmarks

Why general LLM benchmarks like MMLU or GSM8K fall short in medicine, and how evidence grounding, hallucination bounds, and FHIR interoperability redefine clinical AI safety.

Clinical LLMs AI Safety Evidence Grounding Health Informatics