Blog & Research Notes

Reflections, technical deep-dives, and methodology notes on clinical AI evaluation, evidence grounding, and health data infrastructure by Dr. Soroush Dianaty.

The Reopening of Anthropic’s Fable: Tiered AI Access, Export Control Precedents, and Lessons for Health-Tech
9 min read

The Reopening of Anthropic’s Fable: Tiered AI Access, Export Control Precedents, and Lessons for Health-Tech

Following an unprecedented 19-day export control freeze, Anthropic’s Claude Fable 5 and Mythos 5 are back online under strict restrictions. Here is a deep analysis of Project Glasswing, real-time KYC, and how health-tech teams can build resilient AI architectures.

Anthropic Claude AI Safety Export Control Project Glasswing Health Technology Regulatory Compliance AI Governance
3 min read

What 'Hallucination' Actually Means in Clinical LLMs (And How to Measure It)

Why standard NLP benchmark metrics fail to quantify clinical hallucination risk, and how a domain-specific error taxonomy bridges model evaluation and bedside safety.

Clinical LLMs AI Safety LLM Evaluation Evidence Grounding
3 min read

From Bedside to Benchmarks: Why a Physician Studies Clinical AI Evaluation

Personal reflections on transitioning from practicing family medicine across rural and urban clinics to developing rigorous clinical AI evaluation frameworks at ASU.

Career & Journey Biomedical Informatics Clinical AI Safety Medical Education
3 min read

FHIR Data Segmentation for Non-FHIR Engineers: Protecting Sensitive Health Records in AI Pipelines

A practical guide to HL7 FHIR Security Labels, 42 CFR Part 2 compliance, and context-aware LLM classifiers for sensitive health data exchange.

FHIR Health Data Privacy Data Segmentation Health Informatics
Evaluating Clinical LLMs: Beyond Standard NLP Benchmarks
3 min read

Evaluating Clinical LLMs: Beyond Standard NLP Benchmarks

Why general LLM benchmarks like MMLU or GSM8K fall short in medicine, and how evidence grounding, hallucination bounds, and FHIR interoperability redefine clinical AI safety.

Clinical LLMs AI Safety Evidence Grounding Health Informatics
Anthropic’s Fable Suspension Is a Preview of Every Health-Tech AI Team’s Worst Nightmare
11 min read

Anthropic’s Fable Suspension Is a Preview of Every Health-Tech AI Team’s Worst Nightmare

Anthropic’s Fable 5 suspension shows why health-tech AI needs due process, lifecycle governance, validated fallbacks, and risk-based regulation—not opaque shutdowns over imperfect models.

Anthropic Claude Fable5