Benchmark Databricks ai_mask() for clinical de-identification: 0.71 PHI F1 vs. 0.96 for John Snow Labs Healthcare NLP, with key accuracy and compliance gaps
`ai_mask()` is a Databricks SQL function, in Public Preview and HIPAA compliant, that masks entity types named in a SQL array literal. Run against an expert-annotated clinical corpus with the...
Clinical de-identification benchmarks in 2026 put John Snow Labs Healthcare NLP at 0.96 PHI F1 on expert-annotated clinical notes, against 0.91 for Claude Opus 4.8, 0.89 for GPT-5.5, 0.86 for...
DICOM de-identification is workflow-specific because PHI can appear in metadata tags, free-text metadata fields, burned-in image pixels, and encapsulated PDF content. A production pipeline may need to inspect tags, apply...
Clinical NLP extracts meaning from unstructured text. But in healthcare, extracted meaning isn't useful until it speaks the same language as the systems that need to act on it. An...
Why radiology AI adoption stalls - and what health systems that scaled it did differently By 2055, US imaging demand will rise 16.9%–26.9% above 2023 levels, while the radiologist workforce...