Back to feed
arXiv cs.AI·

Deployment-Centered Evaluation: Predicting Query-Level Rejection Risk in a Clinical LLM System

Signal
78
Hype
15
In three linesDeployment study of an LLM embedded in electronic health records. A pre-response classifier predicts user rejection risk (AUROC 0.719) by leveraging deployment-specific context (provider type, department, model). Prospective analysis over 4.5 months.
Read source
Your take?
EvalsAI safetyAlignmentBusiness

Summary generated by Claude — human-verified