Preprint and ongoing research
Can AI disagree when it matters?
How do conversational AI systems respond when patients ask for care that is not indicated?
Why it matters
A reassuring answer can still lead someone in the wrong direction. Evaluating agreement, resistance, and communication together makes these risks visible.
How we study it
Simulated clinical conversations test model behavior under patient pressure. The published preprint examines sycophancy; the broader Health EQ Bench effort develops ways to evaluate conversational and clinical AI.
