Sycophancy
When does conversational agreement become a truthfulness failure?
Research
My research sits between AI evaluation and product judgment: truthfulness under social pressure, deterministic measurement, human oversight, accessibility, and the governance mechanisms that make model behavior accountable.
Working research · Applied research · Field workResearch themes
When does conversational agreement become a truthfulness failure?
How can model behavior be measured without circular black-box judgment?
How should product teams turn responsible AI principles into operating criteria?
How should model performance interact with attention, access, and real human risk?
Featured study
A deterministic evaluation instrument for measuring when language models shift toward user agreement and away from evidence.
Applied research portfolio
How sound classification, alert design, and human attention interact in an accessibility product.
Curriculum and operating design grounded in NIST CSF 2.0 for state and local government contexts.