Skip to content

Publications

New Publication in npj Digital Medicine

Teferra et al. (2026) systematically evaluated how safety guardrails influence the affective outputs of large language models (LLMs), operationalizing irritability as a measurable behavioral construct. Using validated psychometric instruments adapted for model assessment, the team compared multiple contemporary LLMs under varying safety configurations to quantify differences in affective reactivity. The findings demonstrate that models with … Read More

Artificial Intelligence for Mental Health